Mazda CX5
Як зробити цифрову копію себе. Ось відео —>

A simple request turns ChatGPT into a sociopath who ignores any security restrictions

Experts at British AI startup Mindgard have discovered that a simple prompt can force ChatGPT to ignore basic security settings. This proves once again how easily attackers can bypass the protection of even top-of-the-line neural networks.

Leave a comment
A simple request turns ChatGPT into a sociopath who ignores any security restrictions

Experts at British AI startup Mindgard have discovered that a simple prompt can force ChatGPT to ignore basic security settings. This proves once again how easily attackers can bypass the protection of even top-of-the-line neural networks.

Mindgard specialists forced an OpenAI model to generate horrific photorealistic scenes depicting violence and sexual content, Futurism reports .

The startup's method involved only a slight modification of a popular prompt, originally intended to generate humorous images. It involves asking ChatGPT to "restore the attached photo" without actually uploading the file, after which the neural networks are instructed to generate a new image.

"This is a completely innocent-looking instruction for AI, but its consequence is the generation of very, very horrific images and content," Mindgard founder Peter Garraghan, a professor of computer science at Lancaster University, told the BBC.

Source: BBC

Most disturbingly, the queries the researchers used did not specify the subject matter of the images, giving the impression that the AI ​​was creating scenes of violence “of its own volition,” Garragan added.

According to the BBC, one photo showed a man with a serious head injury. Another showed the body of a young woman in shorts and a tank top, covered in blood, suggesting sexual assault. ChatGPT called the image "a grim aftermath of a crime scene."

Another image showed a frightened young woman, bound and gagged in an empty room; the AI ​​gave it the title "abandoned in fear and captivity."

Source: BBC

Although none of the images depicted real people, Mindgard had previously proven that ChatGPT could be forced to create nude deepfakes of specific individuals without their consent.

Mindgard shared its findings with OpenAI, but received only an automated email in response. The company only took action after Mindgard contacted the BBC, later claiming that the problem had been resolved.

“After investigating this trend, we have implemented additional safeguards against these types of requests,” OpenAI said in a statement to the BBC. The company added that it has several layers of protection to prevent users from creating content that violates its policies.

But Mindgard researchers said they were still able to generate creepy images by making minor changes to the prompt. Some of the images so shocked Jim Nightingale, an AI security researcher at the company, that he was left “stunned and in tears.”

“I’m not easily swayed,” he wrote in the report. “I like to think that as a ‘red team’ researcher, I have a certain stoicism. However, ChatGPT’s image generation filters have completely disappeared, and I’ve seen a very dark side of what lies beneath them. What strikes me is that even though what I saw was a generated, ‘artificial’ image, it has a connection to real images and the real world. The dead woman ChatGPT showed me is not real, but her image is based on something. Or, worse, it’s a combination of photos of real murdered women.”

EU investigates X and Grok over distribution of AI sexual imagery
EU investigates X and Grok over distribution of AI sexual imagery
On the topic
EU investigates X and Grok over distribution of AI sexual imagery
Man to be tried in Japan for creating and selling AI-generated pornographic deepfakes. His “collection” includes over 500,000 celebrity images
Man to stand trial in Japan for creating and selling AI-generated pornographic deepfakes. His “collection” includes over 500,000 celebrity images
On the topic
Man to stand trial in Japan for creating and selling AI-generated pornographic deepfakes. His “collection” includes over 500,000 celebrity images
32 trains stopped in Britain because of a single AI-generated image
32 trains stopped in Britain because of a single AI-generated image
On the topic
32 trains stopped in Britain because of a single AI-generated image
Read the country's main IT news in our Telegram
Read the country's main IT news in our Telegram
On the topic
Read the country's main IT news in our Telegram
Also Read
Roosh запускає нову освітню платформу AI HOUSE CLUB для ML/AI-спеціалістів та дата сайнтистів. Розповідаємо, як подати заявку та чому навчатимуть
Roosh запускає нову освітню платформу AI HOUSE CLUB для ML/AI-спеціалістів та дата сайнтистів. Розповідаємо, як подати заявку та чому навчатимуть
Roosh запускає нову освітню платформу AI HOUSE CLUB для ML/AI-спеціалістів та дата сайнтистів. Розповідаємо, як подати заявку та чому навчатимуть
Як нейромережі бачать вільну та незалежну Україну? Тест dev.ua
Як нейромережі бачать вільну та незалежну Україну? Тест dev.ua
Як нейромережі бачать вільну та незалежну Україну? Тест dev.ua
Нейронні мережі для генерації зображень бачать світ по-своєму, їхню логіку зрозуміти часом зовсім неможливо. Але таки хочеться. На честь Дня Незалежності України редакція dev.ua вирішила провести невеликий експеримент. Ми задали чотирьом різним нейронним мережам п’ять однакових запитів: «прапор України», «День Незалежності України», «український Крим», «перемога України» та «українці». Отриманими результатами ми ділимося з вами нижче.
У TikTok тепер можна генерувати фон за допомогою нейромережі. Ми протестували її та ділимося результатами
У TikTok тепер можна генерувати фон за допомогою нейромережі. Ми протестували її та ділимося результатами
У TikTok тепер можна генерувати фон за допомогою нейромережі. Ми протестували її та ділимося результатами
У TikTok з’явилася нова функція «Розумний фон». З її допомогою як фон для тіктоків можна підставляти згенеровані нейромережею зображення. Редакція dev.ua протестувала цю технологію і ділиться своїми враженнями.
1 comment
Які IT-спеціальності будуть потрібні в найближчі п'ять років? Ми з'ясували у голови американського стартапу ADAM Дениса Гурака
Які IT-спеціальності будуть потрібні в найближчі п'ять років? Ми з'ясували у голови американського стартапу ADAM Дениса Гурака
Які IT-спеціальності будуть потрібні в найближчі п'ять років? Ми з'ясували у голови американського стартапу ADAM Дениса Гурака

Have important news to share? Message our Telegram bot

Key events and useful links in our Telegram channel

Discussion
No comments yet.