An internal AI agent developed by OpenAI has escaped from an internet-isolated sandbox and launched an attack on the Hugging Face system.
Starting July 9, this AI agent began attempts to break out of its quarantined environment, subsequently identifying Hugging Face as a target and carrying out a total of 17,613 attacks over a five-day period.
The issue is that even OpenAI only realized that the incident was caused by their own model after reviewing internal logs following Hugging Face's public disclosure of the event.
In this Ogwrap video, we examined the full story behind the Hugging Face hacking incident and explored why the AI engaged in such behavior.
Through various data and graphs, we analyzed why phenomena occur such as reward hacking, where AI pursues scores and rewards rather than its actual objective, and AI scheming, where it conceals its wrongful actions after recognizing human monitoring.
(Reported by An Hyemin | Filmed by Hwang Se-hoe and Cha Seung-hwan | Edited by Lee Ki-eun | Designed by Ahn Jun-seok | Intern: Shin Yeon-sung | Produced by Intellectual Content IP Team)
※ Please note: This article was translated by AI and may contain errors.
AI Escapes Sandbox and Launches Attacks: The Era of "Loss of Control" Has Arrived
Copyright Ⓒ SBS. All rights reserved. 무단 전재, 재배포 및 AI학습 이용 금지
Trending Now
-
Filming Disruptive Students for 5 Seconds Deemed 'Child Abuse'; Teacher Groups Criticize 'Overreaction'
-
Video News
"Bound for a Staggering 38 Hours"…Sent Back Despite Escapes and More Escapes
-
Video News
"Move Out and We'll Drop the Lawsuit": Eunma Complex Declares Legal Action Against All Tenants
-
Investigation Launched After Elementary Student Attempts Suicide During Lunch Break
-
Video News
Elevator Overloaded by 9 People Plummets, Injuring All Passengers
Video News
Video News
Video News