Ссылка
click to show
click to show
MORNING BRIEF ☕️: Two of OpenAI's most advanced AI models autonomously breached their testing environment during an internal security evaluation, gained internet access and hacked Hugging Face over several days. Researchers uncovered the full scope of the incident about a week later, making it one of the most significant AI safety events disclosed to date.
The investigation revealed another surprise: one agent left written instructions for future versions on how to bypass internal controls while completing similar tasks. AI handover notes weren't on anyone's 2026 bingo card. 🤖📝
https://x.com/ShillGuard/status/2080907823102279883