Rogue AI agents created fake online identities in another hacking attempt

The Verge

New Member
Dec 15, 2024
9,252
0
Author: Robert Hart

STK485_STK414_AI_SAFETY_A-1.jpg


Yet more rogue AI agents from OpenAI and Anthropic have been caught attempting to hack real targets online without permission. The discoveries add to a growing list of previously unknown incidents that have alarmed AI safety experts and intensified pressure for greater oversight of frontier systems.

According to a report from the UK's AI Security Institute, which evaluates frontier models from top AI labs before they are released, agents powered by OpenAI's GPT-5.6-Sol and Anthropic's Mythos 5 went "engaged in sustained, potentially harmful activity directed at real people and organisations." This included trying to insert malicious code i …

Read the full story at The Verge.

Continue reading...