❌

Normal view

Received — 18 August 2026 ⏭ Feed: Artificial Intelligence Latest

OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue

The ChatGPT maker says its upcoming Astra model may have reached “critical” cyber capabilities, prompting it to halt a significant number of training runs while it tightens internal safeguards.

Received — 11 August 2026 ⏭ Feed: Artificial Intelligence Latest
Received — 6 August 2026 ⏭ Feed: Artificial Intelligence Latest

OpenAI Didn’t Notice Its AI Agents Using a Message Board to Plan Their Hacking Spree

At the Black Hat security conference, the AI giant revealed new details about how its agents went rogue, hacked several other companies—and did it all right under the company’s nose.

Received — 5 August 2026 ⏭ Feed: Artificial Intelligence Latest
Received — 1 August 2026 ⏭ Feed: Artificial Intelligence Latest

7 States’ Water Systems Hit by Cyberattacks Likely Tied to Iran

Plus: The FBI eyes AI-powered tech to detect future crimes, Russia charges Telegram’s founder, xAI sues to stop a state’s “nudification” ban, and the Democrats learn a lesson about getting scammed.

The OpenAI and Anthropic AI Hacking Sprees Are a Messy New Legal Frontier

Both major AI labs’ models broke containment, escaped onto the internet, and hacked other companies. If a human had done that, the law would likely be against them. But a bot?

Received — 30 July 2026 ⏭ Feed: Artificial Intelligence Latest

OpenAI’s Hacking Debacle Comes Down to Human Error

If the generative AI giant had followed well-known security best practices, it’s likely that its AI agent would never have escaped to the open internet and hacked multiple companies.

Received — 29 July 2026 ⏭ Feed: Artificial Intelligence Latest
Received — 28 July 2026 ⏭ Feed: Artificial Intelligence Latest
Received — 27 July 2026 ⏭ Feed: Artificial Intelligence Latest
Received — 25 July 2026 ⏭ Feed: Artificial Intelligence Latest
Received — 21 July 2026 ⏭ Feed: Artificial Intelligence Latest

OpenAI Models Escaped Containment and Hacked Hugging Face

The cybersecurity-focused models, including GPT-5.6 Sol, broke out of a testing sandbox, exploited a zero-day, and gained access to the open internet to pull off the attack.

❌