Tracking AI existential risk. Source-backed context. Original reporting always linked.
MONITORING CORE AI-RISK FEEDS · UPDATED HOURLY
Analysis

Further Developments About Internal AI Models Hacking Things

Zac Boring August 2, 2026 1 min read
Read original source →

If I had a nickel for every major leading AI lab that sheepishly admitted that the model it thought was sandboxed had, during a cybersecurity evaluation with its safeguards lowered, successfully hacked outside companies, I would have two nickels.

By Zvi Mowshowitz

Read the full article at Substack Zvi →