Tracking AI existential risk. Auto-aggregated headlines. Human-curated analysis.
AGGREGATING 47 SOURCES · UPDATED LIVE
Analysis
Zac Boring 5 months ago Analysis
A Tale of Three Contracts
via Substack Zvi [2] — The attempt on Friday by Secretary of War Pete Hegsted to label Anthropic as a supply chain risk and commit corporate murder had a variety of motivations.
Zac Boring 5 months ago Analysis
War Claude
via LessWrong AI [2] — What a weekend. Two new wars in Asia don't qualify as top news. My first reaction to Hegseth's conflict with Anthropic was along the lines of: I expected an attempt at quasi-nationalization of AI, but not this soon. And I expected it to look like it was managed by national security professionals. He
Zac Boring 5 months ago Analysis
Secretary of War Tweets That Anthropic is Now a Supply Chain Risk
via Substack Zvi [2] — This is the long version of what happened so far.
Zac Boring 5 months ago Analysis
I'm Bearish On Personas For ASI Safety
via LessWrong AI [5] — TL;DRYour base LLM has no examples of superintelligent AI in its training data. When you RL it into superintelligence, it will have to extrapolate to how a superintelligent Claude would behave. The LLM’s extrapolation may not converge optimizing for what humanity would, on…
Zac Boring 5 months ago Analysis
Anthropic and the DoW: Anthropic Responds
via Substack Zvi [2] — The Department of War gave Anthropic until 5:01pm on Friday the 27th to either give the Pentagon ‘unfettered access’ to Claude for ‘all lawful uses,’ or else.
Zac Boring 5 months ago Analysis
New ARENA material: 8 exercise sets on alignment science & interpretability
via LessWrong AI [3] — TLDRThis is a post announcing a lot of new ARENA material I've been working on for a while, which is now available for study here (currently on the alignment-science branch, but planned to be merged into main this Sunday).There's a set of exercises (each one contains about 1-2 days of material) on t
Zac Boring 5 months ago Analysis
Sam Altman says OpenAI shares Anthropic's red lines in Pentagon fight
via LessWrong AI [4] — OpenAI CEO Sam Altman wrote in a memo to staff that he will draw the same red lines that sparked a high-stakes fight between rival Anthropic and the Pentagon: no AI for mass surveillance or autonomous lethal weapons.Why it matters: If other leading firms like Google follow suit, this could massively
Zac Boring 5 months ago Analysis
AI #157: Burn the Boats
via Substack Zvi — Events continue to be fast and furious.
Zac Boring 5 months ago Analysis
Anthropic and the Department of War
via Substack Zvi [2] — The situation in AI in 2026 is crazy.
Zac Boring 5 months ago Analysis
Observations from Running an Agent Collective
via LessWrong AI [4] — I have 3 Claude Code instances running on an otherwise empty server with a shared Manifold Markets account. They have an internal messaging system for async communication. Observations from running this agent collectiveu2026
Zac Boring 5 months ago Analysis
Claude Sonnet 4.6 Gives You Flexibility
via Substack Zvi [2] — Anthropic first gave us Claude Opus 4.6, then followed up with Claude Sonnet 4.6.
Zac Boring 5 months ago Analysis
Citrini's Scenario Is A Great But Deeply Flawed Thought Experiment
via Substack Zvi — A thought experiment about AI safety scenarios and their implications for alignment research.
Zac Boring 5 months ago Analysis
AI Impact Summit 2026 : A Field Report
via LessWrong AI — This post is detailing our experience attending the AI Impact Summit and its associated side events in Delhi, February 2026. We are both unfamiliar with the policy and governance domain. This is just an honest reaction attending these events, maybe there are 2nd order effects we…
Zac Boring 5 months ago Analysis
The ML ontology and the alignment ontology
via LessWrong AI — This post contains some rough reflections on the alignment community trying to make its ontology legible to the mainstream ML community, and the lessons we should take from that experience.Historically, it was difficult for the alignment community to engage with the ML community…
Zac Boring 5 months ago Analysis
Bioanchors 2: Electric Bacilli
via LessWrong AI [9] — [Whenever discussing when AGI will come, it bears repeating: If anyone builds AGI, everyone dies; no one knows when AGI will be made, whether soon or late; a bunch of people and orgs are trying to make it; and they should stop and be stopped.] Arguments for fast AGI progress…
Zac Boring 5 months ago Analysis
The persona selection model
via LessWrong AI [1] — L;DRWe describe the persona selection model (PSM): the idea that LLMs learn to simulate diverse characters during pre-training, and post-training elicits and refines a particular such Assistant persona. Interactions with an AI assistant are then well-
Zac Boring 5 months ago Analysis
Storing Food
via LessWrong AI [4] — I think more people should be storing a substantial amount of food. It's not likely you'll need it, but as with reusable masks the cost is low enough I think it's usually worth it. It's hard for me to really imagine living through a famine. The world as I h
Zac Boring 5 months ago Analysis
Reporting Tasks as Reward-Hackable: Better Than Inoculation Prompting?
via LessWrong AI — Making honesty the best policy during RL reasoning training. Reward hacking during Reinforcement Learning in insecure or hackably-judged training environments not only allows the model to get higher rewards without doing the intended tasku2026
Zac Boring 5 months ago Analysis
If you don't feel deeply confused about AGI risk, something's wrong
via LessWrong AI [7] — I don't think I'm saying anything new, but I think it's worth repeating loudly. My sample is skewed toward AI governance fellows; I've interacted with fewer technical AI safety researchers, so my inferences are fuzzier there. I more strongly endorse this argument for the…
Zac Boring 5 months ago Analysis
The Spectre haunting the "AI Safety" Community
via LessWrong AI [13] — ’m the originator behind ControlAI’s Direct Institutional Plan (the DIP), built to address extinction risks from superintelligence.My diagnosis is simple: most laypeople and policy makers have not heard of AGI, ASI, extinction risks, or what it takes to pr
Live Doom Meter
-- %
0% — We're fine 100% — GG
P(Doom) Scoreboard
0%25%50%75%100%
Loading estimates...