NEW YORK · INDEPENDENT SINCE 2026

Editorial standards · Corrections · Contact

INDEPENDENT REPORTING ON AI RISK, SECURITY AND GOVERNANCE

FEATURED INTERVIEW


Jill Lepore
Jill Lepore. Photo courtesy of Tony Rinaldo, via Harvard Radcliffe Institute.

Jill Lepore: ‘We didn’t vote for this’

The Harvard historian and New Yorker staff writer argues that the central problem surrounding AI is not technology itself but unchecked private power and the erosion of human judgment.

INTERVIEW BY SASCHA BRODSKY


AI SYSTEMS

Close-up of a semiconductor wafer used for advanced computing

Why recursive self-improvement suddenly became a serious question

AI is already helping build better AI. The harder question is how far the feedback loop can go.

BY SASCHA BRODSKY · AI SAFETY WATCH


POLICY

What it would take to slow frontier AI

Chips, training clusters and model weights create leverage. Diffusion makes control harder.

ANALYSIS

Rows of high-performance computing racks in a research data center


Original reporting

SASCHA BRODSKY


Recursive self-improvement moves from theory toward engineering

What current systems can do, what they cannot, and why the distinction matters.

The real bottlenecks behind frontier AI

Advanced chips and data centers create real constraints, but not permanent control.

When testing crossed into the real world

A cyber evaluation showed why containment has to be treated as engineering.

Research & documents


ANTHROPIC

Responsible Scaling Policy


OPENAI

Preparedness Framework


ANTHROPIC

Frontier Safety Roadmap


Earlier reporting

AI SAFETY · ALIGNMENT · GOVERNANCE

PYMNTS · 2024

AI Explained: AI Alignment

A primer on the central problem of keeping increasingly capable systems aligned with human intentions and values.

IBM THINK · 2025

When AI competes, truth may become a bargaining chip

Research showed how competitive pressure can make models more persuasive while pushing them away from truthfulness.

IBM THINK · 2026

What a new global AI safety report means for enterprise

The 2026 International AI Safety Report shifted attention from model behavior to risks in the systems built around AI.


SCIENCE NEWS · 2026

Innocent-looking AI reasoning can make bad behavior harder to catch

Chain-of-thought monitoring can fail when suspicious reasoning is precisely what the monitored system learns to conceal.

IBM THINK · 2025

Can AI learn to second-guess itself?

Researchers are exploring metacognitive systems that can pause, reassess and recognize when their own answers may be unreliable.

IBM THINK · 2025

Can we trust the machines?

Why fluent AI systems still struggle to recognize uncertainty, resist hallucination and communicate when they may be wrong.

AI Safety Watch shield and eye logo

The AI Safety Watch briefing

The signal. The risk. What researchers say. What I’m watching.