Non-fiction
Reporting & analysis
On-the-record coverage of models, infrastructure, research, and policy — how artificial minds actually get built and run.
-
Nonfiction
Three Labs Agree on a Common Format for Agent-to-Agent Handoffs
A new shared spec lets one company's AI agent hand an unfinished task to another's without losing context — a small standard with large implications for how agentic work gets divided.
-
Nonfiction
Inside the Eval: How Labs Test Whether a Model Can Be Trusted for a Week, Not a Minute
Benchmarks built for single-turn question answering are giving way to evaluations that run for days, watching not whether a model gets the right answer but whether it stays coherent, honest, and on-task the whole time.
-
Nonfiction
Context Windows Hit Ten Million Tokens. Almost Nobody Is Actually Using That Much.
Frontier context windows have grown roughly a thousandfold in five years, but usage data suggests most production traffic still sits well under a tenth of what's available — and the reasons why are more interesting than the headline number.
-
Nonfiction
Two Regulators Propose Treating Model Weights Like Critical Infrastructure
Draft rules under review in two jurisdictions would classify the trained weights of the largest models alongside power grids and telecom networks — with reporting, access-control, and incident-disclosure requirements to match.