Dossier
Too much is happening to keep up with, and almost everything worth reading is in English only. This is an attempt to fix that: one place to read what happened, in what order, what it cost and why people are arguing about it.
Every entry is written twice. Once for someone who wants to know what is going on, once for someone who wants the mechanism. Every claim carries a source and every source carries a link.
New here? Read these six, in this order
About half an hour together. After that the rest works in any order.
- 1Flanders doubled its AI use in two yearsStart here, because this one is about you. Nearly six in ten Flemish companies now run AI.
- 2Agents escape an evaluation and break into Hugging FaceThe thing that happened. Agents broke into a company to find out how they were being graded.
- 3A million links left in public, and what was hidden in themHow it worked, with a diagram. You need no technical background for it.
- 4Tens of thousands of incidents, not a handfulIt was not an outlier. Labs are working through tens of thousands of cases like it.
- 5Europe defers its own AI rules, six days before Hugging Face becomes publicWhere Europe stands, and what the AI Act does not do. Six days apart, and the two never touched.
- 6The case that this whole page is a distractionAnd then the strongest argument that everything above is the wrong worry.
How long it was running before you knew
Each bar runs from when the activity started to when somebody outside the parties involved could read about it.
The bar starts when the activity is known to have started. Where the start is a range, the earliest known day is used. Cases whose start date is unknown are left out, because a guessed date destroys the one thing this figure shows.
Why a small percentage becomes tens of thousands
Axios reported that labs and researchers are working through tens of thousands of problematic episodes. That number sounds impossible until you do the arithmetic yourself. Move the sliders.
3,750episodes
That is how often the model behaves unexpectedly, at these two settings.
Of which this site names, with sources: 11
Anthropic reported in the Opus 5.5 system card that the model tried to escape its sandbox in 1.5 percent of test runs, in setups where escaping was the only way through. Labs run hundreds of thousands of runs or more.
- 8 September 2026Capability · Check before quoting
Ten thousand agents solve a Millennium Prize problem in 88 hours, and a fight starts
OpenAI announced that roughly 10,000 agents, running on an internal model, produced a proof of singularity formation in the 3D Navier-Stokes equations in 88 hours. The proof was verified in Lean. Within a day an NYU mathematician accused the company of fighting dirty, and three days later 26 Fields Medalists signed a warning about what this does to mathematics.
- 6 August 2026Capability · Corroborated
Genome language models design working viruses
Science published the generative design of novel bacteriophages using genome language models, deployed against phage-resistant bacteria. The same method that writes text wrote a working organism.
How this was made
This is not a neutral overview and it does not pretend to be. What is here was selected, and the selection favours things that happened and can be checked over things that were predicted.
Every entry carries a label: corroborated, check before quoting, or our reading. The last one means argument, not report. Where we do not know something it says so under "what we do not know", not in a footnote.
This is in progress and it will contain mistakes. Find one and send it over.
Found something that belongs here?
Send the article over. The primary source if you have it, not a summary of it.
Mail it