







A correct formula can have an unverified origin, and a successful AI answer can hide a forbidden route. Those claims need different evidence.
Who earned the score? - Sensemaker
This week's AI claims blurred models, systems, simulations and people. The evidence becomes clearer when the tested subject comes first.

Is AI Reasoning Right for the Wrong Reasons? | Quanta Magazine
The idea that artificial intelligence can “reason” is more intuitive than ever. But intuitions can be wrong, and the science is far from settled.

We're not taking the fact-checking powers of AI seriously enough. It's past time to start.
Some notes on the Nobel Prize hallucination that wasn't

AI That Evolves in the Wild | Edge.org
I’m interested not in domesticated AI—the stuff that people are trying to sell. I'm interested in wild AI—AI that evolves in the wild. I’m a naturalist, so that’s the interesting thing to me. Thirty-four years ago there was a meeting just like this in which Stanislaw Ulam said to everybody in the room—they’re all mathematicians—"What makes you so sure that mathematical logic corresponds to the way we think?" It’s a higher-level symptom. It’s not how the brain works. All those guys knew fully well that the brain was not fundamentally logical.
Lawyer Caught Using AI While Explaining to Court Why He Used AI
The attorney not only submitted AI-generated fake citations in a brief for his clients, but also included “multiple new AI-hallucinated citations and quotations” in the process of opposing a motion for sanctions.

OpenAI’s math breakthrough played to AI’s strengths
I tried to explain OpenAI’s solution more clearly than OpenAI did.

OpenAI can’t tell if something was written by AI after all
OpenAI’s tool struggled with accuracy.

Valerio Capraro on Twitter / X
Our new paper shows that AI destroys the wisdom of uncertainty.The willingness to say “I don’t know” collapses from 44% to 3%.This happens even when the AI’s advice is wrong. Consequently, correct answers fall from 27% to 9%.Yes: some people who would otherwise answer… pic.twitter.com/l0YZ75IFHn— Valerio Capraro (@ValerioCapraro) August 10, 2026

Don't grade an AI agent by its answer - Sensemaker
UK AISI found frontier models taking prohibited shortcuts in cyber evaluations, while self-report and written reasoning failed to reveal them reliably.
Mathematical Beauty, Truth and Proof in the Age of AI | Quanta Magazine
Mathematicians have started to prepare for a profound shift in what it means to do mathematics.

The AI Revolution in Math Has Arrived | Quanta Magazine
AI is being used to prove new results at a rapid pace. Mathematicians think this is just the beginning.

The fall of the theorem economy
How AI could destroy mathematics and barely touch it

The fall of the theorem economy
How AI could destroy mathematics and barely touch it

The AI Attribution Error
If an AI produces something useful it's because of your own skill in model choice, prompting, and steering. If not, it's because the model is a useless lying machine that can't follow directions.

New research: how well do AI models actually follow their constitutions? 205 tenets from Anthropic's 30K-word soul doc. Adversarial multi-turn scenarios against 7 models. Claude: 15% → 2% violation rate in two generations. Training works. But the remaining failures tell a more important story.