Who earned the score? - Sensemaker
This week's AI claims blurred models, systems, simulations and people. The evidence becomes clearer when the tested subject comes first.
AI Index | Stanford HAI
The mission of the AI Index is to provide unbiased, rigorously vetted, and globally sourced data for policymakers, researchers, journalists, executives, and the general public to develop a deeper understanding of the complex field of AI. To achieve this, we track, collate, distill, and visualize dat
A result does not tell you how it was made - Sensemaker
A correct formula can have an unverified origin, and a successful AI answer can hide a forbidden route. Those claims need different evidence.
Lawyer Caught Using AI While Explaining to Court Why He Used AI
The attorney not only submitted AI-generated fake citations in a brief for his clients, but also included “multiple new AI-hallucinated citations and quotations” in the process of opposing a motion for sanctions.

Developer's Honest Assessment of AI at Work Rattles the Official Narrative
A veteran programmer was praised after sharing his brutally honest thoughts about AI's impact on work and productivity.

AI-Summarized News Articles: Readers Want Clear, Reliable Sources
A survey in Japan found that readers of AI-summarized articles on an app felt that having clearly cited sources was the most important factor in assessing reliability.

AI History in Quotes
Each of these presents the clearest, earliest, or most-cited statement of a specific idea in AI.


Standards around generative AI
Accuracy, fairness and speed are the guiding values for AP’s news report, and we believe the mindful use of artificial intelligence can serve these values and over time improve how we work.
Ed Zitron's AI prediction track record
Highly recommend this read from Trezy. I think he did a great job recognizing what exactly felt off during the at. report. Not really from the AI itself, but from how it was midhandled in the eyes of the community it was in front of.
Trezy
Finally done. I was able to talk with @pfrazee.com, as well as a handful of other folks with insight on the situation. #atmosphereconf trezy.com/blog/the-marshmallow-test
If you're going to cite that NBER report from OpenAI about "how people use AI," you've got to at least caveat it. We have no way to directly verify most of the report, and we do have at least one good reason, via indirect evidence and quoted below, to not trust it. nber.org/papers/w34255
Maria Antoniak
Not the point of the OP, but I was curious and took a closer look at this 2025 piece from OpenAI. Much could be said, but it fails my go-to test for all of these pieces about "how people actually use AI": the tokens "sex" "erotic" and "NSFW" do not appear in the paper.