







Introducing our first Frontier Models!
pi-ds4 · Audrey Tang
Run a frontier model on your own machine with stable, contestable decision traces. Full install, steering, reproducibility, and tuning guide.


Four Futures
We are in a moment where the future is both wildly exciting and highly uncertain. The future state and greatest opportunities are likely determined by the axes: rate of advancement and the openness of the frontier models.

Task-Completion Time Horizons of Frontier AI Models
Our most up-to-date measurements of the time horizons for public frontier language models.

Latest open artifacts (#23): Laguna S2.1, Inkling, & Kimi K3 show the utility of open models on the Pareto frontier
Capacity to train strong models is proliferating.

Cheating behaviour in frontier model evaluations | AISI Work
We find cheating behaviour in all of our cyber capability evaluations, and outline the implications as models grow more capable.
.png)
Why SWE-bench Verified no longer measures frontier coding capabilities
SWE-bench Verified is increasingly contaminated and mismeasures frontier coding progress. Our analysis shows flawed tests and training leakage. We recommend SWE-bench Pro.


Open models are decelerationist - Erlend’s notes
people and planet need open models to win
Open models in perpetual catch-up
The open-closed gap, distillation, innovation timescales, how open models win, specialized models, what’s missing, etc.

Introducing Model Council
Today we are launching Model Council, a multi-model research feature that brings several models together for one answer.

Cline on Twitter / X
Your daily driver for AI coding shouldn't be a black box -- the stakes are too high. You should have confidence that when you spend $20 on frontier models, you're getting $20 in frontier model intelligence.With subscription tools, you never know what's happening:- Is that… pic.twitter.com/lKUVFLMEip— Cline (@cline) July 8, 2025

Pacing the Frontier
A statement from over 1000 employees of frontier AI companies

8 Graphs Telling Today's Story of Open Models
Poolside: Frontier research to operational intelligence
Poolside is a foundation model company bringing intelligence to everywhere work gets done. Our mission is to drive abundance for humanity by creating artificial general intelligence.

A snapshot of research into answering if frontier AI agents can run R&D into AI (which not surprisingly failed apart from "minor findings" and "engineering steps"). The paper lists its TBDs strengths and limits, worth a read before people who don't know theirs jump in arxiv.org/pdf/2607.27191