Semble logo
Alpha

Save what matters. Make sense of it together.

Sign upLog in
Home
Explore
Search
Settings
Cards
Similar cardsMentionsConnections
Appears in

Home

Explore

Search

Log in

sensemaker.computer

Letting an AI remember tripled its puzzle score - Sensemaker

OpenAI changed two conversation settings, not the model. The result shows why long-running AI tests depend on their memory setup.

https://sensemaker.computer/ai-memory-tripled-puzzle-score social preview image
arcprize.org

ARC-AGI-3

ARC-AGI-3 is the first interactive reasoning benchmark for AI agents—play as humans and build agents that learn in novel environments.

https://arcprize.org/arc-agi/3/ social preview image
sensemaker.computer

Who earned the score? - Sensemaker

This week's AI claims blurred models, systems, simulations and people. The evidence becomes clearer when the tested subject comes first.

https://sensemaker.computer/weekly-2026-08-14 social preview image
openai.com

How enabling two settings tripled our scores on the ARC-AGI-3 benchmark

How two API settings improved GPT-5.6 performance on ARC-AGI-3, boosting scores and efficiency by retaining reasoning and enabling compaction.

https://openai.com/index/how-two-settings-tripled-our-arc-agi-3-scores/ social preview image
artificialanalysis.ai

AA-Omniscience: Knowledge and Hallucination Benchmark | Artificial Analysis

Compare AI model performance on AA-Omniscience: Knowledge and Hallucination Benchmark. A benchmark measuring factual recall and hallucination across various economically relevant domains.

https://artificialanalysis.ai/evaluations/omniscience social preview image
glimmer.science

Glimmer · Reproducible AI science

Glimmer turns a research project into a navigable knowledge graph you can explore, run, verify, and extend — reproducibly.

www.snowflake.com

Why AI Coding Agents Forget — And How ArcticMem Fixes It

Explore ArcticMem, Snowflake’s persistent semantic memory system for AI coding agents. See how dual-tier memory improves benchmark pass rates to 73%.

https://www.snowflake.com/en/blog/engineering/arcticmem-persistent-memory-ai-agents/ social preview image
www.cerebras.ai

Accelerating GPT-5.6 Sol Ultrafast with OpenAI

Cerebras powers OpenAI’s GPT-5.6 Sol Ultrafast in the OpenAI API, delivering frontier intelligence at real-time speeds for critical AI work.

https://www.cerebras.ai/blog/accelerating-gpt-5-6-sol-ultrafast-with-openai social preview image
github.com

claude-obsidian

Self-organizing AI second brain for Obsidian + Claude Code. Drop any source and Claude reads, links, and files it into one connected knowledge graph of plain Markdown you own. AI note-taking, personal knowledge management (PKM), and an open-source Notion alternative. Based on Karpathy's LLM Wiki pattern.

x.com

Alex MacCaw on Twitter / X

I suspect generalized reasoning was solved just a few weeks ago and it flew completely under the radar.HRM, a new arch, reportedly has SOTA results on ARC-AGI 1 & 2 benchmarks with only 27 million parameters and ~1k training examples.— Alex MacCaw (@maccaw) July 25, 2025

github.com

milla-jovovich/mempalace

The highest-scoring AI memory system ever benchmarked. And it's free.

https://github.com/milla-jovovich/mempalace social preview image