







Comparing formats like GGUF, GPTQ, and AWQ, with different bitwidths
LLM Leaderboard 2026 — Compare Top AI Models
Compare the latest LLM benchmarks for GPT, Claude, Gemini and more. Updated rankings across reasoning, coding, math, and multilingual tasks with pricing and speed data.
LLM Leaderboard - Comparison of over 100 AI models from OpenAI, Google, DeepSeek & others
Comparison and ranking the performance of over 100 AI models (LLMs) across key metrics including intelligence, price, performance and speed (output speed - tokens per second & latency - TTFT), context window & others.


Reverse-engineering GGUF | Post-Training Quantization
A Visual Guide to Quantization
Exploring memory-efficient techniques for LLMs

AI Leaderboard 2026: Compare & Rank 300+ Top AI Models by Intelligence, Speed & Price
The AI Leaderboard — independent rankings of GPT, Claude, Gemini, Llama, DeepSeek and 300+ AI models by intelligence, speed and price. Composite LLM Stats Score updated continuously from public benchmarks and live API metrics.

On-Device LLM Leaderboard
Intelligence × decode speed × memory × quantization retention, under real iPhone limits. Same protocol for every model; Apple's built-in FM on the board.

LLM Rankings | OpenRouter
LLM rankings and AI leaderboard based on benchmarks and real usage data from millions of users. See which AI models developers actually use.
Thireus/GGUF-Tool-Suite
Produce your own Dynamic 3.0 Quants and achieve optimum accuracy & SOTA quantization performance! Input a target size and the toolchain will create a GGUF recipe tuned to your hardware within seconds — flexible model sizing and lowest achievable perplexity/kld for GGUF enthusiasts seeking precise and automated dynamic quant production.
Sebastian Raschka on Twitter / X
While waiting for DeepSeek V4 we got two very strong open-weight LLMs from India yesterday.There are two size flavors, Sarvam 30B and Sarvam 105B model (both reasoning models).Interestingly, the smaller 30B model uses “classic” Grouped Query Attention (GQA), whereas the… https://t.co/OiJVkDCYNz pic.twitter.com/0uqmLxofRE— Sebastian Raschka (@rasbt) March 7, 2026

AI Model Leaderboards & Benchmarks
Explore leaderboards with expert-driven LLM benchmarks and updated AI model rankings across coding, reasoning and more.
Benchmarks | EXO
Transparent benchmarks for LLMs tested on real hardware. Coming soon.
A 4-Bit Model and a 1-Bit Index
Running NVFP4 Nemotron on a CPU, then mapping every embedding-compression method at matched byte budgets. The two quantizations compose.

chad/whichlang
What programming language do LLMs default to when you don't tell them? A small benchmark.
Best LLM for Coding 2026 | AI Coding Model Rankings & Benchmarks
Which AI model writes the best code? We rank every major LLM — open and closed source — across SWE-bench, HumanEval, LiveCodeBench, and Terminal-Bench coding benchmarks. Compare the best LLMs for coding, software engineering, and programming.
