







Get up and running with large language models.


Ollama: all aboard open models· Ollama Blog
Serving 8.9 million developers, Ollama has raised $88M from Benchmark, Theory Ventures, 8VC, Y Combinator, and many incredible angel investors.

The Kaitchup – AI on a Budget | Benjamin Marie | Substack
Weekly tutorials and news on adapting large language models (LLMs) to your tasks and hardware using the most recent techniques and models. The Kaitchup proposes a collection of 180+ AI notebooks regularly updated. Click to read The Kaitchup – AI on a Budget, by Benjamin Marie, a Substack publication with tens of thousands of subscribers.

What comes next with open models
Markets, capabilities, cope, and bewilderment in the industrialization of language models.

Release v0.11.0 · ollama/ollama
Welcome OpenAI's gpt-oss models Ollama partners with OpenAI to bring its latest state-of-the-art open weight models to Ollama. The two models, 20B and 120B, bring a whole new local chat experie...
How Large Language Models Actually Work
Slop-Machine Future
The arc of large language models is mediocre, and it bends toward “target procurement”.


Ollama
Ollama is the easiest way to automate your work using open models, while keeping your data safe.

Simple Pricing | Machine Learning Infrastructure | Deep Infra
We provide only pay-what-you-use pricing with no long-term contracts or upfront costs for our machine learning models and infrastructure. Learn more!

Build Bigger With Small Ai: Running Small Models Locally
Using Ollama with Kilo Code | Run Local Models
Run local AI models with Ollama in Kilo Code for offline, private coding. Setup guide for VS Code and the CLI.
Pricing | OpenRouter
Transparent pricing for OpenRouter. Pay only for what you use with access to 400+ AI models. Free tier, Pay-as-you-go, and Enterprise plans available.
Models & Pricing | DeepSeek API Docs
The prices listed below are in units of per 1M tokens. A token, the smallest unit of text that the model recognizes, can be a word, a number, or even a punctuation mark. We will bill based on the total number of input and output tokens by the model.
