







A simple and easy-to-use wrapper for working with GGML
sgl-project/sglang
SGLang is a high-performance serving framework for large language models and multimodal models.
ggml
AI inference at the edge. ggml has 22 repositories available. Follow their code on GitHub.
llama.cpp/tools/server at master · ggml-org/llama.cpp
LLM inference in C/C++. Contribute to ggml-org/llama.cpp development by creating an account on GitHub.
Motion Core — Svelte-native motion toolkit with GSAP and OGL-powered components, demos, and CLI-driven workflows
Svelte-native motion toolkit with GSAP and OGL-powered components, demos, and CLI-driven workflows.

llama.cpp/docs/build.md at master · ggml-org/llama.cpp
LLM inference in C/C++. Contribute to ggml-org/llama.cpp development by creating an account on GitHub.
The mythical matched modules | Proceedings of the 24th ACM SIGPLAN conference companion on Object oriented programming systems languages and applications
Certified compilers are complex software systems. Like other large systems, they demand modular, extensible designs. While there has been progress in extensible metatheory mechanization, scaling extensibility and reuse to meet the demands of full ...

Rémi in 🌁 for AIEF on Twitter / X
Projects like @antirez's ds4.c show that we can squeeze a lot of performance out of model-specific implementations instead of mapping everything back to generic ggml nodes.This probably means re-thinking the whole inference server: one central controller, many model runners as…— Rémi in 🌁 for AIEF (@remilouf) May 15, 2026
Visualizing GL_NV_shader_sm_builtins
Using GL_NV_shader_sm_builtins to visualize Streaming Multiprocessors and Warps
OpenSpec — A lightweight spec‑driven framework
OpenSpec is a lightweight, spec‑driven framework for coding agents and CLIs — universal, open source, and no API keys or MCP required.
Building personal tools by programming

Introducing gpt-oss
We’re releasing gpt-oss-120b and gpt-oss-20b—two state-of-the-art open-weight language models that deliver strong real-world performance at low cost. Available under the flexible Apache 2.0 license, these models outperform similarly sized open models on reasoning tasks, demonstrate strong tool use capabilities, and are optimized for efficient deployment on consumer hardware.

GGML and llama.cpp join HF to ensure the long-term progress of Local AI
We’re on a journey to advance and democratize artificial intelligence through open source and open science.

Swival — A coding agent for any model
A CLI coding agent built to be reliable with smaller models, including local ones. Designed for tight context windows and limited resources. Works with LM Studio, llama.cpp, HuggingFace, OpenRouter, Google Gemini, Vertex AI, ChatGPT Plus/Pro, AWS Bedrock, and any OpenAI-compatible server.

SLOP - An alternative to MCP. github.com/agnt-gg/slop