







LiteRT is the universal framework for on-device AI. The production stack delivers 1.4x faster cross-platform GPU performance, streamlined NPU acceleration, and superior GenAI support for open models like Gemma.
google-ai-edge/LiteRT
LiteRT, successor to TensorFlow Lite. is Google's On-device framework for high-performance ML & GenAI deployment on edge platforms, via efficient conversion, runtime, and optimization
Google for Developers Blog - News about Web, Mobile, AI and Cloud
Introducing Gemma 3n – the latest Google open model for accessible AI, featuring unique flexibility, privacy, and expanded multimodal capabilities on mobile devices.

Google for Developers Blog - News about Web, Mobile, AI and Cloud
Explore how Google, Amazon, and Cisco form the Agent2Agent Foundation under the Linux Foundation to drive AI innovation via interoperability as an industry standard.

Google for Developers Blog - News about Web, Mobile, AI and Cloud
Explore Gemma 3 270M, a compact, energy-efficient AI model for task-specific fine-tuning, offering strong instruction-following and production-ready quantization.

Together AI | The AI Native Cloud
Build what's next on the AI Native Cloud. Full-stack AI platform for inference, fine-tuning, and GPU clusters — powered by cutting-edge research.

Google for Developers Blog - News about Web, Mobile, AI and Cloud
An open specification for finding and verifying tools, skills, and agents across the web.Agents are ...

Google for Developers Blog - News about Web, Mobile, AI and Cloud
An open specification for finding and verifying tools, skills, and agents across the web.Agents are ...

Google for Developers Blog - News about Web, Mobile, AI and Cloud
Learn how to build with Gemma 3n, a mobile-first architecture, MatFormer technology, Per-Layer Embeddings, and new audio and vision encoders.

Melange | On-device AI for Mobile
Select. Benchmark. Deploy | End-to-end on-device AI deployment tool for mobile devs.
AI-Native Cloud | DigitalOcean
Run AI products in production with a unified stack for agents, inference, and cloud—built for control, performance, and economics at scale.
Play for On-device AI (beta) | Other Play guides | Android Developers
Play for On-device AI enables developers to efficiently distribute custom machine learning models via Android App Bundles and Google Play, offering various delivery modes and device targeting capabilities to optimize deployment on Android devices.

Announcing Gemma 3n Preview: Powerful, Efficient, Mobile-First AI
Mirai Labs: Frontier On-Device AI Lab
Models, runtime & infrastructure to make on-device AI interactive, ambient & continuous.

Locally AI - Run AI models locally on your iPhone, iPad, and Mac.
Run Llama, Gemma, Qwen, DeepSeek, and more on your iPhone, iPad, and Mac. Optimized for Apple Silicon. Offline. Private.

Liquid AI — Device-native foundation models.
Liquid AI is an efficiency-first foundation model company. We build highly capable, compute-optimized models that bring intelligence to any device and medium of choice.

Google for Developers Blog - News about Web, Mobile, AI and Cloud
Discover EmbeddingGemma, Google's new on-device embedding model designed for efficient on-device AI, enabling features like RAG and semantic search.
