







Learn how to use Apigee AI gateway features to govern, optimize, and scale AI traffic.
Overview of model routing | API Gateway | Google Cloud Documentation
Model routing for API Gateway is a managed traffic management layer that accepts OpenAI-compatible prompt requests, transcodes them in-flight, and routes them to specific Gemini Enterprise Agent Platform models. Model routing acts as a managed alternative to client-side proxies such as LiteLLM, providing centralized infrastructure to manage the lifecycle of AI agents.
Google for Developers Blog - News about Web, Mobile, AI and Cloud
Google Cloud API Gateway now offers a model routing feature in Public Preview, allowing developers to dynamically route traffic to models like Gemini, Claude, or OpenAI OSS-GPT without hardcoding endpoints or managing open-source proxies. Developers can easily configure these routing rules directly within their OpenAPI 3.x specifications by mapping virtual model names to specific backend targets on a shared host. Once deployed, the Gateway acts as a serverless ingress layer that accepts standard OpenAI-compatible requests, automatically transcodes the payload to the native schema of the target model, and routes the traffic on the fly.

Agent Gateway overview | Gemini Enterprise Agent Platform | Google Cloud Documentation
Secure and govern AI agent connectivity with Agent Gateway. Centralize access policies, mTLS, and Model Context Protocol (MCP) security for agent-to-agent and agent-to-tool interactions across diverse runtimes.
Helicone/ai-gateway
The fastest, lightest, and easiest-to-integrate AI gateway on the market. Fully open-sourced.
OpenAPI 3.x Extensions in API Gateway | Google Cloud Documentation
API Gateway accepts a set of Google-specific extensions to the OpenAPI specification that configure the behaviors of the gateway. These extensions allow you to specify API management settings, authentication methods, quota limits, and backend integrations directly within your OpenAPI document. Understanding these extensions helps you tailor your service behavior and integrate with API Gateway features.
Google Sunsets Vertex AI, Launches Agent Control Plane | Awesome Agents
Google replaced Vertex AI with the Gemini Enterprise Agent Platform at Cloud Next 2026 - a full-stack control plane that assigns every agent a cryptographic ID and routes all tool calls through a central policy gateway.

Google Agentspace enables the agent-driven enterprise | Google Cloud Blog
Google Agentspace introduces new expert AI agents, no-code Agent Assembler, access via Chrome Enterprise, and more.

Google for Developers Blog - News about Web, Mobile, AI and Cloud
Explore how Google, Amazon, and Cisco form the Agent2Agent Foundation under the Linux Foundation to drive AI innovation via interoperability as an industry standard.

Together AI | The AI Native Cloud
Build what's next on the AI Native Cloud. Full-stack AI platform for inference, fine-tuning, and GPU clusters — powered by cutting-edge research.

Ambassador: Building a Control Plane for an Envoy-Powered API Gateway on Kubernetes
This article provides an insight into the creation of the Ambassador open source API gateway for Kubernetes, and discusses the technical challenges and lessons learned from building a developer-focused control plane for managing ingress or "edge" traffic within microservice-based applications.

Google AI Plans with Cloud Storage - Google One
Explore Google AI Plans. Access our most advanced AI, generate videos from text, and secure cloud storage.
Driving the UK’s next chapter: From AI potential to agentic reality | Google Cloud Blog
Organizations like HSBC, Ineffable Intelligence, Starling, Vodafone, and the UK and local governments are showcasing their transformation with Google Cloud at the London Summit.

AI-Native Cloud | DigitalOcean
Run AI products in production with a unified stack for agents, inference, and cloud—built for control, performance, and economics at scale.
Google Gemini Enterprise: Vertex AI Rebrand at Cloud Next 26 | AI Automation Global
Google rebranded Vertex AI as Gemini Enterprise Agent Platform at Cloud Next 26. A2A protocol v1.0, Workspace Studio, 200+ models, 150 deployments.

AI proxy: fostering a more open ecosystem - Blog - Braintrust
Introducing Braintrust's latest feature: an AI proxy that lets you use open source models like LLaMa 2 and Mistral, as well as all of OpenAI's and Anthropic's models, behind a single interface with caching, security, and API key management built in.