







Practical strategies to reduce CPU consumption of Istio Envoy sidecar proxies for cost-efficient mesh operations.
Istio 1.23 Drops the Sidecar for a Simpler 'Ambient Mesh'
This new edition of the Istio service mesh can be run without sidecars, simplifying deployments and, in some cases, even reducing latency.


From sidecars to sidecarless: Tracing the evolution of service mesh technologies with Istio and Cilium
Learn how technologies like Istio Ambient and Cilium revolutionize microservices networking, offering unprecedented capabilities in traffic management, observability, and security.

Subquadratic launches with $29M to bring 12M-token context windows to AI
Subquadratic launches with $29M to bring 12M-token context windows to AI - SiliconANGLE

The Sidecar Pattern: Why Every Major Tech Company Runs Proxies on Every Pod
“But doesn’t that mean every single request goes through an extra hop?”

Run Workers for up to 5 minutes of CPU-time
Workers now support up to 5 minutes of CPU time per request. Allowing more CPU-intensive workloads.

What Istio Got Wrong: Learnings from the Last Seven Years of Service Mesh - C. Posta, L. Ryan
Alibaba Cloud claims K8s service meshes are resource hogs
SIGCOMM 2024: Built its own replacement – Canal Mesh – that it says leaves Google's Istio and Ambient eating dust

Using Envoy for Egress Traffic
When our forward proxy did not meet evolving deployment strategies, Palantir deployed Envoy to enable granular egress traffic filtering.

Add Cloud Run example with Cloud Endpoints (ESPv2) sidecar by timburks · Pull Request #20 · apigee-apihub-demo/animals
Using Cloud Run sidecars allows us to make stunning simplifications to the documented recommendations for using Cloud Endpoints with Cloud Run, which previously required a multi-step process includ...
Fastino trains AI models on cheap gaming GPUs and just raised $17.5M led by Khosla | TechCrunch
Tech giants like to boast about trillion-parameter AI models that require massive and expensive GPU clusters. But Fastino is taking a different approach.

The Universal Execution Layer for AI
Optimize any AI model on any engine, across all hardware. Dria’s topology-aware compiler and peer-to-peer runtime merge CPUs, GPUs, NPUs & chiplets into one fabric—maximising utilisation, cutting inference cost and ending vendor lock-in.

TheStage AI – Faster, Cheaper AI Inference
Accelerate models on NVIDIA & edge. Full guides for setup, optimization & deploy. ANNA, QLIP, Elastic Models, CLI & API. Built for AI teams & devs.

Turbocharging Web Apps: Efficient AI Model Caching in Chrome
Cloud Native Live: Envoy Gateway 1.1 - Q&A Session
Virtual Event - Envoy Gateway 1.1 simplifies the management and deployment of Envoy, making it more accessible and easier to use as a Kubernetes ingress gateway.