# solo.io > AI-optimized mirror of solo.io containing 50 pages totalling 119,441 words of clean markdown content, structured data, and semantic HTML. Original source: https://solo.io/. Last updated: 2026-06-13T22:58:34.318Z. Each page is available as HTML (with JSON-LD structured data) and Markdown (text-only, ideal for LLMs and RAG). ## Homepage - [Istio](/content/site-root.html): Solo.io is the unified platform for secure, seamless cloud operations - built to power the next generation of intelligent AI agents. (87 words) ## Articles & Blog Posts - [SOLO ENTERPRISE FOR](/content/topics/rate-limiting/index.html): Rate limiting is a technique used to control the rate at which requests are made to a network, server, or other resource. (25 words) - [SOLO ENTERPRISE FOR](/content/topics/ai-connectivity/llm-traffic-governance-gateway-strategies-for-secure-ai.html): Learn how centralized API gateways streamline LLM traffic governance, enforce security & compliance, and accelerate AI adoption. (67 words) - [SOLO ENTERPRISE FOR](/content/company/careers/index.html): Help deliver the future of cloud connectivity. Check our open positions. (22 words) - [Cloud connectivity done right](/content/press-releases/solo-io-advances-application-networking-market-with-135-million-series-c-funding.html): Solo.io receives $135M in Series C funding from Altimeter, Redpoint Ventures, True Ventures to help growth the company. (27 words) - [Quickly Adopting New Models](/content/blog/llms-in-the-enterprise-overcoming-cost-security-and-observability-challenges-with-nvidia-nim-and-gloo-ai-gateway.html): Learn how enterprises can scale LLM adoption while ensuring cost control, security, and observability with NVIDIA NIM and Gloo AI Gateway. Discover how to manage model switching, enforce guardrails, prevent cost overruns, and optimize LLM performance in a Kubernetes environment. (5,811 words) - [Prerequisites](/content/blog/kagent-3-agent-substrate-a-101-installation-configuration-guide.html): Learn how to run AI agents on Kubernetes with Agent Substrate and kagent. This step-by-step guide covers architecture, installation, configuration, and deploying agent workloads using the Substrate runtime. (5,232 words) - [How to: Building Agentgateway to support Multi-LLM providers.](/content/blog/getting-started-with-multi-llm-provider-routing/index.html): Agentgateway makes it simple to route traffic to multiple LLM providers through a single gateway using the Kubernetes Gateway API. This guide walks through setting up agentgateway OSS on a local Kind cluster with xAI, Anthropic, and OpenAI backends, all routed through a listener named llm-providers. (5,701 words) - [Istio Service Mesh](/content/blog/platform-engineering-essential-tools/index.html): Explore our list of the essential tools that empower platform engineers to do their best work. (5,818 words) - [Native capabilities and the reference implementation](/content/blog/building-real-time-ai-cost-controls-with-agentgateway.html): Discover how to build a dollar-denominated AI cost-governance layer on top of agentgateway's native External Processing support. The reference implementation adds hierarchical budgets, pre-request reservations, actual-cost settlement, forecasting, approvals, rate limits, and a management UI. (7,267 words) - [The “why”](/content/blog/keeping-context-and-tokens-low-with-progressive-disclosure-in-agentgateway.html): Learn how to cut MCP token usage by 91% using agentgateway’s progressive disclosure. Reduce cost, control context bloat, and optimize agent workflows. (5,889 words) - [Context awareness at every layer of agentic infrastructure](/content/blog/kagent-enterprise/index.html): Discover how Solo.io's enterprise version of kagent extends Kubernetes to turn cloud-native infrastructure into agent-native infrastructure. (5,169 words) - ["Legacy" proxies](/content/blog/context-aware-security-ai-gateways/index.html): Learn why agentic systems require context-aware gateways to secure and route LLM, MCP, inference, and agent traffic across modern platforms. (5,736 words) - [Why is llm-d Needed?](/content/blog/llm-d-distributed-inference-serving-on-kubernetes/index.html): RedHat, Google, and IBM have launched llm-d, an open-source distributed inference platform built on vLLM and powered by kgateway. Designed to address the inefficiencies of LLM inference, llm-d introduces disaggregated serving, intelligent routing via the Kubernetes Gateway API Inference Extension, and GPU-aware scheduling to significantly boost performance and reduce costs for AI workloads. (5,047 words) - [Why Agent Substrate?](/content/blog/agent-substrate-powers-kubernetes-agents-with-kagent.html): Kubernetes transformed how we run services. Agent Substrate and kagent are transforming how we run AI agents with fast startup, efficient resource utilization, and secure execution. (5,346 words) - [Agent RPC: A Foundation, Not a Full Stack](/content/blog/why-do-we-need-a-new-gateway-for-ai-agents/index.html): Discover why traditional API gateways can't meet the needs of agentic systems built on MCP and A2A protocols. Learn how Agent Gateway provides secure, scalable, stateful communication for AI agents in enterprise environments. (5,283 words) - [What is Istio Ambient Mode?](/content/blog/istio-more-for-less/index.html): Optimize resources whilst simplifying your operations across your service mesh with Istio Ambient mesh - a sidecarless dataplane deployment mode. (6,640 words) - [Goodbye to the service proxy?](/content/blog/ebpf-for-service-mesh-yes-but-envoy-is-here-to-stay.html): eBPF can bring new functionality to Istio Service Mesh, but Envoy sidecar and Ambient mesh sidecar less modes are here to stay. (6,147 words) - [Goodbye to the service proxy?](/content/blog/ebpf-for-service-mesh-zslgt/index.html): eBPF can bring new functionality to Istio Service Mesh, but Envoy sidecar and Ambient mesh sidecar less modes are here to stay. (6,124 words) - [SOLO ENTERPRISE FOR](/content/products/istio/index.html): Deploy Istio faster with sidecarless ambient mode, built-in mTLS, and production-grade support. Secure and observe every workload with zero code changes. (636 words) - [Solo Enterprise for Kgateway](/content/products/kgateway/index.html): Standardize and scale API traffic across your platform. Kgateway provides reliable routing, security, and observability with Kubernetes-native control. (1,195 words) - [SOLO ENTERPRISE FOR](/content/blog/index.html): Fully connect your APIs and services from end – to end user and win in the cloud-native era. (24 words) - [SOLO ENTERPRISE FOR](/content/resources/case-study/index.html): Fully connect your APIs and services from end – to end user and win in the cloud-native era. (22 words) - [SOLO ENTERPRISE FOR](/content/resources/lab/index.html): Fully connect your APIs and services from end – to end user and win in the cloud-native era. (22 words) - [Solo.io products](/content/company/about-us/index.html): Learn about the story, mission, and passionate people behind everything we do. (115 words) - [SOLO ENTERPRISE FOR](/content/topics/nginx/index.html): NGINX is open source software that powers web servers and enables reverse proxying, caching, load balancing, and media streaming. (25 words) - [SOLO ENTERPRISE FOR](/content/resources/video/index.html): Fully connect your APIs and services from end – to end user and win in the cloud-native era. (22 words) - [SOLO ENTERPRISE FOR](/content/resources/report/index.html): Fully connect your APIs and services from end – to end user and win in the cloud-native era. (22 words) - [SOLO ENTERPRISE FOR](/content/topics/istio/index.html): Istio is a leading, open source platform for service mesh, which is an important infrastructure for a new generation of microservices applications. (24 words) - [SOLO ENTERPRISE FOR](/content/topics/openshift/index.html): Red Hat OpenShift is an open source platform for developing, deploying, and managing containerized applications. (24 words) - [SOLO ENTERPRISE FOR](/content/partners/index.html): Learn how your business could benefit from becoming a Gloo Network Partner. Discover the opportunities that introducing Gloo Platform delivers. (22 words) - [SOLO ENTERPRISE FOR](/content/resources/ebook/index.html): Fully connect your APIs and services from end – to end user and win in the cloud-native era. (22 words) - [SOLO ENTERPRISE FOR](/content/topics/microservices/index.html): Microservices are a software architectural style that structures an application as a collection of small, independent services. (27 words) - [SOLE ENTERPRISE FOR](/content/topics/index.html): Learn more about industry topics with our series of articles. (22 words) - [SOLO ENTERPRISE FOR](/content/resources/infographic/index.html): Fully connect your APIs and services from end – to end user and win in the cloud-native era. (22 words) - [SOLO ENTERPRISE FOR](/content/topics/api-management/index.html): API management is the process of deploying, controlling, and monitoring APIs to ensure security, performance, and seamless integration across applications and data. (26 words) - [SOLO ENTERPRISE FOR](/content/customers/index.html): Organizations of all sizes across every industry rely on Solo for their modern application networks and to stay at the forefront of cloud connectivity. (22 words) - [SOLO ENTERPRISE FOR](/content/resources/white-paper/index.html): Fully connect your APIs and services from end – to end user and win in the cloud-native era. (22 words) - [The Ultimate Guide to API Gateways: 5 Key Capabilities You Need to Know](/content/topics/api-gateway/index.html): An API gateway secures, manages, and routes API traffic, acting as a single access point for external consumers and internal microservices. (34 words) - [SOLO ENTERPRISE FOR](/content/resources/resource-kit/index.html): Fully connect your APIs and services from end – to end user and win in the cloud-native era. (22 words) - [SOLO ENTERPRISE FOR](/content/resources/webinar/index.html): Fully connect your APIs and services from end – to end user and win in the cloud-native era. (22 words) - [Agentic Infrastructure](/content/resources/datasheet/index.html): Fully connect your APIs and services from end – to end user and win in the cloud-native era. (22 words) - [Agentic Infrastructure](/content/resources/index.html): Fully connect your APIs and services from end – to end user and win in the cloud-native era. (22 words) - [SOLO ENTERPRISE FOR](/content/topics/nginx/nginx-rate-limiting/index.html): NGINX’s rate-limiting feature employs the leaky bucket algorithm typically used in packet-switched computer networks and telecommunications. (27 words) - [Indirect Prompt Injection Example](/content/blog/mitigating-indirect-prompt-injection-attacks-on-llms.html): Indirect prompt injection attacks can manipulate AI systems by embedding malicious instructions in trusted data sources. Learn how enterprises can mitigate these threats using OWASP best practices and AI Gateways to enforce security policies, validate data, and protect LLM-powered applications. (5,396 words) - [An Agent Mesh](/content/blog/agent-mesh-for-enterprise-agents/index.html): The industry's first complete connectivity solution for AI agent ecosystems supports both agent-to-agent and agent-to-tool communication across any environment (5,993 words) - [Why Dynamic Agent Discovery Matters](/content/blog/agent-discovery-naming-and-resolution-the-missing-pieces-to-a2a.html): While the A2A specification provides the critical first steps toward discovery with Agent Cards, the infrastructure for truly dynamic, scalable agent ecosystems requires additional components that the spec intentionally leaves “up to you.” In this blog, we dig into those missing pieces. (5,695 words) - [**Prerequisites**](/content/blog/security-holes-in-mcp-servers-and-how-to-plug-them/index.html): Learn how to close the major security gaps in Model Context Protocol (MCP) with a proper AI Gateway. This guide walks you through deploying MCP Servers on Kubernetes, adding authentication, locking down tools, and strengthening your organization’s overall MCP security posture. (5,718 words) - [**Meet Our Inference Workload Cluster**](/content/blog/deep-dive-into-llm-d-and-distributed-inference/index.html): Digging into the llm-d project and how it does distributed inference. (5,855 words) - [Understanding Cilium’s Control Plane Architecture](/content/blog/scaling-cilium-to-new-heights-with-xds/index.html): Explore the current Cilium control-plane design, where and why limitations arise, and how to advance the architecture using the CNCF universal data plane (xDS) APIs. (6,881 words) ## Resources - [Full Page Index](/index.html): Browse all cached pages with rich metadata - [About This Cache](/content/about.html): Methodology, technical details, and usage guidelines - [XML Sitemap](/sitemap.xml): Machine-readable sitemap for crawler discovery - [Robots.txt](/robots.txt): Crawler directives