New

Announcing AISIX: The AI-Native AI Gateway for LLMs and AI AgentsLearn More

Learn More

All posts tagged

"Performance"

TLS Key Exchange for API Gateway Origins: Cut Handshake Retries Safely

Technology

September 15, 2026

TLS Key Exchange for API Gateway Origins: Cut Handshake Retries Safely

Learn how TLS 1.3 key-share selection affects API gateway origin latency, post-quantum readiness, compatibility, and rollout safety.

Speculative Decoding Changes AI Gateway Capacity Planning

Technology

September 8, 2026

Speculative Decoding Changes AI Gateway Capacity Planning

Learn how speculative decoding changes LLM throughput, latency, GPU memory, routing, and the metrics an AI Gateway should expose.

API Gateway Cache Memory: Lessons from Cloudflare's 100 TB DNS Optimization

Technology

September 1, 2026

API Gateway Cache Memory: Lessons from Cloudflare's 100 TB DNS Optimization

Learn how cache-key cardinality, entry layout, TTLs, response limits, and realistic benchmarks keep API Gateway caches fast without wasting memory.

Building Web Servers From Scratch vs. Using API Gateways: A Performance Analysis

Technology

May 12, 2026

Building Web Servers From Scratch vs. Using API Gateways: A Performance Analysis

Explore why building custom web servers in assembly isn't scalable. Learn how API Gateways deliver performance at scale.

How Is Apache APISIX Fast?

Technology

June 12, 2023

How Is Apache APISIX Fast?

Taking a look under Apache APISIX's hood to understand how it achieves ultimate performance.