New

Announcing AISIX: The AI-Native AI Gateway for LLMs and AI AgentsLearn More

Learn More

All posts tagged

"AISIX"

AI Gateway High Availability: Survive Control-Plane Outages

Technology

September 29, 2026

AI Gateway High Availability: Survive Control-Plane Outages

Design an AI gateway that safely serves accepted policy through control-plane outages, restarts, stale state, shared-service failures, and recovery drills.

AI Model Gateway: Migrate Models Without Breaking Apps

Technology

September 29, 2026

AI Model Gateway: Migrate Models Without Breaking Apps

Use stable model aliases, contract tests, canary traffic, and rollback gates to migrate LLM providers without coupling applications to constant model churn.

AISIX 1.5.0: Run Non-Chat AI Workloads Through Model Groups

Products

September 29, 2026

AISIX 1.5.0: Run Non-Chat AI Workloads Through Model Groups

AISIX 1.5.0 extends Model Group routing, retry, failover, balancing, and per-attempt usage records to supported single-request AI endpoints.

AI Gateway for Dify, n8n, and Open WebUI

Technology

September 22, 2026

AI Gateway for Dify, n8n, and Open WebUI

Connect Dify, n8n, and Open WebUI to AISIX with stable model aliases, protected provider keys, policy, usage attribution, and workload tests.

AISIX 1.4.0: Redis Startup Resilience and Bounded Usage Retry

Products

September 22, 2026

AISIX 1.4.0: Redis Startup Resilience and Bounded Usage Retry

AISIX 1.4.0 can serve in degraded mode when Redis is unavailable at startup and can retry usage delivery within strict limits when deduplication evidence permits.

OpenAI-Compatible AI Gateway: Test the Contract

Technology

September 22, 2026

OpenAI-Compatible AI Gateway: Test the Contract

Test an OpenAI-compatible AI gateway across auth, model discovery, streaming, tool calls, errors, usage, and provider translation.

AISIX 1.3.0: Run Structured AI Agent Workflows Across Providers

Products

September 18, 2026

AISIX 1.3.0: Run Structured AI Agent Workflows Across Providers

AISIX 1.3.0 expands structured output and Responses translation, with explicit guardrail boundaries and clearer records of completed agent requests.

Private LLM Gateway: Secure vLLM, SGLang, and Ollama

Technology

September 15, 2026

Private LLM Gateway: Secure vLLM, SGLang, and Ollama

Learn how a private LLM gateway secures vLLM, SGLang, and Ollama with stable model aliases, traffic policy, routing, and observability.

What Is an MCP Registry? Discovery and Gateway Control

Technology

September 15, 2026

What Is an MCP Registry? Discovery and Gateway Control

Learn how an MCP Registry catalogs servers, differs from a marketplace and gateway, and supports governed MCP discovery and tool execution.

AI Gateway for Batch Inference and Fine-Tuning

Technology

September 8, 2026

AI Gateway for Batch Inference and Fine-Tuning

Learn how an AI Gateway governs Files, Batch, and Fine-tuning APIs with secure routing, credential isolation, job lifecycle control, and usage attribution.

Multimodal AI Gateway for Image, Audio, and Video

Technology

September 8, 2026

Multimodal AI Gateway for Image, Audio, and Video

Learn how a multimodal AI Gateway governs image, audio, video, and realtime APIs across identity, policy, provider routing, cost, and telemetry.

AISIX 1.0.0: Make AI Guardrails Measurable Before You Rely on Them

Products

September 7, 2026

AISIX 1.0.0: Make AI Guardrails Measurable Before You Rely on Them

AISIX 1.0.0 adds semantic guardrail testing and decision scores, with clearer failure policies and controls for production AI traffic.

AI Gateway Audit Logging: Build Evidence for AI Governance and Compliance

Technology

September 1, 2026

AI Gateway Audit Logging: Build Evidence for AI Governance and Compliance

Learn how AI Gateway audit logging connects requests, policy outcomes, configuration changes, retention, and exports for AI governance and investigations.

PII Redaction in an AI Gateway: Protect Sensitive Data Before It Reaches an LLM

Technology

September 1, 2026

PII Redaction in an AI Gateway: Protect Sensitive Data Before It Reaches an LLM

Learn how AI Gateway PII redaction detects, masks, or blocks sensitive data in LLM requests and responses before it crosses a trust boundary.

Rolling Out AI Gateway Guardrails: Monitor, Block, and Handle Streaming Safely

Technology

August 25, 2026

Rolling Out AI Gateway Guardrails: Monitor, Block, and Handle Streaming Safely

Use AISIX AI Gateway guardrails safely with monitor and block modes, input and output hooks, scoped policies, streaming controls, and failure handling.

Rotate LLM Provider Keys in AISIX Without Breaking Model Aliases

Technology

August 25, 2026

Rotate LLM Provider Keys in AISIX Without Breaking Model Aliases

Rotate LLM provider credentials in AISIX Cloud or open-source AISIX while keeping caller API keys and model aliases stable, tested, and recoverable.

Open Source AI Gateway: Envoy, AISIX, Kong, and Cloudflare Compared

Technology

July 1, 2026

Open Source AI Gateway: Envoy, AISIX, Kong, and Cloudflare Compared

Compare open source AI Gateway options and learn how to evaluate AI traffic control, security, cost governance, and enterprise readiness.

Model Ensembles: Several Low-Cost LLMs, One Better Answer

Products

June 17, 2026

Model Ensembles: Several Low-Cost LLMs, One Better Answer

AISIX model ensembles fan one chat request to a panel of LLMs, then a judge synthesizes one answer — so several low-cost models can rival a frontier model.

Announcing AISIX: The AI-Native AI Gateway

Products

March 31, 2026

Announcing AISIX: The AI-Native AI Gateway

Learn why API7 built AISIX from scratch as an AI-native gateway for LLM workloads, with unified model access, observability, and fine-grained traffic control.