Skip to main content

Overview

Kyma API is an LLM API gateway that routes one key to 101 models and publishes measured uptime per model. It gives you instant access to the best open source and frontier LLMs through a single endpoint. Compatible with both OpenAI and Anthropic SDKs. Multi-provider redundancy means your requests always go through — even when individual providers are down.

101 models, one endpoint

Qwen 3.6, DeepSeek V4, Gemma 4, GPT-OSS, Kimi K2.6, Gemini, Llama, MiniMax, GLM, Perplexity Sonar for live web search, plus GPT Image 2, FLUX, Ideogram, Recraft, MiniMax Image for image generation, ElevenLabs and MiniMax for voice.

Auto-Failover

Multi-provider redundancy. If one fails, your request is automatically retried on another.

OpenAI Compatible

Drop-in replacement. Works with any OpenAI SDK, LangChain, Cursor, and more.

Why Kyma?

The 0.50signupcreditspendsonthefreetier:/GENERATED:freetierthresholds/textmodelsatorunder0.50 signup credit spends on the free tier: {/* GENERATED:free-tier-thresholds */}text models at or under 3/M output and 0.50/Minput,plusmediaatorunder0.50/M input, plus media at or under 0.01 per billing unit. Image, video, music, premium voices and realtime sessions need a top-up first — see Free tier. Updated 2026-09-07.

How it works

Every model has a primary provider and an ordered list of backups. If the primary fails, Kyma automatically retries the next healthy one — you never see the error. No need to manage multiple API keys or monitor provider status.

Quick example

Ready to start?

Get your API key and make your first request in 30 seconds →