AI API
A free AI API and LLM gateway for GPT-5.4, Claude, Gemini, DeepSeek, Llama, and Qwen. One API key — and one Claude API key — through a single OpenAI-compatible endpoint. $5 free credit, pay-per-token pricing, no subscriptions.
No credit card required. Takes 30 seconds.
2,400+
Developers
1.2M+
API calls served
100+
Endpoints
$0.01
Per call
Yep, that's it.
Why use a unified AI API?
One API key, 130+ AI models
YepAPI is an AI models API for text, image, and video generation. Access GPT-5.4, Claude, Gemini, DeepSeek, Llama, Qwen — all through one LLM API key. No juggling providers.
OpenAI compatible API endpoint
Our OpenAI compatible API endpoint means zero code changes. Same request format, same response shape. Drop in YepAPI's base URL and switch models by changing one parameter.
AI gateway & LLM router
YepAPI works as an AI gateway, routing requests to 130+ models through a single endpoint. An LLM API router that handles authentication, rate limits, and failover across all providers.
Cheapest AI API pricing
Straightforward pay-per-token pricing for production workloads — no subscriptions and no monthly minimums. Access GPT-4o, DeepSeek, and Claude API pricing through one key, billed per token.
One API call. Any model.
OpenAI-compatible endpoint — same format you already use.
const res = await fetch("https://api.yepapi.com/v1/ai/chat/completions", {
method: "POST",
headers: {
"Authorization": "Bearer yep_sk_your_key",
"Content-Type": "application/json",
},
body: JSON.stringify({
model: "deepseek/deepseek-chat-v3-0324",
messages: [{ role: "user", content: "Explain quantum computing" }],
}),
});
const data = await res.json();
console.log(data.choices[0].message.content);from openai import OpenAI
client = OpenAI(
base_url="https://api.yepapi.com/v1/ai",
api_key="yep_sk_your_key",
)
response = client.chat.completions.create(
model="deepseek/deepseek-chat-v3-0324",
messages=[{"role": "user", "content": "Explain quantum computing"}],
)
print(response.choices[0].message.content)Switch models by changing one parameter. Works with the OpenAI SDK, LangChain, Vercel AI SDK, or raw HTTP.
Text Generation
GPT-4o Mini
OpenAIAccess GPT-4o Mini through one API key. Fast, cheap, and OpenAI-compatible.
GPT-4o
OpenAIAccess GPT-4o through one API key. Flagship reasoning and multimodal capabilities.
Claude Sonnet 4
AnthropicAccess Claude Sonnet 4 through one API key. Anthropic's best balance of speed and intelligence.
Gemini 2.5 Flash
GoogleAccess Gemini 2.5 Flash through one API key. Google's fastest model with 1M context.
Gemini 2.5 Pro
GoogleAccess Gemini 2.5 Pro through one API key. Google's most capable model with 1M context.
Llama 4 Scout
MetaAccess Llama 4 Scout through one API key. Meta's open-weight model with 512K context.
DeepSeek R1
DeepSeekAccess DeepSeek R1 through one API key. Reasoning-specialized model for complex problems.
GPT-6 Astra
OpenAIAccess GPT-6 Astra through one API key. OpenAI's GPT-6 flagship for analysis, engineering and research.
GPT-5.5
OpenAIAccess GPT-5.5 through one API key. OpenAI's newest and most capable model.
GPT-5.5 Pro
OpenAIAccess GPT-5.5 Pro through one API key. OpenAI's most powerful reasoning model.
GPT-5.4
OpenAIAccess GPT-5.4 through one API key. OpenAI's latest and most capable model.
GPT-5.4 Mini
OpenAIAccess GPT-5.4 Mini through one API key. Smart and affordable with 400K context.
GPT-5.4 Nano
OpenAIAccess GPT-5.4 Nano through one API key. Ultra-cheap for high-volume tasks.
Claude Opus 4.8
AnthropicAccess Claude Opus 4.8 through one API key. Anthropic's newest and most intelligent model.
Claude Fable 5
AnthropicAccess Claude Fable 5 through one API key. Anthropic's creative writing flagship.
Claude Opus 4.7
AnthropicAccess Claude Opus 4.7 through one API key. Anthropic's newest and most intelligent model.
Claude Opus 4.5
AnthropicAccess Claude Opus 4.5 through one API key. OpenAI-compatible, pay per token.
Claude Opus 4.1
AnthropicAccess Claude Opus 4.1 through one API key. OpenAI-compatible, pay per token.
Claude Opus 4.6
AnthropicAccess Claude Opus 4.6 through one API key. Anthropic's previous-generation flagship.
Claude Sonnet 4.6
AnthropicAccess Claude Sonnet 4.6 through one API key. Anthropic's latest balanced model with 1M context.
Grok 4.3
xAIAccess Grok 4.3 through one API key. xAI's latest model with 1M context.
Grok 4.20
xAIAccess Grok 4.20 through one API key. xAI's flagship model with the largest context window.
Gemini 3.5 Flash
GoogleAccess Gemini 3.5 Flash through one API key. Google's latest fast multimodal model.
Gemini 3.1 Pro
GoogleAccess Gemini 3.1 Pro through one API key. Google's next-generation model.
Qwen3.7 Max
QwenAccess Qwen3.7 Max through one API key. Alibaba's most capable Qwen model.
Qwen 3.6 Plus
QwenAccess Qwen 3.6 Plus through one API key. Powerful reasoning with 1M context at a great price.
Mistral Medium 3.5
MistralAccess Mistral Medium 3.5 through one API key. Mistral's balanced mid-tier model.
Mistral Small 4
MistralAccess Mistral Small 4 through one API key. Europe's leading AI model at a great price.
Devstral 2
MistralAccess Devstral 2 through one API key. Mistral's coding-specialized model for developers.
GPT-5.4 Pro
OpenAIAccess GPT-5.4 Pro through one API key. OpenAI's most powerful reasoning model.
Sonar Pro
PerplexityAccess Sonar Pro through one API key. Search-augmented AI with citations built in.
Sonar
PerplexityAccess Sonar through one API key. Affordable search-augmented AI with citations.
Qwen 3 Coder
QwenAccess Qwen 3 Coder through one API key. Coding-specialized model at an ultra-low price.
Gemini 3 Flash
GoogleAccess Gemini 3 Flash through one API key. Google's next-generation fast model.
Grok 4.20 Multi-Agent
xAIAccess Grok 4.20 Multi-Agent through one API key. Multi-agent orchestration with 2M context.
GPT-5.3 Codex
OpenAIAccess GPT-5.3 Codex through one API key. Previous-gen coding model with massive output.
GPT-5.2 Pro
OpenAIAccess GPT-5.2 Pro through one API key. Previous-gen premium reasoning model.
GPT-5.2
OpenAIAccess GPT-5.2 through one API key. Previous-generation flagship model.
GPT-5.2 Chat
OpenAIAccess GPT-5.2 Chat through one API key. Reliable chat model.
GPT-5.2 Codex
OpenAIAccess GPT-5.2 Codex through one API key. Reliable coding model with massive context.
GPT-5.1 Codex Max
OpenAIAccess GPT-5.1 Codex Max through one API key. Max-output coding model.
GPT Audio
OpenAIAccess GPT Audio through one API key. Audio-capable multimodal model.
GPT Audio Mini
OpenAIAccess GPT Audio Mini through one API key. Affordable audio-capable model.
Gemini 3.1 Flash Lite
GoogleAccess Gemini 3.1 Flash Lite through one API key. Ultra-affordable with 1M context.
Fugu Max
Sakana AIAccess Sakana AI's Fugu Max through one API key. Multi-agent orchestration at a workhorse price.
Fugu Ultra v2
Sakana AIAccess Sakana AI's Fugu Ultra v2 through one API key. The top tier of the Fugu multi-agent system.
GPT-6 Sol
OpenAIAccess GPT-6 Sol through one API key. OpenAI GPT-6 Cost-Efficient High-End.
GPT-6 Luna
OpenAIAccess GPT-6 Luna through one API key. OpenAI GPT-6 Fast & Low-Cost.
GPT-6 Astra Pro
OpenAIAccess GPT-6 Astra Pro through one API key. OpenAI GPT-6 Astra Pro Reasoning Mode.
GPT-6 Sol Pro
OpenAIAccess GPT-6 Sol Pro through one API key. OpenAI GPT-6 Sol Pro Reasoning Mode.
GPT-6 Luna Pro
OpenAIAccess GPT-6 Luna Pro through one API key. OpenAI GPT-6 Luna Pro Reasoning Mode.
GPT-5.6 Luna
OpenAIAccess GPT-5.6 Luna through one API key. OpenAI GPT-5.6 Fast & Low-Cost.
GPT-5.6 Luna Pro
OpenAIAccess GPT-5.6 Luna Pro through one API key. OpenAI GPT-5.6 Luna Pro Reasoning Mode.
GPT-5.6 Sol
OpenAIAccess GPT-5.6 Sol through one API key. OpenAI GPT-5.6 Flagship for Coding & Agents.
GPT-5.6 Sol Pro
OpenAIAccess GPT-5.6 Sol Pro through one API key. OpenAI GPT-5.6 Sol Pro Reasoning Mode.
GPT-5.6 Terra
OpenAIAccess GPT-5.6 Terra through one API key. OpenAI GPT-5.6 Balanced Tier.
GPT-5.6 Terra Pro
OpenAIAccess GPT-5.6 Terra Pro through one API key. OpenAI GPT-5.6 Terra Pro Reasoning Mode.
Claude Fable 5.1
AnthropicAccess Claude Fable 5.1 through one API key. Anthropic Frontier Agentic Coding & Knowledge Work.
Claude Opus 5.5
AnthropicAccess Claude Opus 5.5 through one API key. Anthropic Flagship for Reasoning & Coding.
Claude Opus 5
AnthropicAccess Claude Opus 5 through one API key. Anthropic Opus 5 Flagship.
Claude Sonnet 5
AnthropicAccess Claude Sonnet 5 through one API key. Anthropic Sonnet 5 Frontier Mid-Tier.
Claude Haiku 4.5
AnthropicAccess Claude Haiku 4.5 through one API key. Anthropic Fastest Claude.
Gemini 3.8 Flash
GoogleAccess Gemini 3.8 Flash through one API key. Google Most Intelligent Flash.
Gemini 3.7 Flash
GoogleAccess Gemini 3.7 Flash through one API key. Google Fast Agentic Multimodal.
Gemini 3.6 Flash
GoogleAccess Gemini 3.6 Flash through one API key. Google High-Efficiency Coding Flash.
Gemini 3.5 Flash Lite
GoogleAccess Gemini 3.5 Flash Lite through one API key. Google Cheapest Gemini Subagent Model.
Grok 4.7
xAIAccess Grok 4.7 through one API key. xAI Flagship for Coding & Agents.
Grok 4.6
xAIAccess Grok 4.6 through one API key. xAI Grok 4.6 Frontier Coding & STEM.
Muse Spark 1.3
MetaAccess Muse Spark 1.3 through one API key. Meta Multimodal Agentic Reasoning.
DeepSeek V4.1 Flash
DeepSeekAccess DeepSeek V4.1 Flash through one API key. DeepSeek Ultra-Cheap Sparse MoE.
DeepSeek V4 Pro 0813
DeepSeekAccess DeepSeek V4 Pro 0813 through one API key. DeepSeek V4 Pro GA Release.
Qwen3.8 Max Prime
QwenAccess Qwen3.8 Max Prime through one API key. Qwen3.8 Max High-Throughput.
Qwen3.8 Max (0902)
QwenAccess Qwen3.8 Max (0902) through one API key. Qwen3.8 Max Updated Snapshot.
Qwen3.8 Flash
QwenAccess Qwen3.8 Flash through one API key. Qwen3.8 Cheap Multimodal Reasoning.
Qwen3.8 Omni Flash
QwenAccess Qwen3.8 Omni Flash through one API key. Qwen Omni-Modal Audio & Video Agent.
Qwen3.8 27B
QwenAccess Qwen3.8 27B through one API key. Qwen Open-Weight Vision-Language 27B.
GLM 5.3 Prime
Z.aiAccess GLM 5.3 Prime through one API key. Z.ai GLM 5.3 High-Speed Flagship.
GLM 5.3
Z.aiAccess GLM 5.3 through one API key. Z.ai GLM 5.3 Reasoning for Engineering.
GLM 5.3 Flash
Z.aiAccess GLM 5.3 Flash through one API key. Z.ai Cheapest Multimodal GLM.
GLM 5.3 FlashX
Z.aiAccess GLM 5.3 FlashX through one API key. Z.ai GLM 5.3 Flash at 200 tok/s.
MiMo-V2.6-Pro
XiaomiAccess MiMo-V2.6-Pro through one API key. Xiaomi 1T-Parameter Flagship.
MiMo-V2.6-Flash
XiaomiAccess MiMo-V2.6-Flash through one API key. Xiaomi Open-Source Omni-Modal MoE.
MiMo-V2.6-Pro-UltraSpeed
XiaomiAccess MiMo-V2.6-Pro-UltraSpeed through one API key. Xiaomi Flagship at 10× Speed.
Seed 2.1 Turbo
ByteDanceAccess Seed 2.1 Turbo through one API key. ByteDance Multimodal Coding Agent.
Seed-2.0-Code
ByteDanceAccess Seed-2.0-Code through one API key. ByteDance Agentic Coding Model.
Command A+
CohereAccess Command A+ through one API key. Cohere Enterprise Agentic Flagship.
Ember-1
FireworksAccess Ember-1 through one API key. Fireworks Efficient Reasoning on Kimi K3.
Solar Pro 4
UpstageAccess Solar Pro 4 through one API key. Upstage Cost-Efficient Long-Horizon LLM.
Solar Mini 4
UpstageAccess Solar Mini 4 through one API key. Upstage Compact 35B MoE.
Nemotron 3.5 Lightning
NVIDIAAccess Nemotron 3.5 Lightning through one API key. NVIDIA Open High-Throughput MoE.
Hy4 preview
TencentAccess Hy4 preview through one API key. Tencent Hunyuan 4 Preview Coding Agent MoE.
Gemma 4 31B
GoogleAccess Gemma 4 31B through one API key. Google's open-weight model at ultra-low cost.
Gemma 4 26B
GoogleAccess Gemma 4 26B through one API key. Efficient MoE model at the lowest price.
Qwen 3.5 Plus
QwenAccess Qwen 3.5 Plus through one API key. Strong model with 1M context.
Qwen 3.5 397B
QwenAccess Qwen 3.5 397B through one API key. Alibaba's largest MoE model.
Qwen 3.5 122B
QwenAccess Qwen 3.5 122B through one API key. Efficient MoE model.
Qwen 3.5 35B
QwenAccess Qwen 3.5 35B through one API key. Ultra-fast MoE model.
Qwen 3.5 27B
QwenAccess Qwen 3.5 27B through one API key. Dense model for consistent performance.
Qwen 3.5 9B
QwenAccess Qwen 3.5 9B through one API key. Cheapest model for high-volume tasks.
Qwen 3.5 Flash
QwenAccess Qwen 3.5 Flash through one API key. Ultra-fast with 1M context.
Qwen 3 Max Thinking
QwenAccess Qwen 3 Max Thinking through one API key. Reasoning model for complex problems.
Llama 4 Maverick
MetaAccess Llama 4 Maverick through one API key. Meta's most capable open-weight model.
MiniMax M3
MiniMaxAccess MiniMax M3 through one API key. MiniMax's latest flagship with 1M context.
MiniMax M2.7
MiniMaxAccess MiniMax M2.7 through one API key. Bilingual flagship model with massive output.
MiniMax M2.5
MiniMaxAccess MiniMax M2.5 through one API key. Affordable bilingual model.
GLM 5.2
Z.aiAccess GLM 5.2 through one API key. Z.ai's latest flagship with 1M context.
GLM 5.1
Z.aiAccess GLM 5.1 through one API key. Zhipu AI's latest bilingual model.
GLM 5
Z.aiAccess GLM 5 through one API key. Balanced bilingual model from Zhipu AI.
GLM 5 Turbo
Z.aiAccess GLM 5 Turbo through one API key. Fast bilingual model from Zhipu AI.
Seed 2.0 Lite
ByteDanceAccess Seed 2.0 Lite through one API key. ByteDance's latest bilingual model.
Seed 2.0 Mini
ByteDanceAccess Seed 2.0 Mini through one API key. Ultra-cheap from ByteDance.
Seed 1.6
ByteDanceAccess Seed 1.6 through one API key. Reliable model from ByteDance.
Seed 1.6 Flash
ByteDanceAccess Seed 1.6 Flash through one API key. Cheapest model from ByteDance.
DeepSeek V3.2
DeepSeekAccess DeepSeek V3.2 through one API key. Latest DeepSeek model.
Claude Sonnet 4.5
AnthropicAccess Claude Sonnet 4.5 through one API key. Previous-gen Anthropic flagship.
Gemini 2.5 Flash Lite
GoogleAccess Gemini 2.5 Flash Lite through one API key. Google's cheapest model.
GPT-OSS 120B
OpenAIAccess GPT-OSS 120B through one API key. OpenAI's open-source model.
Step 3.5 Flash
StepFunAccess Step 3.5 Flash through one API key. StepFun's fast model.
Kimi K2.7 Code
Moonshot AIAccess Kimi K2.7 Code through one API key. Moonshot AI's coding-specialized model.
Kimi K2.5
Moonshot AIAccess Kimi K2.5 through one API key. Moonshot AI's flagship model.
Nemotron 3 Super
NVIDIAAccess Nemotron 3 Super through one API key. NVIDIA's free model.
DeepSeek V4 Flash
DeepSeekAccess DeepSeek V4 Flash through one API key. 1M context, MoE speed, budget pricing.
DeepSeek V4 Pro
DeepSeekAccess DeepSeek V4 Pro through one API key. Frontier reasoning at a fraction of flagship pricing.
Image Generation
Nano Banana 3.1 Flash
GoogleFast AI image generation via Nano Banana 3.1 Flash. Generate and edit images up to 4K resolution.
Nano Banana 3 Pro
GoogleNano Banana API for AI image generation. Generate photorealistic images from text prompts — an AI image generator API and Midjourney alternative via simple REST API.
FLUX.2 Pro
Black Forest LabsGenerate photorealistic images with FLUX.2 Pro from $0.07 per image — one REST call, no GPU to rent.
FLUX.2 Max
Black Forest LabsFLUX.2 Max for hero-grade AI images from $0.15 per image — the family's highest-fidelity tier via one API key.
FLUX.2 Flex
Black Forest LabsFLUX.2 Flex from $0.13 per image — the tier between Pro and Max, on the same media endpoint.
FLUX.2 Klein
Black Forest LabsFLUX.2 Klein from $0.03 per image — the compact 4B FLUX model for high-volume generation.
Seedream 4.5
ByteDanceSeedream 4.5 image generation from $0.08 — 18 aspect ratios and reference-image editing through one API key.
GPT Image 2
OpenAIOpenAI's GPT Image 2 from $0.07 per image — best-in-class prompt following and in-image text, via one REST call.
GPT Image 1
OpenAIGPT Image 1 from $0.09 per image — OpenAI's proven image generator on the unified media endpoint.
GPT Image 1 Mini
OpenAIGPT Image 1 Mini from $0.02 per image — the budget OpenAI image tier for drafts and volume.
GPT-5.4 Image 2
OpenAIGPT-5.4 Image 2 from $0.07 per image — reasoning-led image generation on the unified media endpoint.
GPT-5 Image
OpenAIGPT-5 Image from $0.09 per image — detailed, prompt-faithful generation through one API key.
GPT-5 Image Mini
OpenAIGPT-5 Image Mini from $0.02 per image — the fast, cheap tier for drafts and volume.
Recraft V4.1
RecraftRecraft V4.1 from $0.07 per image — design-grade illustrations, icons, and graphics via one REST call.
Recraft V4.1 Pro
RecraftRecraft V4.1 Pro from $0.44 per image — client-ready illustration and brand graphics on the media endpoint.
Recraft V4.1 Vector
RecraftText-to-SVG with Recraft V4.1 Vector from $0.17 — real, editable vector art through one API call.
Recraft V4.1 Pro Vector
RecraftRecraft V4.1 Pro Vector from $0.63 — production-grade text-to-SVG for brand and print work.
Recraft V4.1 Utility
RecraftRecraft V4.1 Utility from $0.07 per image — the volume tier for drafts, placeholders, and bulk assets.
Recraft V4.1 Utility Pro
RecraftRecraft V4.1 Utility Pro from $0.44 per image — production-quality everyday design assets.
Recraft V4
RecraftRecraft V4 from $0.08 per image — the previous-generation Recraft design model, still available.
Recraft V4 Pro
RecraftRecraft V4 Pro from $0.53 per image — the professional tier of the previous Recraft generation.
Recraft V4 Vector
RecraftRecraft V4 Vector from $0.17 — editable text-to-SVG on the previous Recraft generation.
Recraft V4 Pro Vector
RecraftRecraft V4 Pro Vector from $0.63 — professional-grade text-to-SVG for brand and print work.
Recraft V3
RecraftRecraft V3 from $0.08 per image — the classic Recraft design model, still on the same endpoint.
Krea 2 Large
KreaKrea 2 Large from $0.10 per image — art-directed AI imagery through one API key.
Krea 2 Medium
KreaKrea 2 Medium from $0.10 per image — the balanced Krea tier for everyday creative work.
Krea 2 Medium Turbo
KreaKrea 2 Medium Turbo from $0.10 per image — the fastest Krea tier for interactive generation.
MAI Image 2.5 Pro
MicrosoftMicrosoft MAI Image 2.5 Pro from $0.29 per image — flagship-quality generation on the unified media endpoint.
MAI Image 2.5
MicrosoftMicrosoft MAI Image 2.5 from $0.13 per image — general-purpose generation through one API key.
Gemini 3.1 Flash Lite Image
GoogleGemini 3.1 Flash Lite Image from $0.08 — fast Google image generation with 14 aspect ratios.
Gemini 2.5 Flash Image
GoogleGemini 2.5 Flash Image — the original Nano Banana — from $0.08 per image via one REST call.
Riverflow 2.5 Pro
SourcefulRiverflow 2.5 Pro from $0.36 per image — Sourceful's flagship design and packaging image model.
Riverflow 2.5 Fast
SourcefulRiverflow 2.5 Fast from $0.04 per image — the budget tier of Sourceful's design image model.
Riverflow 2 Pro
SourcefulRiverflow 2 Pro from $0.70 per image — the previous-generation Sourceful flagship, still available.
Riverflow 2 Fast
SourcefulRiverflow 2 Fast from $0.08 per image — the quick tier of the previous Riverflow generation.
Grok Imagine Image
xAIxAI's Grok Imagine from $0.15 per image — expressive generation with 14 aspect ratios.
Video Generation
Veo 3.1
GoogleAI video generation via Veo 3.1. Generate videos from text or images at up to 4K resolution.
Veo 3.1 Fast
GoogleAI video generation API powered by Google Veo 3.1. Generate videos from text prompts at up to 4K — a Sora, Runway, and Kling alternative via simple REST API.
Veo 3.1 Lite
GoogleBudget AI video generation via Veo 3.1 Lite. Fastest and cheapest option.
Seedance 2.0
ByteDanceAI video generation API powered by ByteDance's Seedance 2.0. Generate videos from text, images, or video references with built-in audio generation.
Seedance 2.0 Fast
ByteDanceFast AI video generation powered by ByteDance's Seedance 2.0. Lower cost, faster output with same flexibility.
Sora 2
OpenAIAI video generation API powered by OpenAI's Sora 2. Generate videos with synced audio from text or images.
Sora 2 Pro
OpenAIPremium AI video generation powered by OpenAI's Sora 2 Pro. Production-quality videos up to 1080p with synced audio.
Kling 3.0 Pro
KlingKling 3.0 Pro video generation from $0.24/sec — 3–15 second clips with audio and keyframe control.
Kling 3.0 Standard
KlingKling 3.0 Standard from $0.18/sec — the value tier of Kling's video family, same features as Pro.
Kling O1
KlingKling O1 from $0.24/sec — reasoning-led video generation for complex, multi-beat prompts.
Hailuo 3
MiniMaxHailuo 3 from $0.27/sec — MiniMax's 2K video model with synced audio and keyframe control.
Hailuo 2.3
MiniMaxHailuo 2.3 from $0.17/sec — MiniMax 1080p video generation with image-to-video support.
Runway Gen-4.5
RunwayRunway Gen-4.5 from $0.25/sec — cinematic AI video via REST, billed per second with no monthly plan.
Runway Aleph 2
RunwayRunway Aleph 2 from $0.59/sec — the premium Runway tier with eight aspect ratios.
Grok Imagine Video 1.5
xAIGrok Imagine Video 1.5 from $0.17/sec — 1–15 second clips at up to 1080p through one API key.
Grok Imagine Video
xAIGrok Imagine Video from $0.11/sec — the budget xAI video tier with seven aspect ratios.
Wan 2.7
AlibabaWan 2.7 from $0.21/sec — Alibaba's flagship video model with synced audio and keyframe control.
Wan 2.6
AlibabaWan 2.6 from $0.08/sec — the most affordable video generation on the platform, audio included.
HappyHorse 1.1
AlibabaHappyHorse 1.1 from $0.21/sec — 3–15 second clips at up to 1080p with seven aspect ratios.
HappyHorse 1.0
AlibabaHappyHorse 1.0 from $0.21/sec — the first HappyHorse release, still on the same endpoint.
Seedance 1.5 Pro
ByteDanceSeedance 1.5 Pro from $0.46/sec — premium ByteDance video with synced audio and keyframe control.
Music Generation
Video Analysis
Text to Speech
MAI-Voice-2
MicrosoftMAI-Voice-2 text to speech — expressive speech in 15 languages, billed on input length.
MAI-Voice-2 Flash
MicrosoftMAI-Voice-2 Flash text to speech — low-latency speech for voice agents, billed on input length.
Deepgram Aura-2
DeepgramDeepgram Aura-2 text to speech — 90 voices across seven languages, billed on input length.
Voxtral Mini TTS
MistralVoxtral Mini TTS text to speech — 30 emotion-tagged voices, billed on input length.
Grok Voice TTS 1.0
xAIGrok Voice TTS 1.0 text to speech — 20+ languages, automatic detection, billed on input length.
Fish Audio S2.1 Pro
Fish AudioFish Audio S2.1 Pro text to speech — stateless voice cloning, billed on input length.
Fish Audio S2 Pro
Fish AudioFish Audio S2 Pro text to speech — expressive multi-speaker narration, billed on input length.
Fish Audio S1
Fish AudioFish Audio S1 text to speech — inline emotional control, billed on input length.
Fish Audio S2.1 Pro Free
Fish AudioFish Audio S2.1 Pro Free text to speech — free tier for prototyping, billed on input length.
Qwen-Audio 3.0 TTS Flash
QwenQwen-Audio 3.0 TTS Flash text to speech — fast Chinese & English speech, billed on input length.
Qwen-Audio 3.0 TTS Plus
QwenQwen-Audio 3.0 TTS Plus text to speech — higher-fidelity Chinese & English speech, billed on input length.
MiniMax Speech 2.8 HD
MiniMaxMiniMax Speech 2.8 HD text to speech — premium high-fidelity speech, billed on input length.
MiniMax Speech 2.8 Turbo
MiniMaxMiniMax Speech 2.8 Turbo text to speech — fast, flexible-voice speech, billed on input length.
Gemini 3.1 Flash TTS
GoogleGemini 3.1 Flash TTS text to speech — 70+ languages, 200+ audio tags, billed on input length.
Kokoro 82M
hexgradKokoro 82M text to speech — the cheapest AI speech available, billed on input length.
Orpheus 3B
Canopy LabsOrpheus 3B text to speech — natural English prosody, cheap, billed on input length.
Sesame CSM 1B
SesameSesame CSM 1B text to speech — conversational speech for assistants, billed on input length.
Zonos v0.1 Hybrid
ZyphraZonos v0.1 Hybrid text to speech — English speech with accent control, billed on input length.
Zonos v0.1 Transformer
ZyphraZonos v0.1 Transformer text to speech — pure-transformer English speech, billed on input length.
The LLM API that gives you 130+ models through one endpoint
YepAPI's LLM API is a unified AI gateway and AI models API that gives you access to 130+ AI models through a single OpenAI compatible API endpoint. Use the GPT-4o API, DeepSeek API, Claude API, Gemini API, Llama API, and Qwen API — all through one LLM API endpoint. Send the same request format you use with OpenAI and switch models by changing one parameter.
Looking for a free AI API to start with? Get $5 free credit on signup — no credit card — then pay as you go, per token, with no subscriptions. The DeepSeek API starts at $0.20/M input tokens. The GPT-4o API costs $3.50/M input. Claude API pricing starts at $1.12/M for Haiku. No monthly minimums makes this the LLM API developers reach for when they want pay-as-you-go pricing with broad model coverage.
YepAPI works as an AI gateway and LLM router — one key, one bill, one dashboard. Building an AI agent, chatbot, or content tool? This LLM API works with any framework — LangChain, LlamaIndex, Vercel AI SDK, or raw HTTP calls. Plus, your same API key covers web scraping, SEO data, SERP results, and YouTube APIs. Compare OpenRouter pricing to YepAPI — same models, better value.
OpenRouter alternative — better pricing, more APIs
Compare OpenRouter pricing to YepAPI: same LLM API access, but YepAPI bundles web scraping, SEO, SERP, and YouTube APIs under one key. The cheapest AI API with no monthly minimums.
AI API — Frequently asked questions
An AI API gives you programmatic access to large language models (LLMs) like GPT-5.4, Claude, Gemini, and DeepSeek. You send a prompt via HTTP request and get back generated text, code, or structured data. YepAPI's AI API is a unified endpoint — one API key accesses 130+ models from OpenAI, Anthropic, Google, Meta, and more.
YepAPI gives you $5 free credit on signup — no credit card required. That's enough for thousands of API calls on cheaper models like DeepSeek, Llama, or Qwen. There's no free tier with hard limits — you get real credits to use on any model, then pay as you go.
Sign up for YepAPI and your single API key works as your Claude API key — the same key also accesses GPT-5.4, Gemini, DeepSeek, and 60+ other models. No Anthropic enterprise contract or separate Claude API key required. You get $5 free credit on signup (no credit card), then pay per token: Claude Haiku 4.5 from $1.12/M input, Claude Sonnet 4.6 at $4.20/M input. Use it with the OpenAI SDK by changing the base URL — your Claude API key works through the same OpenAI-compatible endpoint.
OpenAI API pricing through YepAPI is simple per-token billing. GPT-4o costs $3.50/M input and $14/M output; GPT-4o mini costs $0.21/M input and $0.84/M output; GPT-5.4 and GPT-5.5 are available per token. There are no monthly minimums or subscriptions — your $5 free credit covers thousands of OpenAI API calls, then you pay per token. The same key also accesses Claude, Gemini, and DeepSeek through one OpenAI-compatible endpoint.
An LLM gateway is a unified API layer that sits between your app and multiple AI model providers. Instead of integrating with OpenAI, Anthropic, Google, and Meta separately, you connect to one gateway. YepAPI acts as an LLM gateway and router — one endpoint, one key, one billing dashboard. Switch between models by changing a single parameter.
Through YepAPI, DeepSeek V3.2 costs $0.36/M input tokens and $0.53/M output tokens. DeepSeek R1 (reasoning model) costs $0.77/M input and $3.07/M output. No monthly minimums — pay per token, starting from your $5 free credit.
An OpenAI-compatible API uses the same request and response format as OpenAI's API. Any code written for the OpenAI SDK works with YepAPI — just change the base URL and API key. This means you can access Claude, Gemini, DeepSeek, and Llama models using the same OpenAI client libraries you already use.
YepAPI serves a similar role to OpenRouter — both are LLM gateways that give you access to multiple AI models through one API. The differences: YepAPI bundles web scraping, SEO, SERP, and YouTube APIs under the same key, uses straightforward pay-per-token pricing, and gives you $5 free credit to start. If you want an OpenRouter alternative with broader API coverage, YepAPI is it.
Claude API pricing through YepAPI is per-token. Claude Sonnet 4.6 costs $4.20/M input and $21/M output. Claude Haiku 4.5 costs $1.12/M input and $5.60/M output. Claude Opus 4.8 costs $7/M input and $35/M output. No monthly minimums — pay only for the tokens you use. Access the Claude API without an Anthropic enterprise contract.
The GPT-4o API through YepAPI costs $3.50/M input tokens and $14/M output tokens. GPT-4o mini costs $0.21/M input and $0.84/M output. No monthly subscriptions or minimums. Your $5 free credit covers thousands of GPT-4o API calls. Switch between GPT-4o and any other model by changing one parameter.
YepAPI is an affordable LLM API for accessing multiple AI models. DeepSeek models start at $0.20/M tokens. Llama models start at $0.21/M tokens. Qwen models start at $0.07/M tokens. No monthly minimums, and $5 free credit on signup. Compare that to juggling OpenRouter or going direct to each provider — YepAPI gives you the broadest model coverage under one key.
OpenRouter pricing and YepAPI pricing are both per-token with no monthly minimums. The key differences: YepAPI bundles web scraping, SEO, SERP, and YouTube APIs under the same key — OpenRouter is LLM-only. YepAPI gives you $5 free credit on signup. Both offer the same models (GPT-4o API, DeepSeek API, Claude API, Gemini), but YepAPI is a more complete AI gateway for developers building data-driven apps.
DeepSeek API pricing through YepAPI: DeepSeek V3.2 costs $0.14/M input and $0.28/M output tokens. DeepSeek R1 (reasoning model) costs $0.55/M input and $2.19/M output. The DeepSeek API is one of the cheapest LLM APIs available — $5 free credit covers tens of thousands of calls. No monthly minimums, no subscriptions.
Ready to build with AI?
All 203 models included with one API key. $5 free credit on signup. No credit card required.