GPT-5.6 Luna
Access GPT-5.6 Luna through one API key. OpenAI GPT-5.6 Fast & Low-Cost.
GPT-5.6 Luna is the fast, cost-efficient model in OpenAI's GPT-5.6 series. Luna 5.6 fits routing, extraction, summarisation and the small repeated steps inside agents, where cost per call matters more than peak intelligence.
No credit card required. Takes 30 seconds.
2,400+
Developers
1.2M+
API calls served
100+
Endpoints
$0.01
Per call
Yep, that's it.
Try it live
Send a message and see GPT-5.6 Luna respond in real time.
Maximum tokens in the response.
Real-time tokens
Context Window
1.1M tokens
Max Output
128K tokens
Input Price
$0.30 / 1M tokens
Output Price
$1.77 / 1M tokens
Strengths
$0.30 per 1M input and $1.77 per 1M output tokens, for high-volume pipelines.
Tuned for fast responses in interactive and batch workloads.
Keeps the 1,050,000-token window despite the price.
Function calling and JSON schema outputs for agent sub-steps.
Quick start
Copy this snippet and start making calls with GPT-5.6 Luna.
const res = await fetch('https://api.yepapi.com/v1/ai/chat', {
method: 'POST',
headers: {
'x-api-key': 'YOUR_API_KEY',
'Content-Type': 'application/json',
},
body: JSON.stringify({
"model": "openai/gpt-5.6-luna",
"messages": [
{
"role": "user",
"content": "Explain API gateways in 2 sentences."
}
],
"maxTokens": 256
}),
});
const { data } = await res.json();
console.log(data.message.content);Why use GPT-5.6 Luna through YepAPI?
GPT-5.6 Luna API — pricing, context window & access
GPT-5.6 Luna is the fast, cost-efficient model in OpenAI's GPT-5.6 series. It pairs a 1,050,000-token context window with up to 128,000 output tokens per response, at $0.30 per 1M input and $1.77 per 1M output tokens through YepAPI.
Through YepAPI you call GPT-5.6 Luna on an OpenAI-compatible endpoint with a single key — the same key that reaches GPT-6, Claude, Gemini, Grok and the SEO, SERP and scraping APIs. Input: text, images, files.
What is GPT-5.6 Luna?
GPT-5.6 Luna is the fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification and lightweight agentic workflows, providing capable reasoning at $0.30 / $1.77 per 1M tokens with the same 1,050,000-token context window as the rest of the series. A GPT-5.6 Luna Pro variant runs it in a higher-effort reasoning mode. It accepts text, images, files as input, handles up to 1,050,000 tokens of context and returns up to 128,000 output tokens per call.
Build with GPT-5.6 Luna via YepAPI
YepAPI serves GPT-5.6 Luna on the OpenAI-compatible /v1/ai/chat endpoint. Point your base URL at YepAPI, add your key, and set the model string to gpt-5.6-luna (or the full openai/gpt-5.6-luna). Function calling, structured outputs and reasoning settings pass straight through, and switching to any other model on the platform is a one-string change per request.
GPT-5.6 Luna API pricing — $0.30 / 1M input, $1.77 / 1M output
GPT-5.6 Luna costs $0.30 per 1M input tokens and $1.77 per 1M output tokens through YepAPI, pay per token with no minimums. Failed calls are never charged, and every request is itemised in your dashboard API Logs so you can see exactly what a 128,000-token response cost.
GPT-5.6 Luna use cases
Luna 5.6 fits routing, extraction, summarisation and the small repeated steps inside agents, where cost per call matters more than peak intelligence.
Try GPT-5.6 Luna free
Every new YepAPI account includes $5 in free credit with no card required — enough to test GPT-5.6 Luna on your own prompts before committing to paid usage. Create an account, copy your key, and call gpt-5.6-luna right away.
Start generating in 30 seconds
$5 free credit on signup. No credit card required. Pay per call.
What developers say
“Switched from SerpAPI and cut our SERP costs by 80%. Same data quality, way simpler billing.”
“One API key for AI models, SERP data, and web scraping. Saved us from managing 4 separate providers.”
“The $5 free credit let us prototype our entire rank tracking feature before committing. No other API does that.”
Frequently asked questions
GPT-5.6 Luna is the fast, cost-efficient model in OpenAI's GPT-5.6 series. Luna 5.6 fits routing, extraction, summarisation and the small repeated steps inside agents, where cost per call matters more than peak intelligence.
Input tokens cost $0.30 per 1M tokens and output tokens cost $1.77 per 1M tokens through YepAPI. No monthly minimums — you only pay for what you use.
Sign up for a free API key, then send requests to the /v1/ai/chat endpoint.
GPT-5.6 Luna supports a 1.1M token context window with up to 128K output tokens per request.
Ready to use GPT-5.6 Luna?
$5 free credit on signup. No credit card required. Pay per call.
Explore more models
GPT-4o Mini
OpenAIAccess GPT-4o Mini through one API key. Fast, cheap, and OpenAI-compatible.
GPT-4o
OpenAIAccess GPT-4o through one API key. Flagship reasoning and multimodal capabilities.
GPT-6 Astra
OpenAIAccess GPT-6 Astra through one API key. OpenAI's GPT-6 flagship for analysis, engineering and research.