Ember-1
Access Ember-1 through one API key. Fireworks Efficient Reasoning on Kimi K3.
Ember-1 is a specialised reasoning model from Fireworks Research, built on Moonshot's Kimi K3. Ember-1 is for reasoning-heavy workloads where the cost is dominated by long thinking traces: it delivers the same answers with markedly fewer tokens.
No credit card required. Takes 30 seconds.
2,400+
Developers
1.2M+
API calls served
100+
Endpoints
$0.01
Per call
Yep, that's it.
Try it live
Send a message and see Ember-1 respond in real time.
Maximum tokens in the response.
Real-time tokens
Context Window
1.0M tokens
Max Output
944K tokens
Input Price
$4.43 / 1M tokens
Output Price
$22.15 / 1M tokens
Strengths
Roughly 40% fewer reasoning tokens than its base for equivalent answers.
Inherits Moonshot's Kimi K3 capability with a Fireworks efficiency post-train.
Handles up to 1,048,576 input tokens and returns up to 943,718 output tokens per call.
Function calling and JSON schema outputs.
Quick start
Copy this snippet and start making calls with Ember-1.
const res = await fetch('https://api.yepapi.com/v1/ai/chat', {
method: 'POST',
headers: {
'x-api-key': 'YOUR_API_KEY',
'Content-Type': 'application/json',
},
body: JSON.stringify({
"model": "fireworks/ember-1",
"messages": [
{
"role": "user",
"content": "Explain API gateways in 2 sentences."
}
],
"maxTokens": 256
}),
});
const { data } = await res.json();
console.log(data.message.content);Why use Ember-1 through YepAPI?
Ember-1 API — pricing, context window & access
Ember-1 is a specialised reasoning model from Fireworks Research, built on Moonshot's Kimi K3. It pairs a 1,048,576-token context window with up to 943,718 output tokens per response, at $4.43 per 1M input and $22.15 per 1M output tokens through YepAPI.
Through YepAPI you call Ember-1 on an OpenAI-compatible endpoint with a single key — the same key that reaches GPT-6, Claude, Gemini, Grok and the SEO, SERP and scraping APIs. Input: text, images.
What is Ember-1?
Ember-1 is a specialised reasoning model from Fireworks Research, built on Moonshot's Kimi K3. It is designed to make every token go further: it produces shorter reasoning traces, using roughly 40% fewer tokens than the base model for the same answers, with text and image input, a 1,048,576-token context window and up to 943,718 output tokens at $4.43 / $22.15 per 1M tokens. It accepts text, images as input, handles up to 1,048,576 tokens of context and returns up to 943,718 output tokens per call.
Build with Ember-1 via YepAPI
YepAPI serves Ember-1 on the OpenAI-compatible /v1/ai/chat endpoint. Point your base URL at YepAPI, add your key, and set the model string to ember-1 (or the full fireworks/ember-1). Function calling, structured outputs and reasoning settings pass straight through, and switching to any other model on the platform is a one-string change per request.
Ember-1 API pricing — $4.43 / 1M input, $22.15 / 1M output
Ember-1 costs $4.43 per 1M input tokens and $22.15 per 1M output tokens through YepAPI, pay per token with no minimums. Failed calls are never charged, and every request is itemised in your dashboard API Logs so you can see exactly what a very long response cost.
Ember-1 use cases
Ember-1 is for reasoning-heavy workloads where the cost is dominated by long thinking traces: it delivers the same answers with markedly fewer tokens.
Try Ember-1 free
Every new YepAPI account includes $5 in free credit with no card required — enough to test Ember-1 on your own prompts before committing to paid usage. Create an account, copy your key, and call ember-1 right away.
Start generating in 30 seconds
$5 free credit on signup. No credit card required. Pay per call.
What developers say
“Switched from SerpAPI and cut our SERP costs by 80%. Same data quality, way simpler billing.”
“One API key for AI models, SERP data, and web scraping. Saved us from managing 4 separate providers.”
“The $5 free credit let us prototype our entire rank tracking feature before committing. No other API does that.”
Frequently asked questions
Ember-1 is a specialised reasoning model from Fireworks Research, built on Moonshot's Kimi K3. Ember-1 is for reasoning-heavy workloads where the cost is dominated by long thinking traces: it delivers the same answers with markedly fewer tokens.
Input tokens cost $4.43 per 1M tokens and output tokens cost $22.15 per 1M tokens through YepAPI. No monthly minimums — you only pay for what you use.
Sign up for a free API key, then send requests to the /v1/ai/chat endpoint.
Ember-1 supports a 1.0M token context window with up to 944K output tokens per request.
Ready to use Ember-1?
$5 free credit on signup. No credit card required. Pay per call.
Explore more models
GPT-4o Mini
OpenAIAccess GPT-4o Mini through one API key. Fast, cheap, and OpenAI-compatible.
GPT-4o
OpenAIAccess GPT-4o through one API key. Flagship reasoning and multimodal capabilities.
Claude Sonnet 4
AnthropicAccess Claude Sonnet 4 through one API key. Anthropic's best balance of speed and intelligence.