Qwen3.8 Omni Flash
Access Qwen3.8 Omni Flash through one API key. Qwen Omni-Modal Audio & Video Agent.
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio and video understanding. Omni Flash is the model for meeting recordings, call analysis, video summarisation and any agent that needs to hear and see as well as read, at a very low price.
No credit card required. Takes 30 seconds.
2,400+
Developers
1.2M+
API calls served
100+
Endpoints
$0.01
Per call
Yep, that's it.
Try it live
Send a message and see Qwen3.8 Omni Flash respond in real time.
Maximum tokens in the response.
Real-time tokens
Context Window
1M tokens
Max Output
131K tokens
Input Price
$0.22 / 1M tokens
Output Price
$0.69 / 1M tokens
Strengths
Accepts text, images, audio and video in one prompt.
The first Qwen model built around agent capabilities over rich media.
Handles up to 1,000,000 input tokens and returns up to 131,072 output tokens per call.
$0.22 per 1M input and $0.69 per 1M output tokens.
Quick start
Copy this snippet and start making calls with Qwen3.8 Omni Flash.
const res = await fetch('https://api.yepapi.com/v1/ai/chat', {
method: 'POST',
headers: {
'x-api-key': 'YOUR_API_KEY',
'Content-Type': 'application/json',
},
body: JSON.stringify({
"model": "qwen/qwen3.8-omni-flash",
"messages": [
{
"role": "user",
"content": "Explain API gateways in 2 sentences."
}
],
"maxTokens": 256
}),
});
const { data } = await res.json();
console.log(data.message.content);Why use Qwen3.8 Omni Flash through YepAPI?
Qwen3.8 Omni Flash API — pricing, context window & access
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio and video understanding. It pairs a 1,000,000-token context window with up to 131,072 output tokens per response, at $0.22 per 1M input and $0.69 per 1M output tokens through YepAPI.
Through YepAPI you call Qwen3.8 Omni Flash on an OpenAI-compatible endpoint with a single key — the same key that reaches GPT-6, Claude, Gemini, Grok and the SEO, SERP and scraping APIs. Input: text, images, audio, video.
What is Qwen3.8 Omni Flash?
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio and video understanding. It is suited for audio-video analysis and summarisation and for agents that operate over recordings, with a 1,000,000-token context window, up to 131,072 output tokens and $0.22 / $0.69 per 1M token pricing. It accepts text, images, audio, video as input, handles up to 1,000,000 tokens of context and returns up to 131,072 output tokens per call.
Build with Qwen3.8 Omni Flash via YepAPI
YepAPI serves Qwen3.8 Omni Flash on the OpenAI-compatible /v1/ai/chat endpoint. Point your base URL at YepAPI, add your key, and set the model string to qwen3.8-omni-flash (or the full qwen/qwen3.8-omni-flash). Function calling, structured outputs and reasoning settings pass straight through, and switching to any other model on the platform is a one-string change per request.
Qwen3.8 Omni Flash API pricing — $0.22 / 1M input, $0.69 / 1M output
Qwen3.8 Omni Flash costs $0.22 per 1M input tokens and $0.69 per 1M output tokens through YepAPI, pay per token with no minimums. Failed calls are never charged, and every request is itemised in your dashboard API Logs so you can see exactly what a 131,072-token response cost.
Qwen3.8 Omni Flash use cases
Omni Flash is the model for meeting recordings, call analysis, video summarisation and any agent that needs to hear and see as well as read, at a very low price.
Try Qwen3.8 Omni Flash free
Every new YepAPI account includes $5 in free credit with no card required — enough to test Qwen3.8 Omni Flash on your own prompts before committing to paid usage. Create an account, copy your key, and call qwen3.8-omni-flash right away.
Start generating in 30 seconds
$5 free credit on signup. No credit card required. Pay per call.
What developers say
“Switched from SerpAPI and cut our SERP costs by 80%. Same data quality, way simpler billing.”
“One API key for AI models, SERP data, and web scraping. Saved us from managing 4 separate providers.”
“The $5 free credit let us prototype our entire rank tracking feature before committing. No other API does that.”
Frequently asked questions
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio and video understanding. Omni Flash is the model for meeting recordings, call analysis, video summarisation and any agent that needs to hear and see as well as read, at a very low price.
Input tokens cost $0.22 per 1M tokens and output tokens cost $0.69 per 1M tokens through YepAPI. No monthly minimums — you only pay for what you use.
Sign up for a free API key, then send requests to the /v1/ai/chat endpoint.
Qwen3.8 Omni Flash supports a 1M token context window with up to 131K output tokens per request.
Ready to use Qwen3.8 Omni Flash?
$5 free credit on signup. No credit card required. Pay per call.
Explore more models
Qwen3.7 Max
QwenAccess Qwen3.7 Max through one API key. Alibaba's most capable Qwen model.
Qwen 3.6 Plus
QwenAccess Qwen 3.6 Plus through one API key. Powerful reasoning with 1M context at a great price.
Qwen 3 Coder
QwenAccess Qwen 3 Coder through one API key. Coding-specialized model at an ultra-low price.