YepAPI

Command Palette

Search for a command to run...

GoogleText Generation/v1/ai/chat

Gemini 3.5 Flash Lite

Access Gemini 3.5 Flash Lite through one API key. Google Cheapest Gemini Subagent Model.

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities, suited for subagents that execute focused tasks inside complex multi-agent workflows. Flash Lite is built for the worker roles in a multi-agent system: fetch, extract, classify, summarise, and hand back to an orchestrator running a larger model.

No credit card required. Takes 30 seconds.

2,400+

Developers

1.2M+

API calls served

100+

Endpoints

$0.01

Per call

Yep, that's it.

Try it live

Send a message and see Gemini 3.5 Flash Lite respond in real time.

POST/v1/ai/chat

Maximum tokens in the response.

Real-time tokens

Hit "Send Request" to see the response

Context Window

1.0M tokens

Max Output

66K tokens

Input Price

$0.44 / 1M tokens

Output Price

$3.69 / 1M tokens

Strengths

✓
Cheapest current Gemini

$0.44 per 1M input and $3.69 per 1M output tokens.

✓
Built for subagents

Upgraded agentic capabilities aimed at focused tasks inside larger workflows.

✓
Full multimodal input

Accepts text, images, files, audio and video.

✓
1M token context

Handles up to 1,048,576 input tokens and returns up to 65,536 output tokens per call.

Quick start

Copy this snippet and start making calls with Gemini 3.5 Flash Lite.

const res = await fetch('https://api.yepapi.com/v1/ai/chat', {
  method: 'POST',
  headers: {
    'x-api-key': 'YOUR_API_KEY',
    'Content-Type': 'application/json',
  },
  body: JSON.stringify({
    "model": "google/gemini-3.5-flash-lite",
    "messages": [
      {
        "role": "user",
        "content": "Explain API gateways in 2 sentences."
      }
    ],
    "maxTokens": 256
  }),
});
const { data } = await res.json();
console.log(data.message.content);

Why use Gemini 3.5 Flash Lite through YepAPI?

✓One API key for all models — no separate accounts
✓OpenAI SDK compatible — just change the base URL
✓No monthly minimums — pay per token
✓Switch models with one line of code
✓Full provider passthrough — citations, search results, and all extras included
✓Streaming and non-streaming support on every model
✓Works with Cursor, Claude, LangChain, and any LLM tool
✓Unified billing across all providers

Gemini 3.5 Flash Lite API — pricing, context window & access

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities, suited for subagents that execute focused tasks inside complex multi-agent workflows. It pairs a 1,048,576-token context window with up to 65,536 output tokens per response, at $0.44 per 1M input and $3.69 per 1M output tokens through YepAPI.

Through YepAPI you call Gemini 3.5 Flash Lite on an OpenAI-compatible endpoint with a single key — the same key that reaches GPT-6, Claude, Gemini, Grok and the SEO, SERP and scraping APIs. Input: text, images, files, audio, video.

What is Gemini 3.5 Flash Lite?

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities, suited for subagents that execute focused tasks inside complex multi-agent workflows. At $0.44 / $3.69 per 1M tokens it is the cheapest current Gemini, and it keeps the family's full multimodal input and 1,048,576-token context window. It accepts text, images, files, audio, video as input, handles up to 1,048,576 tokens of context and returns up to 65,536 output tokens per call.

Build with Gemini 3.5 Flash Lite via YepAPI

YepAPI serves Gemini 3.5 Flash Lite on the OpenAI-compatible /v1/ai/chat endpoint. Point your base URL at YepAPI, add your key, and set the model string to gemini-3.5-flash-lite (or the full google/gemini-3.5-flash-lite). Function calling, structured outputs and reasoning settings pass straight through, and switching to any other model on the platform is a one-string change per request.

Gemini 3.5 Flash Lite API pricing — $0.44 / 1M input, $3.69 / 1M output

Gemini 3.5 Flash Lite costs $0.44 per 1M input tokens and $3.69 per 1M output tokens through YepAPI, pay per token with no minimums. Failed calls are never charged, and every request is itemised in your dashboard API Logs so you can see exactly what a 65,536-token response cost.

Gemini 3.5 Flash Lite use cases

Flash Lite is built for the worker roles in a multi-agent system: fetch, extract, classify, summarise, and hand back to an orchestrator running a larger model.

Try Gemini 3.5 Flash Lite free

Every new YepAPI account includes $5 in free credit with no card required — enough to test Gemini 3.5 Flash Lite on your own prompts before committing to paid usage. Create an account, copy your key, and call gemini-3.5-flash-lite right away.

Start generating in 30 seconds

$5 free credit on signup. No credit card required. Pay per call.

What developers say

“Switched from SerpAPI and cut our SERP costs by 80%. Same data quality, way simpler billing.”

Marcus T.

SEO Platform Founder

“One API key for AI models, SERP data, and web scraping. Saved us from managing 4 separate providers.”

Priya S.

Full-Stack Developer

“The $5 free credit let us prototype our entire rank tracking feature before committing. No other API does that.”

Jake R.

Indie Hacker

Frequently asked questions

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities, suited for subagents that execute focused tasks inside complex multi-agent workflows. Flash Lite is built for the worker roles in a multi-agent system: fetch, extract, classify, summarise, and hand back to an orchestrator running a larger model.

Input tokens cost $0.44 per 1M tokens and output tokens cost $3.69 per 1M tokens through YepAPI. No monthly minimums — you only pay for what you use.

Sign up for a free API key, then send requests to the /v1/ai/chat endpoint.

Gemini 3.5 Flash Lite supports a 1.0M token context window with up to 66K output tokens per request.

Ready to use Gemini 3.5 Flash Lite?

$5 free credit on signup. No credit card required. Pay per call.

Explore more models