YepAPI

Command Palette

Search for a command to run...

AnthropicText Generation/v1/ai/chat

Claude Haiku 4.5

Access Claude Haiku 4.5 through one API key. Anthropic Fastest Claude.

Claude Haiku 4.5 is Anthropic's fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Haiku 4.5 fits latency-sensitive chat, high-volume extraction and classification, and the fast sub-agents inside larger Claude-based systems.

No credit card required. Takes 30 seconds.

2,400+

Developers

1.2M+

API calls served

100+

Endpoints

$0.01

Per call

Yep, that's it.

Try it live

Send a message and see Claude Haiku 4.5 respond in real time.

POST/v1/ai/chat

Maximum tokens in the response.

Real-time tokens

Hit "Send Request" to see the response

Context Window

200K tokens

Max Output

64K tokens

Input Price

$1.48 / 1M tokens

Output Price

$7.38 / 1M tokens

Strengths

✓
Sonnet 4 quality, Haiku speed

Matches Claude Sonnet 4 on coding and reasoning at a fraction of the latency and cost.

✓
Fast sub-agent

A natural choice for the worker agents in a multi-agent system where a larger model orchestrates.

✓
200K token context, 64K output

Handles up to 200,000 input tokens and returns up to 64,000 output tokens per call.

✓
Tools, vision and structured output

Function calling, image and file input, and JSON schema outputs.

Quick start

Copy this snippet and start making calls with Claude Haiku 4.5.

const res = await fetch('https://api.yepapi.com/v1/ai/chat', {
  method: 'POST',
  headers: {
    'x-api-key': 'YOUR_API_KEY',
    'Content-Type': 'application/json',
  },
  body: JSON.stringify({
    "model": "anthropic/claude-haiku-4.5",
    "messages": [
      {
        "role": "user",
        "content": "Explain API gateways in 2 sentences."
      }
    ],
    "maxTokens": 256
  }),
});
const { data } = await res.json();
console.log(data.message.content);

Why use Claude Haiku 4.5 through YepAPI?

✓One API key for all models — no separate accounts
✓OpenAI SDK compatible — just change the base URL
✓No monthly minimums — pay per token
✓Switch models with one line of code
✓Full provider passthrough — citations, search results, and all extras included
✓Streaming and non-streaming support on every model
✓Works with Cursor, Claude, LangChain, and any LLM tool
✓Unified billing across all providers

Claude Haiku 4.5 API — pricing, context window & access

Claude Haiku 4.5 is Anthropic's fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. It pairs a 200,000-token context window with up to 64,000 output tokens per response, at $1.48 per 1M input and $7.38 per 1M output tokens through YepAPI.

Through YepAPI you call Claude Haiku 4.5 on an OpenAI-compatible endpoint with a single key — the same key that reaches GPT-6, Claude, Gemini, Grok and the SEO, SERP and scraping APIs. Input: text, images, files.

What is Claude Haiku 4.5?

Claude Haiku 4.5 is Anthropic's fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. It matches Claude Sonnet 4's performance on coding and reasoning while running far faster, with a 200,000-token context window, 64,000 output tokens and $1.48 / $7.38 per 1M token pricing. It accepts text, images, files as input, handles up to 200,000 tokens of context and returns up to 64,000 output tokens per call.

Build with Claude Haiku 4.5 via YepAPI

YepAPI serves Claude Haiku 4.5 on the OpenAI-compatible /v1/ai/chat endpoint. Point your base URL at YepAPI, add your key, and set the model string to claude-haiku-4.5 (or the full anthropic/claude-haiku-4.5). Function calling, structured outputs and reasoning settings pass straight through, and switching to any other model on the platform is a one-string change per request.

Claude Haiku 4.5 API pricing — $1.48 / 1M input, $7.38 / 1M output

Claude Haiku 4.5 costs $1.48 per 1M input tokens and $7.38 per 1M output tokens through YepAPI, pay per token with no minimums. Failed calls are never charged, and every request is itemised in your dashboard API Logs so you can see exactly what a 64,000-token response cost.

Claude Haiku 4.5 use cases

Haiku 4.5 fits latency-sensitive chat, high-volume extraction and classification, and the fast sub-agents inside larger Claude-based systems.

Try Claude Haiku 4.5 free

Every new YepAPI account includes $5 in free credit with no card required — enough to test Claude Haiku 4.5 on your own prompts before committing to paid usage. Create an account, copy your key, and call claude-haiku-4.5 right away.

Start generating in 30 seconds

$5 free credit on signup. No credit card required. Pay per call.

What developers say

“Switched from SerpAPI and cut our SERP costs by 80%. Same data quality, way simpler billing.”

Marcus T.

SEO Platform Founder

“One API key for AI models, SERP data, and web scraping. Saved us from managing 4 separate providers.”

Priya S.

Full-Stack Developer

“The $5 free credit let us prototype our entire rank tracking feature before committing. No other API does that.”

Jake R.

Indie Hacker

Frequently asked questions

Claude Haiku 4.5 is Anthropic's fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Haiku 4.5 fits latency-sensitive chat, high-volume extraction and classification, and the fast sub-agents inside larger Claude-based systems.

Input tokens cost $1.48 per 1M tokens and output tokens cost $7.38 per 1M tokens through YepAPI. No monthly minimums — you only pay for what you use.

Sign up for a free API key, then send requests to the /v1/ai/chat endpoint.

Claude Haiku 4.5 supports a 200K token context window with up to 64K output tokens per request.

Ready to use Claude Haiku 4.5?

$5 free credit on signup. No credit card required. Pay per call.

Explore more models