Connected / Updated 00:00:00

Compare AI model API pricing

Compare GPT, Claude, Gemini, DeepSeek, Grok, Qwen, and other AI API cost signals from the pricing endpoint before routing traffic through Route Key.

967
Models
18
Vendors
6
Endpoints
7
Groups
967 of 967 models
Page size

gpt-3.5-turbo

OpenAI
No model description is provided by the pricing endpoint.
$0.50
Input
$1.50
Output
Context -Group ratio 1x
Per tokenopenai

gpt-3.5-turbo-0125

OpenAI
No model description is provided by the pricing endpoint.
$0.50
Input
$1.50
Output
Context -Group ratio 1x
Per tokenopenai

gpt-3.5-turbo-0301

OpenAI
No model description is provided by the pricing endpoint.
$1.50
Input
$2.00
Output
Context -Group ratio 1x
Per tokenopenai

gpt-3.5-turbo-0613

OpenAI
No model description is provided by the pricing endpoint.
$1.50
Input
$2.00
Output
Context -Group ratio 1x
Per tokenopenai

gpt-3.5-turbo-1106

OpenAI
No model description is provided by the pricing endpoint.
$1.00
Input
$2.00
Output
Context -Group ratio 1x
Per tokenopenai

gpt-3.5-turbo-16k

OpenAI
No model description is provided by the pricing endpoint.
$3.00
Input
$4.00
Output
Context -Group ratio 1x
Per tokenopenai

gpt-3.5-turbo-16k-0613

OpenAI
No model description is provided by the pricing endpoint.
$3.00
Input
$4.00
Output
Context -Group ratio 1x
Per tokenopenai

gpt-3.5-turbo-instruct

OpenAI
No model description is provided by the pricing endpoint.
$1.50
Input
$2.00
Output
Context -Group ratio 1x
Per tokenopenai

gpt-35-turbo-instruct

OpenAI
No model description is provided by the pricing endpoint.
$75.00
Input
$75.00
Output
Context -Group ratio 1x
Per tokenopenai

gpt-4

OpenAI
No model description is provided by the pricing endpoint.
$30.00
Input
$60.00
Output
Context -Group ratio 1x
Per tokenopenai

gpt-4-0125-preview

OpenAI
No model description is provided by the pricing endpoint.
$10.00
Input
$30.00
Output
Context -Group ratio 1x
Per tokenopenai

gpt-4-0314

OpenAI
No model description is provided by the pricing endpoint.
$30.00
Input
$60.00
Output
Context -Group ratio 1x
Per tokenopenai
1-12 of 967
1 / 81

AI API pricing guide

Compare token costs before you route production traffic

AI API pricing is rarely one number. Compare input, output, cache, billing group, and endpoint signals together so you can choose a model that fits both quality and budget. Route Key keeps these model details in one searchable catalog and lets you route compatible requests without changing your application shape.

Input, output, and cache pricing

Input tokens are the prompt and context you send. Output tokens are generated text and often cost more. Cache pricing shows whether repeated context can be served at a different rate. Check all three when estimating a real request.

Token-based versus per-request billing

Most chat and reasoning models bill by tokens, while some endpoints add a per-request fee or use a different quota unit. The billing filter helps you compare like for like before selecting a model.

Group ratio and model ratio

A group ratio represents the account or channel multiplier applied to a model. A model ratio describes the model-specific adjustment. Together they explain why the same model can have different effective costs across groups.

GPT, Claude, Gemini, DeepSeek, and Qwen comparisons

Use the vendor and search filters to compare GPT API cost, Claude API pricing, Gemini pricing, DeepSeek cost, Qwen pricing, and other supported families from the same catalog.

Choose the right API endpoint

Confirm that the endpoint supports your request type, such as chat, responses, embeddings, images, or audio. An enabled model ID and endpoint combination is the safest starting point for an integration.

A practical model selection workflow

Start with task quality, then compare input and output cost, context length, latency expectations, and channel health. Test a small sample, inspect Usage Logs, and only then move the model into a larger routing policy.

AI API pricing FAQ

Common model cost questions

Can I compare GPT, Claude, Gemini, DeepSeek, Grok, and Qwen API pricing in Route Key?

Yes. The Models page reads the pricing endpoint and lets you compare supported model families, vendors, billing groups, endpoints, and price signals before sending production traffic.

Does Route Key help lower AI API cost?

Route Key can route requests through healthy lower-cost channels while keeping an OpenAI-compatible request shape, so teams can test cost-aware routing without rewriting their app.

Is the model pricing page useful before choosing an API model?

Yes. You can search and filter models by vendor, group, billing type, tag, and endpoint, then inspect input, output, cache, and group-ratio pricing signals.

Comparison resources

Continue into model pricing comparisons

Use the live catalog for current model signals, then open the comparison guides when you need a deeper pricing or integration workflow.

Complete model ID indexBrowse all 965 available model IDs

Expand this index to inspect every model ID currently returned by the pricing catalog. Select an ID to filter the live model table.