Google / API Pricing and Routing

Google gemini-2.5-flash-lite-preview-09-2025-thinking-4096 API pricing and routing

Inspect the live model ID, provider, context window, endpoint support, billing group, and Route Key pricing signals before sending production traffic.

Model ID
gemini-2.5-flash-lite-preview-09-2025-thinking-4096

Compare gemini-2.5-flash-lite-preview-09-2025-thinking-4096 API pricing, context length, endpoints, billing groups, and routing availability through Route Key's OpenAI-compatible AI API gateway.

Pricing signals

Input · per token / 1M tokens
$0.10
Output · per token / 1M tokens
$0.40
Cached input · per token / 1M tokens
$0.01

Billing: per token · Available groups: default, gemini

Supported endpoints

gemini, openai

Context window

-

Model technical profile

Vendor
Google
Billing
per token
Pricing schema
shell-api-modelmak-v1
Model ratio
x0.05
Completion ratio
x4
Cache read ratio
x0.1
Audio input ratio
x3
Audio output ratio
x4.8
Image ratio
x1

Use this model with Route Key

Keep your existing OpenAI-compatible request shape and change the gateway, key, and model values.

Base URL
https://api.routekey.ai/v1
API key
Bearer sk-route-key
Model
gemini-2.5-flash-lite-preview-09-2025-thinking-4096

Frequently asked questions about gemini-2.5-flash-lite-preview-09-2025-thinking-4096

What is gemini-2.5-flash-lite-preview-09-2025-thinking-4096?

gemini-2.5-flash-lite-preview-09-2025-thinking-4096 is a currently listed model route from Google. Confirm the live catalog for its exact ID, endpoint, group availability, and pricing signal.

How is gemini-2.5-flash-lite-preview-09-2025-thinking-4096 API pricing calculated?

Route Key exposes the current input, output, and cached-input pricing signals for gemini-2.5-flash-lite-preview-09-2025-thinking-4096. Effective cost can vary by billing group and route ratio.

How do I use gemini-2.5-flash-lite-preview-09-2025-thinking-4096 with Route Key?

Set the Route Key base URL, add a Route Key API key, and send the exact model ID shown below. Start with a small request before enabling streaming or tools.