Google / API Pricing and Routing

Google gemini-2.5-flash-lite-preview-09-2025-thinking-8192 API pricing and routing

Inspect the live model ID, provider, context window, endpoint support, billing group, and Route Key pricing signals before sending production traffic.

Model ID
gemini-2.5-flash-lite-preview-09-2025-thinking-8192

Compare gemini-2.5-flash-lite-preview-09-2025-thinking-8192 API pricing, context length, endpoints, billing groups, and routing availability through Route Key's OpenAI-compatible AI API gateway.

Pricing signals

Input · per token / 1M tokens
$0.10
Output · per token / 1M tokens
$0.40
Cached input · per token / 1M tokens
$0.01

Billing: per token · Available groups: default, gemini

Supported endpoints

gemini, openai

Context window

-

Model technical profile

Vendor
Google
Billing
per token
Pricing schema
shell-api-modelmak-v1
Model ratio
x0.05
Completion ratio
x4
Cache read ratio
x0.1
Audio input ratio
x3
Audio output ratio
x4.8
Image ratio
x1

Use this model with Route Key

Keep your existing OpenAI-compatible request shape and change the gateway, key, and model values.

Base URL
https://api.routekey.ai/v1
API key
Bearer sk-route-key
Model
gemini-2.5-flash-lite-preview-09-2025-thinking-8192

Frequently asked questions about gemini-2.5-flash-lite-preview-09-2025-thinking-8192

What is gemini-2.5-flash-lite-preview-09-2025-thinking-8192?

gemini-2.5-flash-lite-preview-09-2025-thinking-8192 is a currently listed model route from Google. Confirm the live catalog for its exact ID, endpoint, group availability, and pricing signal.

How is gemini-2.5-flash-lite-preview-09-2025-thinking-8192 API pricing calculated?

Route Key exposes the current input, output, and cached-input pricing signals for gemini-2.5-flash-lite-preview-09-2025-thinking-8192. Effective cost can vary by billing group and route ratio.

How do I use gemini-2.5-flash-lite-preview-09-2025-thinking-8192 with Route Key?

Set the Route Key base URL, add a Route Key API key, and send the exact model ID shown below. Start with a small request before enabling streaming or tools.