Gemini 2.5 Flash Lite API via TokenMix

Use Gemini 2.5 Flash Lite from Google as a chat model through the TokenMix AI API relay and multi-model gateway.

Google's fastest and most economical multimodal model. Optimized for low latency and high-volume use cases. Supports adjustable thinking budgets. Deprecated March 31, 2026.

API access

Base URL: https://api.tokenmix.ai/v1
Model ID: gemini-2.5-flash-lite
OpenAI SDK compatible. Change the base URL and use your TokenMix API key.

Pricing

Input $0.097/M tokens, output $0.388/M tokens

Capabilities

Vision, Function calling, JSON mode, Streaming, Reasoning

Model specs

Context: 1049K tokens
Max output: 66K tokens

Availability

1/1 available API endpoints are healthy right now.

Recent performance

TTFT 635ms, latency 1547ms, throughput 182.9 tok/s.

Start using this model

Create an API key, top up from $1 when needed, and call this model through the TokenMix OpenAI-compatible endpoint.

Create API key · View pricing · Quickstart