GPT-4.1 Nano API via TokenMix
Use GPT-4.1 Nano from OpenAI as a chat model through the TokenMix AI API relay and multi-model gateway.
GPT-4.1 Nano is the fastest and cheapest model in the GPT-4.1 series. It supports 1M token context and 32K max output, designed for classification, autocompletion, and lightweight tasks where speed and cost matter most.
API access
- Base URL:
https://api.tokenmix.ai/v1 - Model ID:
gpt-4.1-nano - OpenAI SDK compatible. Change the base URL and use your TokenMix API key.
Pricing
Input $0.097/M tokens, output $0.388/M tokens
Capabilities
Vision, Function calling, JSON mode, Streaming
Model specs
- Context: 1049K tokens
- Max output: 33K tokens
Availability
2/2 available API endpoints are healthy right now.
Recent performance
TTFT 1017ms, latency 2494ms, throughput 135.7 tok/s.
Start using this model
Create an API key, top up from $1 when needed, and call this model through the TokenMix OpenAI-compatible endpoint.