Qwen 3.8 Flash API via TokenMix
Use Qwen 3.8 Flash from Qwen as a chat model through the TokenMix AI API relay and multi-model gateway.
Alibaba Qwen 3.8 Flash multimodal model with 1M context, vision, structured output and tool calling.
API access
- Base URL:
https://api.tokenmix.ai/v1 - Model ID:
qwen3.8-flash - OpenAI SDK compatible. Change the base URL and use your TokenMix API key.
Pricing
Input $0.1017/M tokens, output $0.3438/M tokens
Capabilities
Vision, Function calling, JSON mode, Streaming, Reasoning
Model specs
- Context: 1000K tokens
- Max output: 131K tokens
Availability
1/1 available API endpoints are healthy right now.
Recent performance
TTFT 1701ms, latency 9432ms, throughput 75.4 tok/s.
Start using this model
Create an API key, top up from $1 when needed, and call this model through the TokenMix OpenAI-compatible endpoint.