GPT-4.1 API via TokenMix

Use GPT-4.1 from OpenAI as a chat model through the TokenMix AI API relay and multi-model gateway.

GPT-4.1 is OpenAI's model optimized for coding, instruction following, and long-context tasks. It supports 1M tokens of context with 32K max output, and offers major improvements over GPT-4o in coding benchmarks and multi-step task execution.

API access

  • Base URL: https://api.tokenmix.ai/v1
  • Model ID: gpt-4.1
  • OpenAI SDK compatible. Change the base URL and use your TokenMix API key.

Pricing

Input $1.94/M tokens, output $7.76/M tokens

Capabilities

Vision, Function calling, JSON mode, Streaming

Model specs

  • Context: 1049K tokens
  • Max output: 33K tokens

Availability

3/3 available API endpoints are healthy right now.

Recent performance

TTFT 2867ms, latency 16644ms, throughput 125.7 tok/s.

Start using this model

Create an API key, top up from $1 when needed, and call this model through the TokenMix OpenAI-compatible endpoint.

Create API key · View pricing · Quickstart