Your privacy choices

Allow optional cookies for referral attribution, visit analytics, and Google Ads purchase measurement.

Concept landscape: an Inner Mongolian grassland data center and an animated AI model network

Intelligence,withoutlimits.

GPT, Claude, DeepSeek and Qwen. One API key.

Tutorial centerClient setup guides

One key. Your choice of models.

A clear setup guide for the tools you use.

Make your first request.

A clear setup guide for the tools you use.

Example request

Example request
View models & pricing
curl https://api.llm-token.cn/v1/chat/completions \  -H "Authorization: Bearer YOUR_API_KEY" \  -H "Content-Type: application/json" \  -d '{  "model": "gpt-5.6-sol",  "messages": [    { "role": "user", "content": "Hello" }  ]}'
Chat CompletionsResponsesMessages
Connect without a rewrite

Keep your client. Change the model.

The gateway exposes OpenAI-compatible Chat Completions and Responses routes plus an Anthropic-compatible Messages route. Check the tutorial for the exact Base URL, path, and model ID.

01OpenAI-compatiblePOST /v1/chat/completions
02ResponsesPOST /v1/responses
03MessagesPOST /v1/messages
00Base URLhttps://api.llm-token.cn/v1
Read setup guides
Model catalog · Live prices

Choose a route for your workload

75Model catalog
Pricing

Showing 75 of 75 models

  • Claude Sonnet 4.6
  • GPT-5.6 Sol
  • GPT-5.6 Terra
  • Grok 4.6
  • Grok 4.7
  • Grok 4.5
  • Claude Haiku 4.5
  • Claude Sonnet 5
  • Claude Opus 4.6
  • Claude Opus 4.7
  • Claude Opus 4.8
  • Claude Opus 5
  • Claude Opus 5.5
  • Qwen 3.7 Plus
  • Qwen 3.7 Max
  • Qwen 3.8 Max
  • MiniMax M3
  • Doubao Seed 2.1 Turbo
  • MiMo V2.6 Pro
  • MiMo V2.6 Flash
  • DeepSeek V4 Pro
  • DeepSeek V4 Flash
  • Kimi K3
  • Kimi K2.7
  • GLM 5.2
  • GLM 5.3
  • GLM 5.3 Flash
  • GPT-5.6 Luna
  • GPT-6 Astra
  • DeepSeek V4.1 Flash
  • GPT Image 2.5 Flare
  • GPT Image 2.5 Sunburst
  • GPT-6 Sol
  • GPT-6 Luna
  • claude-opus-4-5-20251101
  • claude-opus-4-5-20251101-thinking
  • claude-opus-4-6-high
  • claude-opus-4-6-low
  • claude-opus-4-6-max
  • claude-opus-4-6-medium
  • claude-opus-4-6-thinking
  • claude-opus-4-7-high
  • claude-opus-4-7-low
  • claude-opus-4-7-max
  • claude-opus-4-7-medium
  • claude-opus-4-7-thinking
  • claude-opus-4-7-xhigh
  • claude-opus-5-high
  • claude-opus-5-low
  • claude-opus-5-max
  • claude-opus-5-medium
  • claude-opus-5-thinking
  • claude-opus-5-xhigh
  • doubao-seed-2.0-code
  • doubao-seed-2.0-lite
  • doubao-seed-2.0-pro
  • gemini-3-flash
  • gemini-3.1-pro
  • gemini-3.5-flash
  • gpt-5.4-openai-compact
  • gpt-5.5-openai-compact
  • gpt-image-2.5
  • jev-1.13.0
  • jev-latest
  • jev-preview
  • kimi-k2.6
  • LongCat-2.5
  • mimo-v2.6
  • MiniMax-M2.7
  • MiniMax-M2.7-highspeed
  • MiniMax-M3-highspeed
  • qwen3.8-flash
  • qwen3.8‑flash
  • 火山-豆包/doubao-seed-2.0-code
  • 火山-豆包/doubao-seed-2.1-turbo
Model catalog

Build with the model you need

  1. Code & agents

    Use Codex CLI, Claude Code, Cursor, OpenCode, or your own app with a familiar API.

    Docs
  2. Reasoning & chat

    Route everyday chat, long-context work, and reasoning to GPT, Claude, DeepSeek, Qwen, Kimi, or GLM.

    Docs
  3. Vision & images

    Choose supported vision and image-generation routes from the live model catalog before you call them.

    Docs
Start in 3 steps

From a prepaid balance to your first request.

  1. Choose a plan

    Pick prepaid credits for the models and volume you expect.

  2. Add credits

    Complete checkout with a payment method shown for your account.

  3. Send a request

    Copy the API key, set the documented Base URL, and send a test call.

Support

Common questions

How do I buy an API key?

Open the access page, choose a package, enter the delivery email, and complete payment. New credentials are delivered according to the checkout instructions; also check the spam folder.

How do I top up an existing API key?

Use the top-up page and enter the existing key carefully. Never paste a full key into a public ticket, group chat, screenshot, or shared document.

Where can I check balance and usage?

Use the official quota portal linked on this site. Enter the key only on that official page and never send the full credential to support.

Which Base URL should I use?

Use https://api.llm-token.cn/v1 everywhere. It is a globally routed entry point (nodes in East China, South China, Hong Kong, Singapore, the US, and Germany) that automatically connects you to the fastest route.

Which API formats are supported?

Chat Completions-compatible, Responses-compatible, and Messages-compatible formats are available. Select the format expected by your library, client, or agent workflow.

Which model families are available?

The catalog covers major GPT, Claude, Grok, DeepSeek, Qwen, Doubao, Kimi, GLM, MiniMax, Hunyuan, StepFun, MiMo, and LongCat families. Check the pricing page for current versions, capabilities, and rates.

Are image generation and multimodal inputs supported?

Yes, for models marked with those capabilities in the live catalog. Confirm whether the selected model accepts image URLs, Base64 input, or produces images, and follow its dedicated request format.

How is pricing calculated?

Credits are prepaid and deducted using the live rate for each model, including input/output tokens, multipliers, or per-task fees where applicable. Review the pricing page before purchase or production use.

Read the full FAQ
LLM API Gateway

Ready to send a real request?

Start with one key, follow the client-specific tutorial, and switch models by changing the model ID.