LLM API Gateway User Guide (English)

English setup reference for API endpoints, model IDs, billing, and supported clients.

Welcome to Guishu LLM Token. This is the English customer edition of the Guishu Token guide. It has been rewritten for international users, with practical setup steps, clear pricing language, and a more familiar support style for UK and overseas customers.

1. Start Here

In most cases, you only need three things: your API key, the compatible API base URL, and the model ID you want to use.

What you needUse this
Buy or top upAccount center
Check balance and usagehttps://token.gpt-agent.cc
International default API base URLhttps://api.llm-token.cn/v1
Chat completions endpointhttps://api.llm-token.cn/v1/chat/completions
Image generation endpointhttps://api.llm-token.cn/v1/images/generations
Image web apphttps://image2.gpt-agent.cc/

Configuration tip: Some clients ask for a Base URL, while others ask for the full endpoint. If a client asks for a Base URL, use https://api.llm-token.cn/v1. If it asks for a complete chat endpoint, use https://api.llm-token.cn/v1/chat/completions.

curl https://api.llm-token.cn/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "claude-sonnet-4-6",
    "messages": [
      {"role": "user", "content": "Please reply with one short sentence."}
    ]
  }'

2. Buying, Balance and Billing

Self-service

Enterprise support

  • For larger top-ups, invoices, enterprise onboarding or dedicated routing, contact your account manager or support contact.
  • Enterprise teams can request account-level guidance, deployment support and usage monitoring.

实时模型与计费口径

输入、输出和缓存输入可能采用不同单价。图片、视频或按次计费的模型请核对计费单位。模型 ID 和价格以实时定价页为准,实际扣费以账单记录为准。

查看实时模型价格

Price Monitoring Skill

If your team uses OpenClaw or an automation bot, you can monitor the public model and pricing table for changes. The monitor detects new models, removed models, rate changes and price changes, then produces a notification that can be sent to a team chat or automation channel.

clawhub install feishu-public-table-monitor

Already installed? Update to the latest version:

clawhub --workdir ~/.openclaw/workspace update feishu-public-table-monitor --force

3. Model Catalogue

从实时列表复制完整模型 ID。不同 API Key 或线路的访问范围可能不同,接入前请确认所选模型。

查看实时模型价格

当前显示 20 / 20 个模型
当前 20 个正式模型;点击模型行可展开详情。
模型厂商价格上下文图片能力复制操作
阿里云 / Qwen3 倍率 · ¥1.2 / 百万 Tokens1M支持图片
阿里云 / Qwen4 倍率 · ¥1.6 / 百万 Tokens1M支持图片
阿里云 / Qwen10 倍率 · ¥4 / 百万 Tokens1M文本输入
阿里云 / Qwen20 倍率 · ¥8 / 百万 Tokens1M支持图片
阿里云 / Qwen5 倍率 · ¥2 / 百万 Tokens1M支持图片
美团 / LongCat1 倍率 · ¥0.4 / 百万 Tokens1M文本输入
MiniMax1 倍率 · ¥0.4 / 百万 Tokens1M支持图片
阶跃星辰 / StepFun1 倍率 · ¥0.4 / 百万 Tokens256K支持图片
字节跳动 / Doubao3 倍率 · ¥1.2 / 百万 Tokens128K支持图片
小米 / MiMo4 倍率 · ¥1.6 / 百万 Tokens1M文本输入
小米 / MiMo4 倍率 · ¥1.6 / 百万 Tokens1M支持图片
小米 / MiMo2 倍率 · ¥0.8 / 百万 Tokens来源有差异1M支持图片
DeepSeek7.5 倍率 · ¥3 / 百万 Tokens来源有差异1M文本输入
DeepSeek2.5 倍率 · ¥1 / 百万 Tokens来源有差异1M文本输入
月之暗面 / Kimi25 倍率 · ¥10 / 百万 Tokens1M支持图片
智谱 / Z.ai8 倍率 · ¥3.2 / 百万 Tokens1M支持图片
智谱 / Z.ai8 倍率 · ¥3.2 / 百万 Tokens1M文本输入
智谱 / Z.ai8 倍率 · ¥3.2 / 百万 Tokens来源有差异1M支持图片
DeepSeek2.5 倍率 · ¥1 / 百万 Tokens来源有差异1M支持图片
阿里巴巴输入¥0.8/ 百万 Tokens输出¥4/ 百万 Tokens缓存¥0.08/ 百万 Tokens—待确认
阿里云 / Qwen1M
支持图片

3 倍率 · ¥1.2 / 百万 Tokens

阿里云 / Qwen1M
支持图片

4 倍率 · ¥1.6 / 百万 Tokens

阿里云 / Qwen1M
文本输入

10 倍率 · ¥4 / 百万 Tokens

阿里云 / Qwen1M
支持图片

20 倍率 · ¥8 / 百万 Tokens

阿里云 / Qwen1M
支持图片

5 倍率 · ¥2 / 百万 Tokens

美团 / LongCat1M
文本输入

1 倍率 · ¥0.4 / 百万 Tokens

MiniMax1M
支持图片

1 倍率 · ¥0.4 / 百万 Tokens

阶跃星辰 / StepFun256K
支持图片

1 倍率 · ¥0.4 / 百万 Tokens

字节跳动 / Doubao128K
支持图片

3 倍率 · ¥1.2 / 百万 Tokens

小米 / MiMo1M
文本输入

4 倍率 · ¥1.6 / 百万 Tokens

小米 / MiMo1M
支持图片

4 倍率 · ¥1.6 / 百万 Tokens

小米 / MiMo1M
支持图片来源有差异

2 倍率 · ¥0.8 / 百万 Tokens

DeepSeek1M
文本输入来源有差异

7.5 倍率 · ¥3 / 百万 Tokens

DeepSeek1M
文本输入来源有差异

2.5 倍率 · ¥1 / 百万 Tokens

月之暗面 / Kimi1M
支持图片

25 倍率 · ¥10 / 百万 Tokens

智谱 / Z.ai1M
支持图片

8 倍率 · ¥3.2 / 百万 Tokens

智谱 / Z.ai1M
文本输入

8 倍率 · ¥3.2 / 百万 Tokens

智谱 / Z.ai1M
支持图片来源有差异

8 倍率 · ¥3.2 / 百万 Tokens

DeepSeek1M
支持图片来源有差异

2.5 倍率 · ¥1 / 百万 Tokens

阿里巴巴非 Token 模型
待确认

输入¥0.8/ 百万 Tokens输出¥4/ 百万 Tokens缓存¥0.08/ 百万 Tokens

数据口径更新于 2026-09-25T11:10:04.896Z

从实时列表复制完整模型 ID。不同 API Key 或线路的访问范围可能不同,接入前请确认所选模型。

4. API Access for Developers

Regional API Base URLs

All regions use the same globally routed Base URL. Your API key, model IDs and request format stay the same wherever you are.

Customer locationRecommended API Base URLNotes
All regionshttps://api.llm-token.cn/v1Globally routed entry (nodes in East China, South China, Hong Kong, Singapore, the US, and Germany) that automatically picks the fastest route.

If access is ever unstable, retry with the same Base URL — routing will pick the next fastest node automatically.

Chat Completions

curl https://api.llm-token.cn/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gpt-5.5",
    "messages": [
      {"role": "user", "content": "Summarise this project in three bullet points."}
    ]
  }'

Image Generation

curl https://api.llm-token.cn/v1/images/generations \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gpt-image-2",
    "prompt": "A clean modern product banner for an AI platform, blue and white colour palette, professional SaaS style",
    "response_format": "url",
    "n": 1,
    "aspect_ratio": "16:9"
  }'

4K image output: Check the current model list for available routes, capabilities and billing units.

Image tip: URLs returned by image APIs may be temporary. Download or store generated images if they are needed for a website, campaign or production workflow.

Common Request Fields

FieldPurpose
modelThe model ID. Copy it exactly from the model list.
messagesUsed for chat-style requests.
promptUsed for image generation or simple text-style tasks.
response_formatFor images, url is usually easiest for testing and download.
aspect_ratioImage ratio, for example 1:1, 16:9 or 9:16.

5. Tool Setup Guides

Universal setup pattern: choose an OpenAI-compatible or custom provider, enter the Guishu API base URL, paste your API key, then choose the model ID.

Tool or guideEnglish setup instructions
Image generationUse https://api.llm-token.cn/v1/images/generations. Choose gpt-image-2 according to the required image workflow. Start with one image per request for testing.
OpenClawOpen settings, add an OpenAI-compatible provider, set Base URL to https://api.llm-token.cn/v1, paste your API key, then select a model such as claude-sonnet-4-6, MiniMax-M3 or gpt-5.5. For long-context work, confirm the client context window is not capped locally.
Tencent CloudWhen using a Tencent Cloud function, workflow or server-side app, store the Guishu API key as an environment variable and call the compatible endpoint from your backend. Avoid placing keys in frontend code.
Cherry StudioAdd a custom OpenAI-compatible provider. Use Base URL https://api.llm-token.cn/v1, paste your API key, and enter the model ID exactly as listed.
Claude CodeUse the supplied one-click script where available, or configure a compatible model route through your local provider settings. Recommended default: claude-sonnet-4-6. Use claude-opus-4-8 only for high-value work.
CodexConfigure an OpenAI-compatible endpoint with https://api.llm-token.cn/v1. Use coding-focused models such as gpt-5.6-terra, gpt-5.5, qwen3.7-plus or deepseek-v4-pro.
HermesConfigure the gateway or client provider with the Guishu base URL and API key. Recommended models are claude-sonnet-4-6 for high concurrency and MiniMax-M3 for long-context workflows.
gpt-image-2 image modelUse the image generation endpoint and set model to gpt-image-2. This model is best for higher-quality images, design assets and polished marketing visuals.
WorkBuddyAdd a custom AI provider. Set the Base URL to https://api.llm-token.cn/v1, paste the API key, then choose a model. If a saved model name fails, copy a model ID directly from this guide.
OpenCodeUse an OpenAI-compatible provider. For fast coding choose gpt-5.6-terra or step-3.7-flash. For complex refactors choose deepseek-v4-pro, gpt-5.5 or claude-opus-4-8.
VS CodeFor extensions that support custom OpenAI-compatible providers, enter the Base URL, API key and model ID. Test with a short prompt before starting a long coding session.
ChatboxSelect a custom OpenAI-compatible source. Use Base URL https://api.llm-token.cn/v1, paste your API key, and select a model. If the client asks for the full endpoint, use https://api.llm-token.cn/v1/chat/completions.
TraeUse the custom provider or OpenAI-compatible configuration. Choose a coding model and run a small test prompt before applying it to an existing repository.
CursorUse custom model/provider settings where available. Add the Guishu base URL and API key, then select a coding model. Keep high-rate models for complex planning or review rather than every autocomplete request.
Image web appOpen image2.gpt-agent.cc, choose the image model, enter your prompt and download the result after generation.
ScenarioRecommended models
Daily automation and high concurrencyclaude-sonnet-4-6, gpt-5.6-terra, MiniMax-M3
Complex reasoning and high-quality writinggpt-5.6-sol, qwen3.8-max, deepseek-v4-pro
Top-tier long-running Agent tasksclaude-opus-5
Fast code iterationgpt-5.6-terra, LongCat-2.0, qwen3.7-plus
Long-context documentsMiniMax-M3, qwen3.7-plus, LongCat-2.0, deepseek-v4-pro, kimi-k3, glm-5.2, glm-5.3, glm-5.3-flash
Image generationgpt-image-2, gpt-image-2-4k

6. Image Prompting Guide

For image generation, a clear prompt normally produces a better result. A useful structure is:

A [ratio] [use case] image of [subject], in [style], with [colour palette], [composition], [important constraints].

Example:

A 16:9 marketing banner for an AI platform, clean SaaS style, blue and white colour palette, a subtle data visualisation background, premium and professional, no text.

For social media or article covers, use 16:9. For app icons and avatars, use 1:1. For mobile posters, use 9:16.


7. Troubleshooting

IssueWhat to check
Authentication failedCheck that the header is exactly Authorization: Bearer YOUR_API_KEY. Make sure there are no extra spaces before or after the key.
Model not foundCopy the model ID exactly from the model list. Model names are case-sensitive and hyphen-sensitive.
Client rejects the URLIf the client asks for a Base URL, use https://api.llm-token.cn/v1. If it asks for a full endpoint, use https://api.llm-token.cn/v1/chat/completions.
Image URL expiredDownload generated images promptly. Temporary image URLs are not intended as permanent asset hosting.
Context window appears too smallSome local clients cap context length in their own settings. Increase the local context window where supported, especially in OpenClaw or similar Agent tools.
Slow streaming or timeoutImage generation and long Agent tasks can take longer than normal chat. Increase client timeout settings and test with a shorter request first.
Balance seems wrongCheck usage at token.gpt-agent.cc. Billing is based on actual model rate and token usage recorded by the system.

OpenClaw 16k Context Fix

If OpenClaw reports a rate limit or context issue while the selected model supports a much larger context window, the local client may still be capped at 16k. Check the OpenClaw model or gateway configuration and raise the context setting to match the model you intend to use.

Support Case Style

When reporting an issue to support, please include:

  • Your API key name or account identifier, but never paste the full secret key in a public chat.
  • The model ID you used.
  • The client or tool name.
  • The Base URL or endpoint you configured.
  • The exact error message and approximate time of the request.

8. Coverage of the Chinese Source Documents

This English guide consolidates and localises the main Chinese guide and its linked child documents. The content has been rewritten for international customers rather than translated word-for-word.

Chinese source areaWhere it appears in this English guide
Main Guishu Token user guideSections 1, 2, 4 and 7
Model listSection 3
Model pricing and billingSection 2 and Section 3
MiniMax image generation guideSection 4 and Section 6
OpenClaw setup guideSection 5 and OpenClaw context fix in Section 7
Tencent Cloud setup guideSection 5
Cherry Studio setup guideSection 5
Claude Code deployment guideSection 5
Codex deployment guideSection 5
Hermes deployment guideSection 5
gpt-image-2 image model guideSection 4 and Section 6
WorkBuddy setup guideSection 5
OpenCode setup guideSection 5
VS Code setup guideSection 5
Chatbox setup guideSection 5
Trae setup guideSection 5
Cursor setup guideSection 5
Web image model guideSection 5 and Section 6
FAQ and customer casesSection 7

Need help? For enterprise onboarding, invoicing, larger top-ups, dedicated routing or team deployment support, please contact your Guishu account manager or customer support contact.