Full model list

80 callable models — 67 text across 10 vendors, plus 6 image and 7 video, with pricing, context and capabilities

The platform currently offers 80 callable models: 67 text chat models across 10 vendors, plus 6 image and 7 video generators. Text pricing is in USD per 1M tokens; images bill per image and video per second — see Image / video / music APIs.

Live data comes from the public catalog endpoint GET https://jiuye.zsopc.com/api/v1/models/public. This page is synced per release; the online model plaza is authoritative.

Reading the tables#

ColumnMeaning
capability toolsfunction calling
capability visionimage input
capability web_searchbuilt-in web search tool
capability image_gen / video_gengenerate images / video inside a conversation
capability thinkingreasoning / thinking mode
capability jsonenforced JSON output
cachedthe price when input hits the prompt cache

Cached pricing needs no opt-in: when upstream usage reports cached_tokens, the gateway settles that portion of the input at the cached rate automatically — clients don't (and can't) enable it explicitly. Models without a cached figure bill all input at the Input price.

Limited-time promotion: the whole Anthropic Claude line is charged at 40% off (list price × 0.6) and the whole OpenAI GPT text line at 70% off (× 0.3), with cached rates discounted equally. The tables show list prices; actual billing always uses the discounted rate, for both API-key calls and in-platform conversations. Discounted models carry an "X% OFF" badge on the model plaza.

Anthropic (7)#

Model IDNameContextInput (积分/1M)Output (积分/1M)Capabilities
claude-haiku-4-5-20251001Claude Haiku 4.5200K404.4 (cached 40.44)2022tools / vision / web_search / thinking / json
claude-opus-4-5-thinkingClaude Opus 4.5 Thinking200K6066 (cached 606.6)30330tools / vision / web_search / thinking / json
claude-opus-4-6Claude Opus 4.61M6066 (cached 606.6)30330tools / vision / web_search / thinking / json
claude-opus-4-6-thinkingClaude Opus 4.6 Thinking1M6066 (cached 606.6)30330tools / vision / web_search / thinking / json
claude-opus-4-7Claude Opus 4.71M6066 (cached 606.6)30330tools / vision / web_search / thinking / json
claude-opus-4-8Claude Opus 4.81M6066 (cached 606.6)30330tools / vision / web_search / thinking / json
claude-sonnet-4-6Claude Sonnet 4.61M1213.2 (cached 121.32)6066tools / vision / web_search / thinking / json

OpenAI (2)#

Model IDNameContextInput (积分/1M)Output (积分/1M)Capabilities
gpt-5.4GPT-5.4400K1011 (cached 101.1)6066tools / vision / web_search / image_gen / thinking / json
gpt-5.5GPT-5.51M202212132tools / vision / web_search / image_gen / thinking / json

Google (8)#

Model IDNameContextInput (积分/1M)Output (积分/1M)Capabilities
gemini-2.5-flashGemini 2.5 Flash1M121.32 (cached 30.33)1011tools / vision / json
gemini-2.5-flash-liteGemini 2.5 Flash Lite1M40.44 (cached 10.11)161.76tools / vision / web_search / json
gemini-2.5-flash-thinkingGemini 2.5 Flash Thinking1M121.32 (cached 30.33)1011tools / vision / web_search / thinking / json
gemini-2.5-proGemini 2.5 Pro2M505.5 (cached 125.36)4044tools / vision / thinking / json
gemini-3-flashGemini 3 Flash1M161.76 (cached 40.44)1213.2tools / vision / web_search / thinking / json
gemini-3-flash-previewGemini 3 Flash (Preview)1M161.76 (cached 40.44)1213.2tools / vision / json
gemini-3-pro-previewGemini 3 Pro (Preview)2M606.6 (cached 151.65)4852.8tools / vision / thinking / json
gemini-3.1-pro-lowGemini 3.1 Pro (Low)2M606.6 (cached 151.65)4852.8tools / vision / web_search / thinking / json

xAI (13)#

Model IDNameContextInput (积分/1M)Output (积分/1M)Capabilities
grok-3-miniGrok 3 Mini131K121.32202.2tools / thinking / json
grok-3-mini-fastGrok 3 Mini Fast131K242.641617.6tools / thinking / json
grok-4Grok 4256K1213.26066tools / vision / thinking / json
grok-4-fast-non-reasoningGrok 4 Fast (Non-Reasoning)2M80.88202.2tools / json
grok-4-fast-reasoningGrok 4 Fast (Reasoning)2M80.88202.2tools / thinking / json
grok-4.1Grok 4.1256K202210110tools / vision / thinking / json
grok-4.2Grok 4.2256K202210110tools / vision / thinking / json
grok-4.20-0309-non-reasoningGrok 4.20 (Non-Reasoning)256K1213.26066tools / vision / json
grok-4.20-0309-reasoningGrok 4.20 (Reasoning)256K1213.26066tools / vision / thinking / json
grok-4.20-multi-agent-0309Grok 4.20 Multi-Agent256K202210110tools / vision / thinking / json
grok-4.3Grok 4.3256K1213.26066tools / vision / video_gen / thinking / json
grok-build-0.1Grok Build 0.1256K1213.26066tools / thinking / json
grok-composer-2.5-fastGrok Composer 2.5 Fast131K404.42022

Zhipu GLM (5)#

Model IDNameContextInput (积分/1M)Output (积分/1M)Capabilities
glm-4-7-251222GLM-4.7 (X4)200K283.08283.08tools / json / open-source
glm-4.7GLM-4.7200K283.08283.08tools / thinking / json / open-source
glm-5GLM-5200K456.971593.34tools / thinking / json / open-source
glm-5-turboGLM-5-Turbo200K485.281617.6tools / json / open-source
glm-5.1GLM-5.1200K456.971593.34tools / thinking / json / open-source

DeepSeek (6)#

Model IDNameContextInput (积分/1M)Output (积分/1M)Capabilities
deepseek-v3-0324DeepSeek V3 (0324)128K109.19 (cached 28.31)444.84tools / json / open-source
deepseek-v3-2-251201DeepSeek V3.2 (X4)64K109.19444.84tools / json / open-source
deepseek-v3.1-terminusDeepSeek V3.1 Terminus128K109.19 (cached 28.31)444.84tools / json / open-source
deepseek-v3.2DeepSeek V3.264K109.19 (cached 28.31)444.84tools / json / open-source
deepseek-v4-flashDeepSeek V4 Flash128K60.66242.64tools / json
deepseek-v4-proDeepSeek V4 Pro128K202.2808.8tools / thinking / json

ByteDance Doubao (19)#

Model IDNameContextInput (积分/1M)Output (积分/1M)Capabilities
doubao-1-5-lite-32k-250115Doubao 1.5 Lite (32k)32K17.3934.37tools / json
doubao-1-5-pro-32k-250115Doubao 1.5 Pro (32k)32K45.7114.04tools / json
doubao-1-5-pro-32k-character-250715Doubao 1.5 Pro Character (32k)32K45.7114.04tools / json
doubao-1-5-vision-pro-32k-250115Doubao 1.5 Vision Pro (32k)32K171.06512.78tools / vision / json
doubao-seed-1-6-250615Doubao Seed 1.6128K45.7455.76tools / vision / json
doubao-seed-1-6-251015Doubao Seed 1.6 (10-15)128K45.7455.76tools / vision / json
doubao-seed-1-6-flash-250615Doubao Seed 1.6 Flash128K8.985.73tools / vision / json
doubao-seed-1-6-flash-250828Doubao Seed 1.6 Flash (08-28)128K8.985.73tools / vision / json
doubao-seed-1-6-vision-250815Doubao Seed 1.6 Vision128K45.7455.76tools / vision / json
doubao-seed-1-8-251228Doubao Seed 1.8128K45.7455.76tools / vision / json
doubao-seed-2-0-code-preview-260215Doubao Seed 2.0 Code Preview256K182.38 (cached 36.4)911.52tools / vision / thinking / json
doubao-seed-2-0-lite-260215Doubao Seed 2.0 Lite128K34.37 (cached 6.87)205.03tools / vision / json
doubao-seed-2-0-lite-260428Doubao Seed 2.0 Lite (04-28)128K34.37 (cached 6.87)205.03tools / vision / json
doubao-seed-2-0-mini-260215Doubao Seed 2.0 Mini64K11.73 (cached 2.43)114.04tools / vision / json
doubao-seed-2-0-mini-260428Doubao Seed 2.0 Mini (04-28)64K11.73 (cached 2.43)114.04tools / vision / json
doubao-seed-2-0-pro-260215Doubao Seed 2.0 Pro256K182.38 (cached 36.4)911.52tools / vision / thinking / json
doubao-seed-character-251128Doubao Seed Character32K45.7455.76tools / json
doubao-seed-code-preview-251028Doubao Seed Code Preview128K68.34455.76tools / vision / json
doubao-seed-translation-250915Doubao Seed Translation32K68.34205.03

Tencent Hunyuan (3)#

Model IDNameContextInput (积分/1M)Output (积分/1M)Capabilities
hunyuan-2.0-instruct-20251111Tencent Hunyuan 2.0 Instruct144K202.2808.8tools / json
hunyuan-2.0-thinking-20251109Tencent Hunyuan 2.0 Thinking192K202.2808.8tools / thinking / json
hunyuan-role-latestTencent Hunyuan Role32K109.19444.84tools / json

Moonshot Kimi (2)#

Model IDNameContextInput (积分/1M)Output (积分/1M)Capabilities
kimi-k2.5Kimi K2.5256K242.641011tools / thinking / json / open-source
kimi-k2.6Kimi K2.6256K242.641011tools / thinking / json / open-source

MiniMax (2)#

Model IDNameContextInput (积分/1M)Output (积分/1M)Capabilities
MiniMax-M2.5MiniMax M2.5200K121.32606.6tools / json
MiniMax-M2.7MiniMax M2.7200K121.32606.6tools / json

Image generation models (9)#

Billed per image through POST /v1/images/generations. See Image / video APIs.

Model IDNamePriceEndpoint
doubao-seedream-4-0-250828Seedream 4.011.73/image/v1/images/generations (size ≥ 960²)
doubao-seedream-4-5-251128Seedream 4.514.96/image/v1/images/generations (size ≥ 1920²)
doubao-seedream-5-0-260128Seedream 5.012.94/image/v1/images/generations (size ≥ 1920²)
doubao-seedream-5-0-pro-260628Seedream 5.0 Pro17.79 per image for output ≤ 2.36 MP, 35.59 above, plus 1.21 per input reference image/v1/images/generations (size ≥ 960²)
grok-imagine-imageGrok Imagine (Image)28.31/image/v1/images/generations
grok-imagine-image-qualityGrok Imagine (Quality)28.31/image/v1/images/generations
nano-bananaNano Banana (Gemini 2.5 Flash Image)15.77/image/v1/images/generations
nano-banana-proNano Banana Pro (Gemini 3 Pro Image)54.19/image/v1/images/generations
nano-banana-2Nano Banana 2 (Gemini 3.1 Flash Image)16.18/image/v1/images/generations

Since 2026-07-08 the three nano entries are the only way in: the older ids gemini-2.5-flash-image, gemini-3-pro-image-preview, gemini-3.1-flash-image(-preview) and the tvod-nano-* family all keep working as aliases resolving onto the canonical entries above. The former chat / Gemini-native surface for gemini-3.1-flash-image has been retired — call /v1/images/generations instead.

Video generation models (10)#

Billed per second as an async task through POST /v1/videos/generations. See Image / video APIs.

Model IDNamePrice
doubao-seedance-1-0-pro-fast-251015Seedance 1.0 Pro Fast32.35/second
doubao-seedance-1-0-pro-250528Seedance 1.0 Pro60.66/second
doubao-seedance-1-5-pro-251215Seedance 1.5 Pro72.79/second
doubao-seedance-2-0-fast-260128Seedance 2.0 Fast48.53/second
doubao-seedance-2-0-260128Seedance 2.088.97/second
doubao-seedance-2.0Seedance 2.0480p 33.16 / 720p 59.45 / 1080p 147.61 / 2k 291.17 / 4k 355.87 per second
doubao-seedance-2.0-fastSeedance 2.0 Fast480p 23.86 / 720p 47.72 / 1080p 117.28 / 2k 141.54 / 4k 169.85 per second
doubao-seedance-2.0-miniSeedance 2.0 Mini480p 14.96 / 720p 29.93 per second
grok-imagine-videoGrok Imagine Video283.08/second
grok-imagine-video-1.5-previewGrok Imagine Video 1.5586.38/second

Picking by job#

JobSuggested modelWhy
Strongest reasoning / agentsclaude-opus-4-8, claude-opus-4-71M context plus tools, vision and thinking
Everyday codingclaude-sonnet-4-6, gpt-5.4, doubao-seed-2-0-code-preview-260215a good price/performance balance
Sub-second responsesgemini-2.5-flash-lite, doubao-seed-1-6-flash-250615low latency, low price
Chinese-language workglm-5.1, deepseek-v3.2, hunyuan-2.0-instruct-20251111domestic vendors
Long contextgemini-2.5-pro (2M), grok-4-fast-* (2M), claude-opus-4-6+ (1M), gpt-5.5 (1M)context window
Cheap at volumedoubao-seed-1-6-flash-250615 (8.9/M), doubao-1-5-lite-32k-250115 (17.39/M)around 8.09–20.22/M
Multimodal visionclaude-*, gpt-5.x, gemini-*, grok-4.x, doubao-seed-1-6-visionvision input supported
Role-playdoubao-seed-character-251128, hunyuan-role-latesttuned for character work
Translationdoubao-seed-translation-250915tuned for translation

Cross-vendor protocol support#

Different models support different client SDKs, and the compatibility matrix is authoritative on what is actually reachable. The catalog's supported_protocols field (returned by GET /api/v1/models/public, the same source as the model plaza cards) reflects only the native / primary path; many models are additionally reachable on other protocols through gateway translation or an upstream channel (the response then carries an X-Protocol-Translation header), and those appear only in the matrix. Native paths as of June 2026:

  • OpenAI Chat (/v1/chat/completions) — 77 of 78 text models (only doubao-seed-translation-250915 is excluded), by far the widest coverage
  • Anthropic Messages (/v1/messages) — natively: all 9 Claude models plus 4 Gemini SKUs (gemini-2.5-flash-lite, gemini-2.5-flash-thinking, gemini-3-flash, gemini-3.1-pro-low). GPT-5.x and most third-party models are also reachable via upstream conversion or translation — see the matrix
  • OpenAI Responses (/v1/responses) — gpt-5.4 / gpt-5.5 (the primary path for Codex CLI) plus claude-opus-4-6, claude-opus-4-7, claude-opus-4-8, claude-sonnet-4-6 and claude-haiku-4-5-20251001
  • Gemini Native (/v1beta/...) — natively: gemini-2.5-flash, gemini-2.5-pro, gemini-3-flash-preview, gemini-3-pro-preview. Some other models are reachable via upstream conversion — see the matrix

For a model × protocol combination with no reachable channel, the gateway returns 503 no_channel_available (the model exists but has no channel on that protocol — not a 404); switch to a supported protocol.

Live data#

The model list and prices track upstream changes. For live data see https://jiuye.zsopc.com/models or the public catalog endpoint (no auth required):

curl https://jiuye.zsopc.com/api/v1/models/public | jq .

The response looks like { "models": [...] }. Note that this endpoint returns every registered entry — including placeholder SKUs that aren't live yet and non-chat SKUs — so filter on the three fields below when consuming it from a script rather than looping over the whole list:

FieldMeaning
callablefalse means a placeholder SKU (listed but not live); calling it returns model_not_found
endpoint_typenull means an ordinary chat model. images_generations, videos_generations, contents_generations_tasks, embeddings and similar mean the model uses its own dedicated endpoint and cannot be sent to /v1/chat/completions — see Image / video / music APIs
supported_protocolsthe client protocols available for this model (see "Cross-vendor protocol support" above)

To take only the models you can send straight to /v1/chat/completions:

curl -s https://jiuye.zsopc.com/api/v1/models/public \
  | jq -r '.models[] | select(.callable and (.supported_protocols | index("openai_chat"))) | .id'