Full model list

80 callable models — 67 text across 10 vendors, plus 6 image and 7 video, with pricing, context and capabilities

The platform currently offers 80 callable models: 67 text chat models across 10 vendors, plus 6 image and 7 video generators. Text pricing is in USD per 1M tokens; images bill per image and video per second — see Image / video / music APIs.

Live data comes from the public catalog endpoint GET https://model.zsopc.com/api/v1/models/public. This page is synced per release; the online model plaza is authoritative.

Reading the tables#

ColumnMeaning
capability toolsfunction calling
capability visionimage input
capability web_searchbuilt-in web search tool
capability image_gen / video_gengenerate images / video inside a conversation
capability thinkingreasoning / thinking mode
capability jsonenforced JSON output
cachedthe price when input hits the prompt cache

Cached pricing needs no opt-in: when upstream usage reports cached_tokens, the gateway settles that portion of the input at the cached rate automatically — clients don't (and can't) enable it explicitly. Models without a cached figure bill all input at the Input price.

Limited-time promotion: the whole Anthropic Claude line is charged at 40% off (list price × 0.6) and the whole OpenAI GPT text line at 70% off (× 0.3), with cached rates discounted equally. The tables show list prices; actual billing always uses the discounted rate, for both API-key calls and in-platform conversations. Discounted models carry an "X% OFF" badge on the model plaza.

Anthropic (7)#

Model IDNameContextInput ($/1M)Output ($/1M)Capabilities
claude-haiku-4-5-20251001Claude Haiku 4.5200K$1 (cached $0.1)$5tools / vision / web_search / thinking / json
claude-opus-4-5-thinkingClaude Opus 4.5 Thinking200K$15 (cached $1.5)$75tools / vision / web_search / thinking / json
claude-opus-4-6Claude Opus 4.61M$15 (cached $1.5)$75tools / vision / web_search / thinking / json
claude-opus-4-6-thinkingClaude Opus 4.6 Thinking1M$15 (cached $1.5)$75tools / vision / web_search / thinking / json
claude-opus-4-7Claude Opus 4.71M$15 (cached $1.5)$75tools / vision / web_search / thinking / json
claude-opus-4-8Claude Opus 4.81M$15 (cached $1.5)$75tools / vision / web_search / thinking / json
claude-sonnet-4-6Claude Sonnet 4.61M$3 (cached $0.3)$15tools / vision / web_search / thinking / json

OpenAI (2)#

Model IDNameContextInput ($/1M)Output ($/1M)Capabilities
gpt-5.4GPT-5.4400K$2.5 (cached $0.25)$15tools / vision / web_search / image_gen / thinking / json
gpt-5.5GPT-5.51M$5$30tools / vision / web_search / image_gen / thinking / json

Google (8)#

Model IDNameContextInput ($/1M)Output ($/1M)Capabilities
gemini-2.5-flashGemini 2.5 Flash1M$0.3 (cached $0.075)$2.5tools / vision / json
gemini-2.5-flash-liteGemini 2.5 Flash Lite1M$0.1 (cached $0.025)$0.4tools / vision / web_search / json
gemini-2.5-flash-thinkingGemini 2.5 Flash Thinking1M$0.3 (cached $0.075)$2.5tools / vision / web_search / thinking / json
gemini-2.5-proGemini 2.5 Pro2M$1.25 (cached $0.31)$10tools / vision / thinking / json
gemini-3-flashGemini 3 Flash1M$0.4 (cached $0.1)$3tools / vision / web_search / thinking / json
gemini-3-flash-previewGemini 3 Flash (Preview)1M$0.4 (cached $0.1)$3tools / vision / json
gemini-3-pro-previewGemini 3 Pro (Preview)2M$1.5 (cached $0.375)$12tools / vision / thinking / json
gemini-3.1-pro-lowGemini 3.1 Pro (Low)2M$1.5 (cached $0.375)$12tools / vision / web_search / thinking / json

xAI (13)#

Model IDNameContextInput ($/1M)Output ($/1M)Capabilities
grok-3-miniGrok 3 Mini131K$0.3$0.5tools / thinking / json
grok-3-mini-fastGrok 3 Mini Fast131K$0.6$4tools / thinking / json
grok-4Grok 4256K$3$15tools / vision / thinking / json
grok-4-fast-non-reasoningGrok 4 Fast (Non-Reasoning)2M$0.2$0.5tools / json
grok-4-fast-reasoningGrok 4 Fast (Reasoning)2M$0.2$0.5tools / thinking / json
grok-4.1Grok 4.1256K$5$25tools / vision / thinking / json
grok-4.2Grok 4.2256K$5$25tools / vision / thinking / json
grok-4.20-0309-non-reasoningGrok 4.20 (Non-Reasoning)256K$3$15tools / vision / json
grok-4.20-0309-reasoningGrok 4.20 (Reasoning)256K$3$15tools / vision / thinking / json
grok-4.20-multi-agent-0309Grok 4.20 Multi-Agent256K$5$25tools / vision / thinking / json
grok-4.3Grok 4.3256K$3$15tools / vision / video_gen / thinking / json
grok-build-0.1Grok Build 0.1256K$3$15tools / thinking / json
grok-composer-2.5-fastGrok Composer 2.5 Fast131K$1$5

Zhipu GLM (5)#

Model IDNameContextInput ($/1M)Output ($/1M)Capabilities
glm-4-7-251222GLM-4.7 (X4)200K$0.7$0.7tools / json / open-source
glm-4.7GLM-4.7200K$0.7$0.7tools / thinking / json / open-source
glm-5GLM-5200K$1.13$3.94tools / thinking / json / open-source
glm-5-turboGLM-5-Turbo200K$1.2$4tools / json / open-source
glm-5.1GLM-5.1200K$1.13$3.94tools / thinking / json / open-source

DeepSeek (6)#

Model IDNameContextInput ($/1M)Output ($/1M)Capabilities
deepseek-v3-0324DeepSeek V3 (0324)128K$0.27 (cached $0.07)$1.1tools / json / open-source
deepseek-v3-2-251201DeepSeek V3.2 (X4)64K$0.27$1.1tools / json / open-source
deepseek-v3.1-terminusDeepSeek V3.1 Terminus128K$0.27 (cached $0.07)$1.1tools / json / open-source
deepseek-v3.2DeepSeek V3.264K$0.27 (cached $0.07)$1.1tools / json / open-source
deepseek-v4-flashDeepSeek V4 Flash128K$0.15$0.6tools / json
deepseek-v4-proDeepSeek V4 Pro128K$0.5$2tools / thinking / json

ByteDance Doubao (19)#

Model IDNameContextInput ($/1M)Output ($/1M)Capabilities
doubao-1-5-lite-32k-250115Doubao 1.5 Lite (32k)32K$0.043$0.085tools / json
doubao-1-5-pro-32k-250115Doubao 1.5 Pro (32k)32K$0.113$0.282tools / json
doubao-1-5-pro-32k-character-250715Doubao 1.5 Pro Character (32k)32K$0.113$0.282tools / json
doubao-1-5-vision-pro-32k-250115Doubao 1.5 Vision Pro (32k)32K$0.423$1.268tools / vision / json
doubao-seed-1-6-250615Doubao Seed 1.6128K$0.113$1.127tools / vision / json
doubao-seed-1-6-251015Doubao Seed 1.6 (10-15)128K$0.113$1.127tools / vision / json
doubao-seed-1-6-flash-250615Doubao Seed 1.6 Flash128K$0.022$0.212tools / vision / json
doubao-seed-1-6-flash-250828Doubao Seed 1.6 Flash (08-28)128K$0.022$0.212tools / vision / json
doubao-seed-1-6-vision-250815Doubao Seed 1.6 Vision128K$0.113$1.127tools / vision / json
doubao-seed-1-8-251228Doubao Seed 1.8128K$0.113$1.127tools / vision / json
doubao-seed-2-0-code-preview-260215Doubao Seed 2.0 Code Preview256K$0.451 (cached $0.09)$2.254tools / vision / thinking / json
doubao-seed-2-0-lite-260215Doubao Seed 2.0 Lite128K$0.085 (cached $0.017)$0.507tools / vision / json
doubao-seed-2-0-lite-260428Doubao Seed 2.0 Lite (04-28)128K$0.085 (cached $0.017)$0.507tools / vision / json
doubao-seed-2-0-mini-260215Doubao Seed 2.0 Mini64K$0.029 (cached $0.006)$0.282tools / vision / json
doubao-seed-2-0-mini-260428Doubao Seed 2.0 Mini (04-28)64K$0.029 (cached $0.006)$0.282tools / vision / json
doubao-seed-2-0-pro-260215Doubao Seed 2.0 Pro256K$0.451 (cached $0.09)$2.254tools / vision / thinking / json
doubao-seed-character-251128Doubao Seed Character32K$0.113$1.127tools / json
doubao-seed-code-preview-251028Doubao Seed Code Preview128K$0.169$1.127tools / vision / json
doubao-seed-translation-250915Doubao Seed Translation32K$0.169$0.507

Tencent Hunyuan (3)#

Model IDNameContextInput ($/1M)Output ($/1M)Capabilities
hunyuan-2.0-instruct-20251111Tencent Hunyuan 2.0 Instruct144K$0.5$2tools / json
hunyuan-2.0-thinking-20251109Tencent Hunyuan 2.0 Thinking192K$0.5$2tools / thinking / json
hunyuan-role-latestTencent Hunyuan Role32K$0.27$1.1tools / json

Moonshot Kimi (2)#

Model IDNameContextInput ($/1M)Output ($/1M)Capabilities
kimi-k2.5Kimi K2.5256K$0.6$2.5tools / thinking / json / open-source
kimi-k2.6Kimi K2.6256K$0.6$2.5tools / thinking / json / open-source

MiniMax (2)#

Model IDNameContextInput ($/1M)Output ($/1M)Capabilities
MiniMax-M2.5MiniMax M2.5200K$0.3$1.5tools / json
MiniMax-M2.7MiniMax M2.7200K$0.3$1.5tools / json

Image generation models (9)#

Billed per image through POST /v1/images/generations. See Image / video APIs.

Model IDNamePriceEndpoint
doubao-seedream-4-0-250828Seedream 4.0$0.029/image/v1/images/generations (size ≥ 960²)
doubao-seedream-4-5-251128Seedream 4.5$0.037/image/v1/images/generations (size ≥ 1920²)
doubao-seedream-5-0-260128Seedream 5.0$0.032/image/v1/images/generations (size ≥ 1920²)
doubao-seedream-5-0-pro-260628Seedream 5.0 Pro$0.044 per image for output ≤ 2.36 MP, $0.088 above, plus $0.003 per input reference image/v1/images/generations (size ≥ 960²)
grok-imagine-imageGrok Imagine (Image)$0.07/image/v1/images/generations
grok-imagine-image-qualityGrok Imagine (Quality)$0.07/image/v1/images/generations
nano-bananaNano Banana (Gemini 2.5 Flash Image)$0.039/image/v1/images/generations
nano-banana-proNano Banana Pro (Gemini 3 Pro Image)$0.134/image/v1/images/generations
nano-banana-2Nano Banana 2 (Gemini 3.1 Flash Image)$0.04/image/v1/images/generations

Since 2026-07-08 the three nano entries are the only way in: the older ids gemini-2.5-flash-image, gemini-3-pro-image-preview, gemini-3.1-flash-image(-preview) and the tvod-nano-* family all keep working as aliases resolving onto the canonical entries above. The former chat / Gemini-native surface for gemini-3.1-flash-image has been retired — call /v1/images/generations instead.

Video generation models (10)#

Billed per second as an async task through POST /v1/videos/generations. See Image / video APIs.

Model IDNamePrice
doubao-seedance-1-0-pro-fast-251015Seedance 1.0 Pro Fast$0.08/second
doubao-seedance-1-0-pro-250528Seedance 1.0 Pro$0.15/second
doubao-seedance-1-5-pro-251215Seedance 1.5 Pro$0.18/second
doubao-seedance-2-0-fast-260128Seedance 2.0 Fast$0.12/second
doubao-seedance-2-0-260128Seedance 2.0$0.22/second
doubao-seedance-2.0Seedance 2.0480p $0.082 / 720p $0.147 / 1080p $0.365 / 2k $0.72 / 4k $0.88 per second
doubao-seedance-2.0-fastSeedance 2.0 Fast480p $0.059 / 720p $0.118 / 1080p $0.29 / 2k $0.35 / 4k $0.42 per second
doubao-seedance-2.0-miniSeedance 2.0 Mini480p $0.037 / 720p $0.074 per second
grok-imagine-videoGrok Imagine Video$0.70/second
grok-imagine-video-1.5-previewGrok Imagine Video 1.5$1.45/second

Picking by job#

JobSuggested modelWhy
Strongest reasoning / agentsclaude-opus-4-8, claude-opus-4-71M context plus tools, vision and thinking
Everyday codingclaude-sonnet-4-6, gpt-5.4, doubao-seed-2-0-code-preview-260215a good price/performance balance
Sub-second responsesgemini-2.5-flash-lite, doubao-seed-1-6-flash-250615low latency, low price
Chinese-language workglm-5.1, deepseek-v3.2, hunyuan-2.0-instruct-20251111domestic vendors
Long contextgemini-2.5-pro (2M), grok-4-fast-* (2M), claude-opus-4-6+ (1M), gpt-5.5 (1M)context window
Cheap at volumedoubao-seed-1-6-flash-250615 ($0.022/M), doubao-1-5-lite-32k-250115 ($0.043/M)around $0.02–$0.05/M
Multimodal visionclaude-*, gpt-5.x, gemini-*, grok-4.x, doubao-seed-1-6-visionvision input supported
Role-playdoubao-seed-character-251128, hunyuan-role-latesttuned for character work
Translationdoubao-seed-translation-250915tuned for translation

Cross-vendor protocol support#

Different models support different client SDKs, and the compatibility matrix is authoritative on what is actually reachable. The catalog's supported_protocols field (returned by GET /api/v1/models/public, the same source as the model plaza cards) reflects only the native / primary path; many models are additionally reachable on other protocols through gateway translation or an upstream channel (the response then carries an X-Protocol-Translation header), and those appear only in the matrix. Native paths as of June 2026:

  • OpenAI Chat (/v1/chat/completions) — 77 of 78 text models (only doubao-seed-translation-250915 is excluded), by far the widest coverage
  • Anthropic Messages (/v1/messages) — natively: all 9 Claude models plus 4 Gemini SKUs (gemini-2.5-flash-lite, gemini-2.5-flash-thinking, gemini-3-flash, gemini-3.1-pro-low). GPT-5.x and most third-party models are also reachable via upstream conversion or translation — see the matrix
  • OpenAI Responses (/v1/responses) — gpt-5.4 / gpt-5.5 (the primary path for Codex CLI) plus claude-opus-4-6, claude-opus-4-7, claude-opus-4-8, claude-sonnet-4-6 and claude-haiku-4-5-20251001
  • Gemini Native (/v1beta/...) — natively: gemini-2.5-flash, gemini-2.5-pro, gemini-3-flash-preview, gemini-3-pro-preview. Some other models are reachable via upstream conversion — see the matrix

For a model × protocol combination with no reachable channel, the gateway returns 503 no_channel_available (the model exists but has no channel on that protocol — not a 404); switch to a supported protocol.

Live data#

The model list and prices track upstream changes. For live data see https://model.zsopc.com/models or the public catalog endpoint (no auth required):

curl https://model.zsopc.com/api/v1/models/public | jq .

The response looks like { "models": [...] }. Note that this endpoint returns every registered entry — including placeholder SKUs that aren't live yet and non-chat SKUs — so filter on the three fields below when consuming it from a script rather than looping over the whole list:

FieldMeaning
callablefalse means a placeholder SKU (listed but not live); calling it returns model_not_found
endpoint_typenull means an ordinary chat model. images_generations, videos_generations, contents_generations_tasks, embeddings and similar mean the model uses its own dedicated endpoint and cannot be sent to /v1/chat/completions — see Image / video / music APIs
supported_protocolsthe client protocols available for this model (see "Cross-vendor protocol support" above)

To take only the models you can send straight to /v1/chat/completions:

curl -s https://model.zsopc.com/api/v1/models/public \
  | jq -r '.models[] | select(.callable and (.supported_protocols | index("openai_chat"))) | .id'