Skip to content
Catalog

Model market

Browse chat, image and video models from domestic labs — compare prices and pick an ID to call.

55 modelsPrices in CNY (¥).
ImageSave 5%

ByteDance Seedream 4.0

ByteDance's mainline text-to-image model — reliable prompt following and clean text rendering, billed per image.

seedream-4-0-t2i
ByteDanceText to image
¥0.199¥0.210/run
ImageSave 5%

ByteDance Seedream 4.5

An upgraded Seedream tier with finer detail and stronger layout control than 4.0.

seedream-4-5-t2i
ByteDanceText to image
¥0.237¥0.250/run
ChatSave 5%

Kimi K3

Moonshot AI's flagship — reasoning with native vision understanding and a 1M-token context window.

kimi-k3
Moonshot AIChat
In ¥19.00Out ¥95.00¥100.00/M tokens
ChatSave 15%

Qwen3 235B A22B Instruct 2507

A 235B MoE (22B active) instruct model from the July 2025 Qwen3 refresh.

qwen3-235b-a22b-instruct-2507
AlibabaChat
In ¥1.23Out ¥4.93¥5.80/M tokens
ChatSave 15%

Qwen3 Coder 480B A35B Instruct

Qwen's largest coding model — a 480B MoE (35B active) built for repo-scale codegen and agents.

qwen3-coder-480b-a35b-instruct
AlibabaChat
In ¥3.40Out ¥13.60¥16.00/M tokens
ChatSave 15%

Qwen3.5 397B A17B

The big open Qwen3.5 MoE — 397B parameters total, 17B active per token.

qwen3-5-397b-a17b
AlibabaChat
In ¥1.02Out ¥6.12¥7.20/M tokens
ChatSave 15%

Qwen3.5 Plus

Alibaba's hosted mainline Qwen3.5 tier — balanced quality, speed and price.

qwen3-5-plus
AlibabaChat
In ¥0.680Out ¥4.08¥4.80/M tokens
ImageSave 5%

ByteDance Seedream 3.0

The earlier Seedream generation — a budget text-to-image option that still renders text well.

seedream-3-0-t2i
ByteDanceText to image
¥0.028¥0.030/run
ImageSave 5%

ByteDance Seedream 5.0 Lite

The light tier of the newest Seedream line — fast, low-cost images for high-volume use.

seedream-5-0-lite-t2i
ByteDanceText to image
¥0.233¥0.245/run
ChatSave 5%

DeepSeek Prover V2 671B

A 671B model specialized in formal theorem proving rather than general chat.

deepseek-prover-v2-671b
DeepSeekChat
In ¥0.665Out ¥2.38¥2.50/M tokens
ChatSave 15%

DeepSeek R1 (0528)

The 0528 refresh of DeepSeek's R1 reasoning model — long chain-of-thought answers for hard problems.

deepseek-r1-0528
DeepSeekChat
In ¥0.340Out ¥0.552¥0.650/M tokens
ChatSave 5%

DeepSeek R1 Distill Llama 70B

R1 reasoning distilled into a 70B Llama base — much cheaper reasoning at some quality cost.

deepseek-r1-distill-llama-70b
DeepSeekChat
In ¥0.760Out ¥0.760¥0.800/M tokens
ChatSave 5%

DeepSeek R1 Turbo

An R1 variant on faster serving, tuned for lower-latency reasoning workloads.

deepseek-r1-turbo
DeepSeekChat
In ¥0.665Out ¥2.38¥2.50/M tokens
ChatSave 15%

DeepSeek V3 (0324)

The 0324 snapshot of DeepSeek V3, the general-purpose MoE chat model.

deepseek-v3-0324
DeepSeekChat
In ¥1.70Out ¥6.80¥8.00/M tokens
ChatSave 15%

DeepSeek V3 Turbo

DeepSeek V3 on faster serving — lower latency for everyday chat and tool use.

deepseek-v3-turbo
DeepSeekChat
In ¥1.70Out ¥6.80¥8.00/M tokens
ChatSave 5%

DeepSeek V3.1

V3.1 merges V3's chat strengths with R1-style thinking in one hybrid model.

deepseek-v3-1
DeepSeekChat
In ¥0.257Out ¥1.04¥1.10/M tokens
ChatSave 5%

DeepSeek V3.1 Terminus

The final V3.1 build, with tool-use and agent fixes, before the V3.2 line.

deepseek-v3-1-terminus
DeepSeekChat
In ¥0.257Out ¥1.04¥1.10/M tokens
ChatSave 15%

DeepSeek V3.2

DeepSeek V3.2 — sparse attention makes long-context inference markedly cheaper.

deepseek-v3-2
DeepSeekChat
In ¥1.70Out ¥2.55¥3.00/M tokens
ChatSave 15%

DeepSeek V3.2 Exp

The experimental V3.2 build that introduced DeepSeek Sparse Attention, priced aggressively low.

deepseek-v3-2-exp
DeepSeekChat
In ¥1.70Out ¥2.55¥3.00/M tokens
Chat

DeepSeek V4 Flash

The fast, low-cost tier of the V4 line — the default pick for high-volume chat.

deepseek-v4-flash
DeepSeekChat
In ¥1.00Out ¥2.00/M tokens
Chat

DeepSeek V4 Pro

DeepSeek's current flagship — the strongest general model of the V4 line.

deepseek-v4-pro
DeepSeekChat
In ¥3.00Out ¥6.00/M tokens
VideoSave 15%

Doubao-Seedance 2.0 (Image-to-Video)

Animates a still frame into a short clip with ByteDance's Seedance 2.0.

seedance-2-i2v-amux
ByteDanceImage to video
¥2.55¥3.00/sec
VideoSave 15%

Doubao-Seedance 2.0 (Text-to-Video)

Text-to-video with Seedance 2.0 — short clips straight from a prompt.

seedance-2-t2v-amux
ByteDanceText to video
¥2.55¥3.00/sec
VideoSave 15%

Doubao-Seedance 2.0 Fast (Image-to-Video)

The fast Seedance tier — quicker, cheaper image-to-video for drafts and volume work.

seedance-2-fast-i2v-amux
ByteDanceImage to video
¥1.19¥1.40/sec
ChatSave 5%

GLM-4.5

Zhipu's agent-focused flagship of the 4.5 generation — chat, reasoning and tool use in one model.

glm-4-5
Z.AIChat
In ¥3.80Out ¥15.20¥16.00/M tokens
ChatSave 5%

GLM-4.5-Air

The light GLM-4.5 tier — most of the capability at a fraction of the price.

glm-4-5-air
Z.AIChat
In ¥1.14Out ¥1.90¥2.00/M tokens
ChatSave 5%

GLM-4.5V

The vision variant of GLM-4.5 — screenshots, documents and visual grounding.

glm-4-5v
Z.AIChat
In ¥3.80Out ¥11.40¥12.00/M tokens
ChatSave 15%

GLM-4.6

GLM-4.6 extends 4.5 with a longer context window and stronger coding.

glm-4-6
Z.AIChat
In ¥3.40Out ¥13.60¥16.00/M tokens
ChatSave 15%

GLM-4.6V

The vision build of the GLM-4.6 generation.

glm-4-6v
Z.AIChat
In ¥1.70Out ¥5.10¥6.00/M tokens
ChatSave 15%

GLM-4.7

The last 4.x generation before GLM-5 — a further coding and agent refinement.

glm-4-7
Z.AIChat
In ¥1.70Out ¥6.80¥8.00/M tokens
ChatSave 10%

GLM-5

Zhipu's fifth-generation flagship — a step up in reasoning and agentic coding over 4.x.

glm-5
Z.AIChat
In ¥3.60Out ¥16.20¥18.00/M tokens
ChatSave 10%

GLM-5.1

A GLM-5 line refresh — incremental quality gains over the base model.

glm-5-1
Z.AIChat
In ¥5.40Out ¥21.60¥24.00/M tokens
ChatSave 5%

GLM-5.2

The current head of the GLM-5 line.

glm-5-2
Z.AIChat
In ¥1.33Out ¥4.18¥4.40/M tokens
ChatSave 15%

Kimi K2 0905

The September 2025 snapshot of K2, Moonshot's trillion-parameter open MoE for agentic work.

kimi-k2-0905
Moonshot AIChat
In ¥3.40Out ¥13.60¥16.00/M tokens
ChatSave 15%

Kimi K2 Instruct

The instruction-tuned K2 base build — general chat and tool calling.

kimi-k2-instruct
Moonshot AIChat
In ¥3.40Out ¥13.60¥16.00/M tokens
ChatSave 15%

Kimi K2 Thinking

K2 with explicit long-horizon thinking — strong on multi-step agent and search tasks.

kimi-k2-thinking
Moonshot AIChat
In ¥3.40Out ¥13.60¥16.00/M tokens
ChatSave 5%

Kimi K2.5

A K2-line refresh that sharpens agentic tool use over the original K2.

kimi-k2-5
Moonshot AIChat
In ¥3.80Out ¥19.95¥21.00/M tokens
ChatSave 5%

Kimi K2.6

The latest K2-line refresh before K3 — incremental gains over K2.5.

kimi-k2-6
Moonshot AIChat
In ¥6.17Out ¥25.65¥27.00/M tokens
ChatSave 15%

MiniMax M1 80k

MiniMax's earlier reasoning MoE, with an 80k-token thinking budget.

minimax-m1-80k
MiniMaxChat
In ¥3.40Out ¥13.60¥16.00/M tokens
ChatSave 15%

MiniMax M2

A compact, fast MoE built for coding and agent loops — MiniMax's price-performance play.

minimax-m2
MiniMaxChat
In ¥1.78Out ¥7.14¥8.40/M tokens
Chat

MiniMax M2.1

An M2-line refresh — better coding and agent reliability over M2.

minimax-m2-1
MiniMaxChat
In ¥2.10Out ¥8.40/M tokens
ChatSave 10%

MiniMax M2.5

A further M2-line upgrade, balancing speed against quality.

minimax-m2-5
MiniMaxChat
In ¥1.89Out ¥7.56¥8.40/M tokens
ChatSave 10%

MiniMax M2.7

The newest M2-line model — MiniMax's current mainline chat tier.

minimax-m2-7
MiniMaxChat
In ¥1.89Out ¥7.56¥8.40/M tokens
Video

MiniMax-H3 (Image-to-Video)

Animate a still first frame into a 2K clip with synchronized audio — MiniMax's flagship image-to-video.

minimax-h3-i2v
MiniMaxImage to video
Video

MiniMax-H3 (Reference-to-Video)

Keep a subject identity consistent across a 2K clip from one reference image.

minimax-h3-r2v
MiniMaxReference to video
Video

MiniMax-H3 (Text-to-Video)

Flagship text-to-video with synchronized audio, 2K output, and 4–15s clips.

minimax-h3-t2v
MiniMaxText to video
ChatSave 5%

Qwen2.5 72B Instruct

The dense 72B workhorse of the earlier Qwen2.5 generation — stable and widely deployed.

qwen2-5-72b-instruct
AlibabaChat
In ¥2.61Out ¥2.74¥2.88/M tokens
ChatSave 15%

Qwen3 Coder 30B A3B Instruct

A light coding MoE (30B total, 3B active) for cheap autocomplete and small agents.

qwen3-coder-30b-a3b-instruct
AlibabaChat
In ¥1.91Out ¥7.65¥9.00/M tokens
ChatSave 15%

Qwen3 Coder Next

The rolling "next" build of Qwen's coder line — the newest coding capabilities as they land.

qwen3-coder-next
AlibabaChat
In ¥1.19Out ¥8.92¥10.50/M tokens
ChatSave 15%

Qwen3 VL 30B A3B Instruct

A compact vision-language MoE that reads images, UI screenshots and documents.

qwen3-vl-30b-a3b-instruct
AlibabaChat
In ¥0.637Out ¥2.55¥3.00/M tokens
ChatSave 15%

Qwen3.5 122B A10B

A mid-size Qwen3.5 MoE (122B total, 10B active) — near-flagship quality at lower cost.

qwen3-5-122b-a10b
AlibabaChat
In ¥0.680Out ¥5.44¥6.40/M tokens
ChatSave 15%

Qwen3.5 27B

A dense 27B Qwen3.5 — small enough to be cheap, big enough to be useful.

qwen3-5-27b
AlibabaChat
In ¥0.510Out ¥4.08¥4.80/M tokens
ChatSave 15%

Qwen3.5 35B A3B

The efficiency tier of Qwen3.5 — a 35B MoE with 3B active, made for high-throughput serving.

qwen3-5-35b-a3b
AlibabaChat
In ¥0.340Out ¥2.72¥3.20/M tokens
Video

Vidu Q3 Pro (Image-to-Video)

Image-to-video on Vidu's Q3 Pro tier — animate one still into a clip.

vidu-q3-pro-i2v-novita
ViduImage to video
Video

Vidu Q3 Pro (Start-End-to-Video)

Give a start and an end frame and Q3 Pro interpolates the motion between them.

vidu-q3-pro-f2v-novita
ViduImage to video