ByteDance Seedream 4.0
ByteDance's mainline text-to-image model — reliable prompt following and clean text rendering, billed per image.
seedream-4-0-t2iBrowse chat, image and video models from domestic labs — compare prices and pick an ID to call.
ByteDance Seedream 4.0
ByteDance's mainline text-to-image model — reliable prompt following and clean text rendering, billed per image.
seedream-4-0-t2iByteDance Seedream 4.5
An upgraded Seedream tier with finer detail and stronger layout control than 4.0.
seedream-4-5-t2iKimi K3
Moonshot AI's flagship — reasoning with native vision understanding and a 1M-token context window.
kimi-k3Qwen3 235B A22B Instruct 2507
A 235B MoE (22B active) instruct model from the July 2025 Qwen3 refresh.
qwen3-235b-a22b-instruct-2507Qwen3 Coder 480B A35B Instruct
Qwen's largest coding model — a 480B MoE (35B active) built for repo-scale codegen and agents.
qwen3-coder-480b-a35b-instructQwen3.5 397B A17B
The big open Qwen3.5 MoE — 397B parameters total, 17B active per token.
qwen3-5-397b-a17bQwen3.5 Plus
Alibaba's hosted mainline Qwen3.5 tier — balanced quality, speed and price.
qwen3-5-plusByteDance Seedream 3.0
The earlier Seedream generation — a budget text-to-image option that still renders text well.
seedream-3-0-t2iByteDance Seedream 5.0 Lite
The light tier of the newest Seedream line — fast, low-cost images for high-volume use.
seedream-5-0-lite-t2iDeepSeek Prover V2 671B
A 671B model specialized in formal theorem proving rather than general chat.
deepseek-prover-v2-671bDeepSeek R1 (0528)
The 0528 refresh of DeepSeek's R1 reasoning model — long chain-of-thought answers for hard problems.
deepseek-r1-0528DeepSeek R1 Distill Llama 70B
R1 reasoning distilled into a 70B Llama base — much cheaper reasoning at some quality cost.
deepseek-r1-distill-llama-70bDeepSeek R1 Turbo
An R1 variant on faster serving, tuned for lower-latency reasoning workloads.
deepseek-r1-turboDeepSeek V3 (0324)
The 0324 snapshot of DeepSeek V3, the general-purpose MoE chat model.
deepseek-v3-0324DeepSeek V3 Turbo
DeepSeek V3 on faster serving — lower latency for everyday chat and tool use.
deepseek-v3-turboDeepSeek V3.1
V3.1 merges V3's chat strengths with R1-style thinking in one hybrid model.
deepseek-v3-1DeepSeek V3.1 Terminus
The final V3.1 build, with tool-use and agent fixes, before the V3.2 line.
deepseek-v3-1-terminusDeepSeek V3.2
DeepSeek V3.2 — sparse attention makes long-context inference markedly cheaper.
deepseek-v3-2DeepSeek V3.2 Exp
The experimental V3.2 build that introduced DeepSeek Sparse Attention, priced aggressively low.
deepseek-v3-2-expDeepSeek V4 Flash
The fast, low-cost tier of the V4 line — the default pick for high-volume chat.
deepseek-v4-flashDeepSeek V4 Pro
DeepSeek's current flagship — the strongest general model of the V4 line.
deepseek-v4-proDoubao-Seedance 2.0 (Image-to-Video)
Animates a still frame into a short clip with ByteDance's Seedance 2.0.
seedance-2-i2v-amuxDoubao-Seedance 2.0 (Text-to-Video)
Text-to-video with Seedance 2.0 — short clips straight from a prompt.
seedance-2-t2v-amuxDoubao-Seedance 2.0 Fast (Image-to-Video)
The fast Seedance tier — quicker, cheaper image-to-video for drafts and volume work.
seedance-2-fast-i2v-amuxGLM-4.5
Zhipu's agent-focused flagship of the 4.5 generation — chat, reasoning and tool use in one model.
glm-4-5GLM-4.5-Air
The light GLM-4.5 tier — most of the capability at a fraction of the price.
glm-4-5-airGLM-4.5V
The vision variant of GLM-4.5 — screenshots, documents and visual grounding.
glm-4-5vGLM-4.6
GLM-4.6 extends 4.5 with a longer context window and stronger coding.
glm-4-6GLM-4.6V
The vision build of the GLM-4.6 generation.
glm-4-6vGLM-4.7
The last 4.x generation before GLM-5 — a further coding and agent refinement.
glm-4-7GLM-5
Zhipu's fifth-generation flagship — a step up in reasoning and agentic coding over 4.x.
glm-5GLM-5.1
A GLM-5 line refresh — incremental quality gains over the base model.
glm-5-1GLM-5.2
The current head of the GLM-5 line.
glm-5-2Kimi K2 0905
The September 2025 snapshot of K2, Moonshot's trillion-parameter open MoE for agentic work.
kimi-k2-0905Kimi K2 Instruct
The instruction-tuned K2 base build — general chat and tool calling.
kimi-k2-instructKimi K2 Thinking
K2 with explicit long-horizon thinking — strong on multi-step agent and search tasks.
kimi-k2-thinkingKimi K2.5
A K2-line refresh that sharpens agentic tool use over the original K2.
kimi-k2-5Kimi K2.6
The latest K2-line refresh before K3 — incremental gains over K2.5.
kimi-k2-6MiniMax M1 80k
MiniMax's earlier reasoning MoE, with an 80k-token thinking budget.
minimax-m1-80kMiniMax M2
A compact, fast MoE built for coding and agent loops — MiniMax's price-performance play.
minimax-m2MiniMax M2.1
An M2-line refresh — better coding and agent reliability over M2.
minimax-m2-1MiniMax M2.5
A further M2-line upgrade, balancing speed against quality.
minimax-m2-5MiniMax M2.7
The newest M2-line model — MiniMax's current mainline chat tier.
minimax-m2-7MiniMax-H3 (Image-to-Video)
Animate a still first frame into a 2K clip with synchronized audio — MiniMax's flagship image-to-video.
minimax-h3-i2vMiniMax-H3 (Reference-to-Video)
Keep a subject identity consistent across a 2K clip from one reference image.
minimax-h3-r2vMiniMax-H3 (Text-to-Video)
Flagship text-to-video with synchronized audio, 2K output, and 4–15s clips.
minimax-h3-t2vQwen2.5 72B Instruct
The dense 72B workhorse of the earlier Qwen2.5 generation — stable and widely deployed.
qwen2-5-72b-instructQwen3 Coder 30B A3B Instruct
A light coding MoE (30B total, 3B active) for cheap autocomplete and small agents.
qwen3-coder-30b-a3b-instructQwen3 Coder Next
The rolling "next" build of Qwen's coder line — the newest coding capabilities as they land.
qwen3-coder-nextQwen3 VL 30B A3B Instruct
A compact vision-language MoE that reads images, UI screenshots and documents.
qwen3-vl-30b-a3b-instructQwen3.5 122B A10B
A mid-size Qwen3.5 MoE (122B total, 10B active) — near-flagship quality at lower cost.
qwen3-5-122b-a10bQwen3.5 27B
A dense 27B Qwen3.5 — small enough to be cheap, big enough to be useful.
qwen3-5-27bQwen3.5 35B A3B
The efficiency tier of Qwen3.5 — a 35B MoE with 3B active, made for high-throughput serving.
qwen3-5-35b-a3bVidu Q3 Pro (Image-to-Video)
Image-to-video on Vidu's Q3 Pro tier — animate one still into a clip.
vidu-q3-pro-i2v-novitaVidu Q3 Pro (Start-End-to-Video)
Give a start and an end frame and Q3 Pro interpolates the motion between them.
vidu-q3-pro-f2v-novita