DeepSeek
DeepSeek-V3.2
FeaturedFlagship non-thinking model. Fast, cheap, and excellent at general chat and tool use.
- Context
- 128K
- Input / 1M
- $0.28
- Output / 1M
- $0.42
deepseek-chatTry itNew Kimi K2 Thinking & Qwen3-Max are live. Browse models →
Model catalog
Every model below is callable through the same OpenAI-compatible API and the same key. Prices are indicative, in USD per 1M tokens — 3 models are free.
Showing 24 of 24 models
DeepSeek
Flagship non-thinking model. Fast, cheap, and excellent at general chat and tool use.
deepseek-chatTry itDeepSeek
Reasoning model with visible chain-of-thought for math, code, and hard problems.
deepseek-reasonerTry itAlibaba Qwen
Alibaba’s largest flagship. Top-tier quality for demanding agentic and chat workloads.
qwen3-maxTry itAlibaba Qwen
Open-weights MoE flagship with hybrid thinking mode.
qwen3-235b-a22bTry itAlibaba Qwen
Dense open-weights all-rounder with thinking and non-thinking modes.
qwen3-32bTry itAlibaba Qwen
Code-specialized model built for agentic coding tools and repository-scale edits.
qwen3-coder-plusTry itAlibaba Qwen
Vision-language model for image understanding, OCR, and chart reasoning.
qwen-vl-maxTry itAlibaba Qwen
Multilingual embedding model for search and RAG pipelines.
text-embedding-v4Try itZhipu GLM
Z.ai flagship. Strong coding, writing, and agentic planning; 200K context.
glm-4.6Try itZhipu GLM
Lighter GLM tuned for cost-sensitive production traffic.
glm-4.5-airTry itZhipu GLM
Free tier of the GLM family — great for prototypes and hobby projects.
glm-4.5-flashTry itMoonshot Kimi
Trillion-parameter MoE tuned for agentic tool use and long documents.
kimi-k2-0905Try itMoonshot Kimi
Reasoning variant of K2 with interleaved thinking for multi-step tasks.
kimi-k2-thinkingTry itByteDance Doubao
Volcano Engine’s value flagship with controllable thinking depth.
doubao-seed-1-6Try itByteDance Doubao
Ultra-cheap fast tier — near-free for high-volume workloads.
doubao-seed-1-6-flashTry itByteDance Doubao
Multimodal Seed model for image understanding at Flash prices.
doubao-seed-1-6-visionTry itMiniMax
Efficient agentic MoE — a favorite for coding agents on a budget.
MiniMax-M2Try itBaidu ERNIE
Hybrid-thinking multimodal model at aggressive prices.
ernie-4.5-turboTry itBaidu ERNIE
Baidu’s reasoning flagship for planning and tool orchestration.
ernie-4.5-x1Try itTencent Hunyuan
Tencent’s hybrid fast model — switches thinking on only when needed.
hunyuan-turbosTry itTencent Hunyuan
Deep-thinking model with long chain-of-thought for hard reasoning.
hunyuan-t1Try itiFlytek Spark
iFlytek’s reasoning model with strong multilingual understanding.
spark-x1Try itiFlytek Spark
Free lightweight tier for chat and simple tools.
spark-liteTry itStepFun
Multimodal flagship from StepFun with interleaved thinking.
step-3Try itPrices are indicative list prices in USD per 1M tokens and may differ from the rates configured in your console. Context windows are the maximum supported by each upstream model.