Models & Pricing

Every model, its limits, and its exact price
ModelBest forContext windowMax outputInput $/1MOutput $/1M
deepseek-v4-proHardest reasoning tasks1,000,000384,000$0.392$0.783
deepseek-v4-flashFast, cheap default workhorse1,000,000384,000$0.081$0.162
deepseek-v3.2Solid general model128,000128,000$0.186$0.280
kimi-k2.7-codeCoding and agentic coding262,144262,144$0.639$3.150
kimi-k2.6General and long-form262,144262,144$0.581$2.448
gpt-5.6-solFrontier flagship272,000128,000$5.00$30.00
gpt-5.6-terraFrontier, balanced cost272,000128,000$2.50$15.00
gpt-5.6-lunaFrontier, budget272,000128,000$1.00$6.00

Prices are per million tokens, input and output priced separately.

gpt-5.6-sol, gpt-5.6-terra, and gpt-5.6-luna are the only models that serve /v1/responses, in addition to chat completions. The rest serve chat completions only.

Requests that omit the output-limit field get a per-model default: 4,096 tokens for deepseek-v3.2, 8,192 tokens for every other model listed above.

GET /v1/models is the authoritative live list, including current per-token pricing. Prices on this page are current as of 2026-08-03. See Models & Receipts for the endpoint reference.