Discounted AI Models Up to 80% off
Direct official access · curated models · discounted rates
Series currently on discount: gpt · gemini · seedance · glm · doubao
Current discounts
Sorted by discount, largest first, synced with official vendor pricing. Struck-through prices are official list prices; orange is the price you pay. Promotions follow vendor changes and may end at any time.
| Model | Discount | Input | Output | Cached input | Capabilities |
|---|---|---|---|---|---|
openai/gpt-5.6-luna | -80% | $0.2/M$1/M | $1.2/M$6/M | $0.02/M$0.1/M | TextReasoningVision |
google/gemini-3.7-flash | -50% | $0.75/M$1.5/M | $3.75/M$7.5/M | $0.075/M$0.15/M | TextReasoningVision |
google/gemini-3.6-flash | -50% | $0.75/M$1.5/M | $3.75/M$7.5/M | $0.075/M$0.15/M | TextReasoningVision |
OpenAI: GPT-5.6 SolLMArena #10 openai/gpt-5.6-sol | -50% | $2.5/M$5/M | $15/M$30/M | $0.25/M$0.5/M | TextReasoningVision |
bytedance/seedance-2.0-mini | -50% | — | from $0.02/sfrom $0.04/s | — | Video |
Z.ai: GLM-5.2LMArena #35 z-ai/glm-5.2 | -30% | $0.98/M$1.4/M | $3.08/M$4.4/M | $0.182/M$0.26/M | TextReasoningTools |
bytedance/seedance-2.0-fast | -30% | — | from $0.042/sfrom $0.06/s | — | Video |
bytedance/seedance-2.5 | -20% | — | $0.568/s · 1080p$0.71/s · 1080p | — | Video |
openai/gpt-5.6-terra | -20% | $2/M$2.5/M | $12/M$15/M | $0.2/M$0.25/M | TextReasoningVision |
volcengine/doubao-seed-2.1-pro | -20% | $0.7072/M$0.884/M | $3.536/M$4.42/M | $0.1416/M$0.177/M | TextReasoningVision |
volcengine/doubao-seed-2.1-turbo | -20% | $0.3536/M$0.442/M | $1.7696/M$2.212/M | $0.068/M$0.085/M | TextReasoningVision |
bytedance/seedance-2.0 | -10% | — | from $0.063/sfrom $0.07/s | — | Video |
z-ai/glm-5.3 | -10% | $1.26/M$1.4/M | $3.96/M$4.4/M | $0.234/M$0.26/M | TextReasoningTools |
Rankings from LMArena (2026-07-12, CC BY 4.0) lmarena.ai
Browse by use case
The same discounted models, grouped by capability — tool calling, context window, modality.
Coding & agents
Tool calling and reasoning, compatible with Claude Code, Codex and Cline. Discounted rates apply to agent workloads as well.
Set up your coding tool →Long documents & RAG
500K+ context windows at the same discounted rates — whole-repo and long-document workloads.
Compare long-context models →High volume, low cost
The lowest current output prices among the discounted chat models.
Find the cheapest fit →Best value at list price
The lowest standing output prices per 1M tokens — standard pricing, no promotion.
Ready to build at discounted rates ?
3 minutes to integrate — discounts apply automatically.
Get API Key