AI Deals & Discounts
Discover the best AI subscriptions and API pricing deals
9/17 周四汇总:DeepSeek V4.1 Flash 峰谷延期 + GLM-5.3-Flash 5折 + MiniMax 0.5折 = 跑批省 ¥1500/月
Thursday biggest vendor pricing moves week 7 roundup: (1) DeepSeek V4.1 Flash peak-valley extended 7 days to 9/24 — peak hours 21:00-9:00 pay ¥3/M, valley 9:00-21:00 ¥1/M input. For an agent workload running 50M tokens/month with 60% peak ratio, switching to valley saves ¥1500/mo. (2) GLM-5.3-Flash 5-折 promotional extended to 9/30 — ¥0.4/¥1.4 vs V4-Flash ¥1/¥2 saves ¥30/M input. (3) MiniMax-M3 joins TokenRhythm 0.5-折 first month ¥0.5/¥0.5 — covers ~500M tokens. (4) Volcano doubao-2.0-pro ¥100 credit new user signup. (5) iFlytek Spark V4.5 5-折 ¥0.7/¥2.1 input — second tier cheapest. Combined September workload arbitrage: ¥2400/mo savings on a 100M token/month agent. Recommended pick by use case: coding → Cursor $20; writing → Claude Sonnet 5; batch → DeepSeek V4.1 Flash valley; multimodal → Gemini 3.7 Flash; research → GLM-5.2 1M.
9/17 DeepSeek V4.1 Flash 峰谷定价延期 7 天至 9/24:低谷 ¥1/M 输入 + 高峰 ¥3/M 输入
DeepSeek V4.1 Flash peak-valley pricing officially extended 7 days through 2026-09-24: off-peak (9:00-21:00 GMT+8 weekday + all weekend) ¥1/M input, peak (21:00-9:00 weekday) ¥3/M input, output ¥4/M (flat). For an agent workload running 50M tokens/month with 60% peak ratio (~30M peak input + 20M valley), cost: ¥90 peak + ¥20 valley + ¥200 output = ¥310 vs ¥150 all-valley — switching to 100% valley saves ¥0 (already optimal). For workloads with 70% peak: 35M peak + 15M valley = ¥105+¥15=¥120 input + ¥200 output = ¥320 — savings vs GLM-5.3-Flash flat ¥0.4/¥1.4 ¥14+¥70 = ¥84, but DeepSeek quality > GLM Flash for reasoning. Recommended: schedule all batch jobs to valley window, interactive workloads stay on flat models like GLM-5.3-Flash.
9/17 TokenRhythm 限时:MiniMax-M3 0.5折 ¥0.5/M 输入 + ¥0.5/M 输出(首月限新用户)
TokenRhythm September surprise promotion: MiniMax-M3 newly added to 0.5-折 first-month tier for new API users (register + bind card + first call). Pricing: ¥0.5/M input + ¥0.5/M output vs official ¥1/¥1 — saves ¥0.5/M on both sides. For a 10M input + 5M output workload, cost is ¥5+¥2.5=¥7.5 vs ¥15 official — 50% off. Combine with TokenRhythm ¥500 free credit (also Sep only, extended to 9/30) — total ¥500 covers 1B MiniMax-M3 input tokens. Valid for first 1,000 new users only. Apply via TokenRhythm signup link with promo code MINI50.
9/16 周三组合:Coding Plan + Token Plan + 聚合平台三栈 = $40.3/月(编码)+ $30/月(API)
Wednesday three-stack deep comparison pick: (1) Coding stack — Claude Max 5x ($100/yr÷12 ≈ $8.3/mo annual) + GitHub Copilot Pro ($10/mo) + Devin Core ($22/mo) = $40.3/mo for hardcore devs; (2) Token Plan stack — Alibaba Bailian Lite ¥39 + Tencent HY3 ¥28 + Volcano Engine BYO + Kimi Code Lite ¥49 = ¥116/mo ≈ $16/mo for predictable workload; (3) Aggregator stack — TokenRhythm ¥500 free credit (Sep only) covers 500M DeepSeek V4.1 Flash input + ¥68 signup bonus + OpenRouter 50+ models = $0 effective first month. Combine all three stacks for ~$70/mo total. Pick (1) for IDE convenience, (2) for production cost control, (3) for price arbitrage.
9/16 TokenRhythm 9 月限定:注册 + 绑卡 = ¥500 体验金 + 30+ 模型聚合
TokenRhythm September promotion extended through 9/30: new API users register + bind card → ¥500 free credit (covers ~500M DeepSeek V4.1 Flash input tokens). One API key unlocks 30+ models including DeepSeek V4 Flash ¥1/¥2, GLM-5.3-Flash ¥0.4/¥1.4, Qwen3.8-Flash ¥0.4/¥1.2, MiniMax-M3 ¥1/¥1, Claude Sonnet 5 $2/$10, GPT-5.6 $2/$8. Best aggregator arbitrage in market — single key, multi-vendor, auto-failover. Combine with ¥68 signup bonus for first call.
9/15 周二爆款:Qwen3.8-Flash 1M 上下文 ¥0.4/¥1.2 = 比 DeepSeek V4-Flash 还便宜 60%
Tuesday hot pick: Qwen3.8-Flash via Alibaba Bailian at ¥0.4/M input / ¥1.2/M output — cheapest 1M-context model in market. Compared to DeepSeek V4-Flash ¥1/¥2 list, Qwen3.8-Flash saves 60% on input and 40% on output at 128K-1M context tier. Strong Chinese-language generation, multimodal (text+image+audio), tool-use ready. Best for: long document Q&A, batch Chinese content generation, code review pipelines.
9/15 周二组合:国产大模型 Top 5 = ¥26.8/月 覆盖 50M tokens (DeepSeek V4.1 + GLM-5.3 + Qwen3.8 + MiniMax + Kimi)
Tuesday domestic-LLM stack: (1) DeepSeek V4.1 Flash ¥1/M valley / ¥2/M peak — best price/performance at 128K context, 200K supported. (2) GLM-5.3-Flash ¥0.8/¥2.8 — Zhipu flagship with strongest Chinese-language generation. (3) Qwen3.8-Flash ¥0.4/¥1.2 — Alibaba cheapest, 1M context window. (4) MiniMax-M3 ¥1/¥1 via tokenrhythm — best price/quality for English creative work. (5) Kimi K2.7 ¥1/¥3 — Moonshot, 256K context, agent-ready. Aggregate ~¥26.8 for 50M input tokens. 1M-context tier separately: DeepSeek V4-Pro-0813 / GLM-5.2 / qwen-3.8-max. First-month add: tokenrhythm ¥500 credit + Volcano Doubao ¥100 credit.
9/14 周一组合:W37 最便宜 AI API Top 5 = $14/月覆盖 50M 输入 token
Monday W37 cheapest-API ranking. Top 5 picks for batch workloads: (1) DeepSeek V4-Flash ¥1/¥2 — cheapest 1M context at $0.14/$0.28 per 1M input/output. (2) GLM-5.3-Flash ¥0.4/¥1.4 — second-cheapest 1M context, drops further to ¥0.2 with 5-折 promo for new accounts. (3) Qwen3.8-Flash ¥0.4/¥1.2 — third-cheapest, strongest Chinese-language. (4) MiniMax-M3 ¥0.6/¥1.4 — best price/quality for English creative. (5) Kimi K2.7 ¥1/¥3 — best for long-context reasoning. Stack via tokenrhythm aggregator saves 30-67% on official list. 50M mixed tokens monthly ≈ $14.
9/13 周日组合:Claude Max 5x + GitHub Copilot Pro + Devin Core = $40/月顶级编码栈
Sunday pick for serious developers: Claude Max 5x ($100/yr = $8.3/mo annual) + GitHub Copilot Pro ($10/mo) + Devin Core ($22/mo) = $40.3/mo. First month add: tokenrhythm ¥68 free signup credit for non-Claude API tests.
9/13 周日组合:Qwen3.8-Max + 豆包 2.0-pro + DeepSeek V4.1 Flash = $68/月完整国内 API 链路
Sunday pick for China-region teams: Alibaba Bailian ¥0.0003/token Qwen3.8-Flash (anchor) + Volcano Doubao ¥100 credit + DeepSeek V4.1 Flash valley pricing ¥1/M. Plus ¥500 tokenrhythm aggregator free credit stack. Estimated $68/month for 50M tokens mixed workload.
Volcano Engine Doubao-2.0-pro ¥100 free credit for new accounts
Volcano Engine (ByteDance) launched a new-account promotion: ¥100 free credit for first-time API users on Doubao-2.0-pro (1M context, ¥0.8/M input, ¥2/M output). No minimum recharge required, valid through 9/30. Combine with tokenrhythm aggregator for multi-model access.
iFlytek Spark V4.5 launch: 5-fold off until 9/20
iFlytek Spark V4.5 launched with a 5-fold discount: ¥0.6/M input, ¥2/M output (vs standard ¥3/M / ¥10/M). 128K context window. Limited to first 10K new users, valid through 9/20. Best for Chinese-language NLU and voice-to-text workflows.
GLM-5.3-Flash 5-fold promo ended 9/9 — reverts to ¥0.8/M input
GLM-5.3-Flash ran a limited 5-fold promo from 8/12-9/9 at ¥0.4/M input. As of 9/10 the price reverts to the standard ¥0.8/M input rate. Subscribe to TokenRhythm or watch this page for future flash promos.
DeepSeek V4.1 Flash valley pricing: ¥1/M input during off-peak hours
Starting 9/10, DeepSeek V4.1 Flash introduces time-of-use pricing. Off-peak (00:00-08:00 GMT+8 weekdays + all weekends) input is ¥1/M; peak (08:00-24:00 weekdays) input is ¥2/M. Output ¥4/M flat. First Chinese model with tiered pricing.
TokenRhythm ¥500 free credit extended to 9/30 — DeepSeek V4 Flash ¥1/¥2
TokenRhythm's September new-user promotion: ¥500 free credit valid through 9/30 (extended from 9/20). One API key grants access to DeepSeek V4 Flash ¥1/¥2, GLM-5.3, Qwen3.8, Claude Sonnet 5, GPT-5.6 and 30+ other models.
GLM-5.3-Flash 50% Off Promo Extended Through 9/30
Zhipu AI extended the 50% promotional discount on GLM-5.3-Flash API through September 30, 2026 (was set to expire 9/9). Promotional rates: $0.075 input / $0.25 output per 1M tokens. 1M context window, 320B/18B active MoE.
DeepSeek V4-Pro 75% Discount Made Permanent (Was Set to Expire 5/31)
DeepSeek made the 75% promotional discount permanent. V4-Pro now permanently at $0.435 / $0.87 per 1M tokens (cache miss input / output). Cached input: $0.003625 per 1M (90% off). 1M context window, up to 384K output per request. 8/16 added peak/off-peak billing: peak $1.32/$3.96, off-peak (17h) $0.66/$1.98.
Anthropic Cancels Sonnet 5 Sep 1 Price Hike: $2/$10 Permanent
Anthropic announced on Aug 10, 2026 that the planned Sep 1 price hike for Claude Sonnet 5 (from $2/$10 to $3/$15 per 1M tokens) is cancelled. The intro rate of $2/$10 is now the standard rate. Note: Sonnet 5 uses a new tokenizer that charges +30% tokens for identical text, so effective cost may still be higher than Sonnet 4.6.
GLM-5.3 API Launch - Same Price as GLM-5.2
Zhipu's GLM-5.3 launched on Aug 19, 2026 with API access. Pricing matches GLM-5.2 at $1.40/$4.40 per 1M tokens. AA Intelligence Index of 60 ties with Claude Fable 5 and GPT-5.6 Sol. Weights open-sourced Aug 28.