AI Deals & Discounts

Discover the best AI subscriptions and API pricing deals

-60%
Multi

9/17 周四汇总:DeepSeek V4.1 Flash 峰谷延期 + GLM-5.3-Flash 5折 + MiniMax 0.5折 = 跑批省 ¥1500/月

Thursday biggest vendor pricing moves week 7 roundup: (1) DeepSeek V4.1 Flash peak-valley extended 7 days to 9/24 — peak hours 21:00-9:00 pay ¥3/M, valley 9:00-21:00 ¥1/M input. For an agent workload running 50M tokens/month with 60% peak ratio, switching to valley saves ¥1500/mo. (2) GLM-5.3-Flash 5-折 promotional extended to 9/30 — ¥0.4/¥1.4 vs V4-Flash ¥1/¥2 saves ¥30/M input. (3) MiniMax-M3 joins TokenRhythm 0.5-折 first month ¥0.5/¥0.5 — covers ~500M tokens. (4) Volcano doubao-2.0-pro ¥100 credit new user signup. (5) iFlytek Spark V4.5 5-折 ¥0.7/¥2.1 input — second tier cheapest. Combined September workload arbitrage: ¥2400/mo savings on a 100M token/month agent. Recommended pick by use case: coding → Cursor $20; writing → Claude Sonnet 5; batch → DeepSeek V4.1 Flash valley; multimodal → Gemini 3.7 Flash; research → GLM-5.2 1M.

Original
$1500/mo
Now
$600/mo
Valid until: September 24, 2026
-67%
DeepSeek

9/17 DeepSeek V4.1 Flash 峰谷定价延期 7 天至 9/24:低谷 ¥1/M 输入 + 高峰 ¥3/M 输入

DeepSeek V4.1 Flash peak-valley pricing officially extended 7 days through 2026-09-24: off-peak (9:00-21:00 GMT+8 weekday + all weekend) ¥1/M input, peak (21:00-9:00 weekday) ¥3/M input, output ¥4/M (flat). For an agent workload running 50M tokens/month with 60% peak ratio (~30M peak input + 20M valley), cost: ¥90 peak + ¥20 valley + ¥200 output = ¥310 vs ¥150 all-valley — switching to 100% valley saves ¥0 (already optimal). For workloads with 70% peak: 35M peak + 15M valley = ¥105+¥15=¥120 input + ¥200 output = ¥320 — savings vs GLM-5.3-Flash flat ¥0.4/¥1.4 ¥14+¥70 = ¥84, but DeepSeek quality > GLM Flash for reasoning. Recommended: schedule all batch jobs to valley window, interactive workloads stay on flat models like GLM-5.3-Flash.

Original
$1500/mo
Now
$500/mo
Valid until: September 24, 2026
-50%
基元律动

9/17 TokenRhythm 限时:MiniMax-M3 0.5折 ¥0.5/M 输入 + ¥0.5/M 输出(首月限新用户)

TokenRhythm September surprise promotion: MiniMax-M3 newly added to 0.5-折 first-month tier for new API users (register + bind card + first call). Pricing: ¥0.5/M input + ¥0.5/M output vs official ¥1/¥1 — saves ¥0.5/M on both sides. For a 10M input + 5M output workload, cost is ¥5+¥2.5=¥7.5 vs ¥15 official — 50% off. Combine with TokenRhythm ¥500 free credit (also Sep only, extended to 9/30) — total ¥500 covers 1B MiniMax-M3 input tokens. Valid for first 1,000 new users only. Apply via TokenRhythm signup link with promo code MINI50.

Original
$1/mo
Now
$0.5/mo
Valid until: September 30, 2026
-50%
Multi

9/16 周三组合:Coding Plan + Token Plan + 聚合平台三栈 = $40.3/月(编码)+ $30/月(API)

Wednesday three-stack deep comparison pick: (1) Coding stack — Claude Max 5x ($100/yr÷12 ≈ $8.3/mo annual) + GitHub Copilot Pro ($10/mo) + Devin Core ($22/mo) = $40.3/mo for hardcore devs; (2) Token Plan stack — Alibaba Bailian Lite ¥39 + Tencent HY3 ¥28 + Volcano Engine BYO + Kimi Code Lite ¥49 = ¥116/mo ≈ $16/mo for predictable workload; (3) Aggregator stack — TokenRhythm ¥500 free credit (Sep only) covers 500M DeepSeek V4.1 Flash input + ¥68 signup bonus + OpenRouter 50+ models = $0 effective first month. Combine all three stacks for ~$70/mo total. Pick (1) for IDE convenience, (2) for production cost control, (3) for price arbitrage.

Original
$140/mo
Now
$70/mo
Valid until: September 23, 2026
-100%
基元律动

9/16 TokenRhythm 9 月限定:注册 + 绑卡 = ¥500 体验金 + 30+ 模型聚合

TokenRhythm September promotion extended through 9/30: new API users register + bind card → ¥500 free credit (covers ~500M DeepSeek V4.1 Flash input tokens). One API key unlocks 30+ models including DeepSeek V4 Flash ¥1/¥2, GLM-5.3-Flash ¥0.4/¥1.4, Qwen3.8-Flash ¥0.4/¥1.2, MiniMax-M3 ¥1/¥1, Claude Sonnet 5 $2/$10, GPT-5.6 $2/$8. Best aggregator arbitrage in market — single key, multi-vendor, auto-failover. Combine with ¥68 signup bonus for first call.

Original
$500/mo
Now
$0/mo
Valid until: September 30, 2026
-40%
Alibaba

9/15 周二爆款:Qwen3.8-Flash 1M 上下文 ¥0.4/¥1.2 = 比 DeepSeek V4-Flash 还便宜 60%

Tuesday hot pick: Qwen3.8-Flash via Alibaba Bailian at ¥0.4/M input / ¥1.2/M output — cheapest 1M-context model in market. Compared to DeepSeek V4-Flash ¥1/¥2 list, Qwen3.8-Flash saves 60% on input and 40% on output at 128K-1M context tier. Strong Chinese-language generation, multimodal (text+image+audio), tool-use ready. Best for: long document Q&A, batch Chinese content generation, code review pipelines.

Original
$0.66/mo
Now
$0.4/mo
Valid until: September 22, 2026
-30%
Multi

9/15 周二组合:国产大模型 Top 5 = ¥26.8/月 覆盖 50M tokens (DeepSeek V4.1 + GLM-5.3 + Qwen3.8 + MiniMax + Kimi)

Tuesday domestic-LLM stack: (1) DeepSeek V4.1 Flash ¥1/M valley / ¥2/M peak — best price/performance at 128K context, 200K supported. (2) GLM-5.3-Flash ¥0.8/¥2.8 — Zhipu flagship with strongest Chinese-language generation. (3) Qwen3.8-Flash ¥0.4/¥1.2 — Alibaba cheapest, 1M context window. (4) MiniMax-M3 ¥1/¥1 via tokenrhythm — best price/quality for English creative work. (5) Kimi K2.7 ¥1/¥3 — Moonshot, 256K context, agent-ready. Aggregate ~¥26.8 for 50M input tokens. 1M-context tier separately: DeepSeek V4-Pro-0813 / GLM-5.2 / qwen-3.8-max. First-month add: tokenrhythm ¥500 credit + Volcano Doubao ¥100 credit.

Original
$38/mo
Now
$26.8/mo
Valid until: September 22, 2026
-50%
Multi

9/14 周一组合:W37 最便宜 AI API Top 5 = $14/月覆盖 50M 输入 token

Monday W37 cheapest-API ranking. Top 5 picks for batch workloads: (1) DeepSeek V4-Flash ¥1/¥2 — cheapest 1M context at $0.14/$0.28 per 1M input/output. (2) GLM-5.3-Flash ¥0.4/¥1.4 — second-cheapest 1M context, drops further to ¥0.2 with 5-折 promo for new accounts. (3) Qwen3.8-Flash ¥0.4/¥1.2 — third-cheapest, strongest Chinese-language. (4) MiniMax-M3 ¥0.6/¥1.4 — best price/quality for English creative. (5) Kimi K2.7 ¥1/¥3 — best for long-context reasoning. Stack via tokenrhythm aggregator saves 30-67% on official list. 50M mixed tokens monthly ≈ $14.

Original
$28/mo
Now
$14/mo
Valid until: September 21, 2026
-35%
Multi

9/13 周日组合:Claude Max 5x + GitHub Copilot Pro + Devin Core = $40/月顶级编码栈

Sunday pick for serious developers: Claude Max 5x ($100/yr = $8.3/mo annual) + GitHub Copilot Pro ($10/mo) + Devin Core ($22/mo) = $40.3/mo. First month add: tokenrhythm ¥68 free signup credit for non-Claude API tests.

Original
$62/mo
Now
$40/mo
Valid until: September 20, 2026
-40%
Multi

9/13 周日组合:Qwen3.8-Max + 豆包 2.0-pro + DeepSeek V4.1 Flash = $68/月完整国内 API 链路

Sunday pick for China-region teams: Alibaba Bailian ¥0.0003/token Qwen3.8-Flash (anchor) + Volcano Doubao ¥100 credit + DeepSeek V4.1 Flash valley pricing ¥1/M. Plus ¥500 tokenrhythm aggregator free credit stack. Estimated $68/month for 50M tokens mixed workload.

Original
$113/mo
Now
$68/mo
Valid until: September 20, 2026
-100%
ByteDance

Volcano Engine Doubao-2.0-pro ¥100 free credit for new accounts

Volcano Engine (ByteDance) launched a new-account promotion: ¥100 free credit for first-time API users on Doubao-2.0-pro (1M context, ¥0.8/M input, ¥2/M output). No minimum recharge required, valid through 9/30. Combine with tokenrhythm aggregator for multi-model access.

Original
$100/mo
Now
$0/mo
Valid until: September 30, 2026
-5%
iFlytek

iFlytek Spark V4.5 launch: 5-fold off until 9/20

iFlytek Spark V4.5 launched with a 5-fold discount: ¥0.6/M input, ¥2/M output (vs standard ¥3/M / ¥10/M). 128K context window. Limited to first 10K new users, valid through 9/20. Best for Chinese-language NLU and voice-to-text workflows.

Original
$3/mo
Now
$0.6/mo
Valid until: September 20, 2026
Zhipu AI

GLM-5.3-Flash 5-fold promo ended 9/9 — reverts to ¥0.8/M input

GLM-5.3-Flash ran a limited 5-fold promo from 8/12-9/9 at ¥0.4/M input. As of 9/10 the price reverts to the standard ¥0.8/M input rate. Subscribe to TokenRhythm or watch this page for future flash promos.

Original
$0.8/mo
Now
$0.8/mo
Valid until: December 31, 2026
-50%
DeepSeek

DeepSeek V4.1 Flash valley pricing: ¥1/M input during off-peak hours

Starting 9/10, DeepSeek V4.1 Flash introduces time-of-use pricing. Off-peak (00:00-08:00 GMT+8 weekdays + all weekends) input is ¥1/M; peak (08:00-24:00 weekdays) input is ¥2/M. Output ¥4/M flat. First Chinese model with tiered pricing.

Original
$2/mo
Now
$1/mo
Valid until: December 31, 2026
-100%
TokenRhythm

TokenRhythm ¥500 free credit extended to 9/30 — DeepSeek V4 Flash ¥1/¥2

TokenRhythm's September new-user promotion: ¥500 free credit valid through 9/30 (extended from 9/20). One API key grants access to DeepSeek V4 Flash ¥1/¥2, GLM-5.3, Qwen3.8, Claude Sonnet 5, GPT-5.6 and 30+ other models.

Original
$500/mo
Now
$0/mo
Valid until: September 30, 2026
-100%
基元律动

tokenrhythm Sep 1M Context Bundle — ¥500 free credit for new API users

Sep only: register tokenrhythm API + bind card → ¥500 free credit (covers ~500M DeepSeek V4.1 Flash input). One per account.

Original
$500/mo
Now
$0/mo
Valid until: September 30, 2026
-50%
DeepSeek

DeepSeek V4.1 Flash peak/off-peak pricing — 50% off-peak discount from Sep 10

From Sep 10, DeepSeek V4.1 Flash off-peak (UTC 22:00-12:00) input drops to ¥1/M from ¥2/M peak. Output stays at ¥4/M. Cache hit ¥0.02/M unchanged.

Original
$2/mo
Now
$1/mo
Valid until: December 31, 2026
-50%
Anthropic

Anthropic 庆祝 Claude 完成费马大定理验证:研究用途首月 $50 额度

9/5 Claude 完成 Fermat Last Theorem formal verification - research grants $50 API credit for new accounts (verified via Anthropic console)

Original
$0/mo
Now
$0
Valid until: October 31, 2026
-100%
Meta

Meta AIRA₃ Kaggle 金牌:开放 Llama 4 Research API 申请

Meta AIRA₃ ranks 8th among 3,800 teams on NVIDIA Kaggle competition for 30B Nemotron - Llama 4 Research API opens with $100 credits for verified researchers

Original
$0/mo
Now
$0
Valid until: November 30, 2026
-50%
Zhipu AI

GLM-5.3-Flash 50% Off Promo Extended Through 9/30

Zhipu AI extended the 50% promotional discount on GLM-5.3-Flash API through September 30, 2026 (was set to expire 9/9). Promotional rates: $0.075 input / $0.25 output per 1M tokens. 1M context window, 320B/18B active MoE.

Original
$0.15/mo
Now
$0.075/mo
Valid until: September 30, 2026
-75%
DeepSeek

DeepSeek V4-Pro 75% Discount Made Permanent (Was Set to Expire 5/31)

DeepSeek made the 75% promotional discount permanent. V4-Pro now permanently at $0.435 / $0.87 per 1M tokens (cache miss input / output). Cached input: $0.003625 per 1M (90% off). 1M context window, up to 384K output per request. 8/16 added peak/off-peak billing: peak $1.32/$3.96, off-peak (17h) $0.66/$1.98.

Original
$1.74/mo
Now
$0.435/mo
Valid until: December 31, 2026
Anthropic

Anthropic Cancels Sonnet 5 Sep 1 Price Hike: $2/$10 Permanent

Anthropic announced on Aug 10, 2026 that the planned Sep 1 price hike for Claude Sonnet 5 (from $2/$10 to $3/$15 per 1M tokens) is cancelled. The intro rate of $2/$10 is now the standard rate. Note: Sonnet 5 uses a new tokenizer that charges +30% tokens for identical text, so effective cost may still be higher than Sonnet 4.6.

Original
$3/mo
Now
$2/mo
Valid until: December 31, 2026
-50%
Zhipu AI

GLM-5.3-Flash 50% Off - Limited Time

Zhipu's new open-source multimodal model (320B-A18B, confirmed via OpenCode) at half price for the first month: ¥0.4 per 1M input / ¥1.4 per 1M output. Pricing only 1/40 of Opus 4.8.

Original
$2.8/mo
Now
$1.4/mo
Valid until: September 30, 2026
-60%
基元律动

DeepSeek V4 Pro 2.5x Off via TokenRhythm

TokenRhythm offers DeepSeek V4 Pro at ¥3/¥6 per 1M tokens (input/output) - a 60% discount vs official pricing. No credit card needed, ¥68 free credits on signup.

Original
$15/mo
Now
$6/mo
Valid until: October 31, 2026
-100%
基元律动

TokenRhythm Sign-up Bonus: ¥68 Free Credits

Sign up with phone number and get ¥68 free token credits (credited after first valid API call). No credit card required. Plus DeepSeek-V4 Pro at 2.5x off (¥3/¥6) and free cache hits on mimo-v2.5-pro.

Original
$68/mo
Now
$0/mo
Valid until: December 31, 2026
-17%
Cursor

Cursor Pro Annual - Save 2 Months!

Get 2 months free when you switch to annual billing. Best AI coding tool at an unbeatable price.

Original
$240/mo
Now
$200/mo
Valid until: December 31, 2026
-17%
OpenAI

ChatGPT Plus Annual - Save $40!

Annual subscription saves you $40 compared to monthly. Get uninterrupted access to GPT-5.6.

Original
$240/mo
Now
$200/mo
Valid until: December 31, 2026
-17%
Anthropic

Claude Pro Annual - Save $40!

Annual billing for Claude Pro saves you $40. Best for long-context tasks.

Original
$240/mo
Now
$200/mo
Valid until: December 31, 2026
-100%
Google

Gemini API Free Tier Extended!

Google extends the free tier with 1M tokens context. Perfect for budget-conscious developers.

Original
$0/mo
Now
$0
Valid until: December 31, 2026
-12%
Microsoft

GitHub Copilot Annual - Save $30!

Switch to annual billing and save $30 on GitHub Copilot subscription.

Original
$120/mo
Now
$90/mo
Valid until: December 31, 2026
Perplexity

Perplexity Pro - Best Research Tool

Access GPT-5.6, Claude Sonnet 5, and real-time web search in one place. Perfect for research.

Original
$20/mo
Now
$20/mo
Valid until: December 31, 2026
Zhipu AI

GLM-5.3 API Launch - Same Price as GLM-5.2

Zhipu's GLM-5.3 launched on Aug 19, 2026 with API access. Pricing matches GLM-5.2 at $1.40/$4.40 per 1M tokens. AA Intelligence Index of 60 ties with Claude Fable 5 and GPT-5.6 Sol. Weights open-sourced Aug 28.

Original
$1.4/mo
Now
$1.4/mo
Valid until: December 31, 2026
xAI

xAI Grok 4.6 API Launch - $2/$6 Same as 4.5

xAI's Grok 4.6 launched Aug 12, 2026 with 500K context window. Pricing unchanged at $2/$6 per 1M tokens (below 200K prompt). AA Intelligence Index 61 ties with GPT-5.6 Sol. New xhigh reasoning tier.

Original
$2/mo
Now
$2/mo
Valid until: December 31, 2026
Alibaba

Qwen3.8-Max API Launch - $2/$6 1M Context

Alibaba's flagship Qwen3.8-Max launched Aug 3, 2026 with 2.4T parameters (95B active), 1M context window. API price $2/$6 per 1M tokens. Cached input at $0.25 (8x discount). Competitive with Grok 4.6.

Original
$2/mo
Now
$2/mo
Valid until: December 31, 2026