Build more, spend less — one plan, every modality. The Model Studio Token Plan for Individual is now live — featuring the exclusive debut of Qwen3.8-Max-Preview, plus access to every text, video, image, and audio model on the platform at up to 3x the usage compared to pay-as-you-go pricing. Early-bird launch pricing starts at just $6/month — and if you're already using off-peak rates, those discounts stack with your Token Plan credits.
Whether you're using Qwen3.8-Max-Preview for full-stack development and data analysis, generating video content with HappyHorse 1.1, or prototyping multimodal applications, the Token Plan eliminates per-call cost anxiety and lets you focus on shipping.
Qwen3.8-Max-Preview debuts exclusively on the Model Studio Token Plan for Individual — the latest and most capable model in the Qwen family, available here before anywhere else.
Whether you're doing full-stack development, data analysis, or automating Office workflows, Qwen3.8-Max-Preview handles it with ease. It combines frontier-level text reasoning with vision understanding, making it equally capable of reading a complex codebase, interpreting a spreadsheet screenshot, or drafting a slide deck outline from a rough sketch.
First-mover advantage: Qwen3.8-Max-Preview is available exclusively through the Token Plan — get early access to the most powerful Qwen model while it's in preview, at a fraction of what flagship models cost elsewhere.
Pay-as-you-go pricing works — until it doesn't. Once you're running AI agents that chain multiple model calls, generating video with HappyHorse, or building voice-enabled applications, per-token billing becomes unpredictable and expensive.
The Token Plan solves this with a single monthly credit allocation that covers every model and modality on the platform. You get a predictable budget, dramatically lower effective cost per token, and the freedom to experiment across capabilities without worrying about a surprise bill.
The core value: For the same spend, Token Plan users get approximately 3x more usage than pay-as-you-go — and that's before off-peak discounts stack on top.
Three tiers designed for different usage levels, all with limited-time early-bird pricing:
Tier | Credits per 5 Hours | Credits per 7 Days | Original Price | Launch Price |
Lite | 700 | 2,500 | $8/mo | $6/mo |
Standard (Recommended) | 3,000 | 10,000 | $25/mo | $18/mo |
Pro | 12,000 | 40,000 | $80/mo | $68/mo |
How credit allocation works: Like Codex and other leading AI subscription products, the Token Plan uses a rolling reset model — not a flat monthly pool. Your credits refresh on two cycles:
This means you always have credits available for your next work session — the plan is designed for sustained daily use, not a lump-sum budget you burn through in the first week.
Early-bird pricing is limited time. Lock in your rate now — prices will return to standard once the launch period ends.
Commit quarterly or yearly to unlock additional savings on top of early-bird pricing:
Period | Lite | Standard | Pro |
Monthly | $6/mo | $18/mo | $68/mo |
Quarterly | $17/quarter (~5% off) | $51/quarter (~6% off) | $195/quarter (~4% off) |
Yearly | $65/year (~10% off) | $195/year (~10% off) | $770/year (~6% off) |
The Standard tier at yearly billing works out to roughly $16.25/month — that's 10,000 credits per week (refreshing every 7 days) of multimodal AI for the price of a lunch.
Unlike single-model subscriptions, the Token Plan gives you access to the full spectrum of Model Studio capabilities with one API key:
Your Token Plan credits also cover platform-native tools that your agents can call directly:
One subscription. Every modality. No separate billing for text, video, image, and audio — it's all covered under your credit allocation.
Dollar for dollar, the Token Plan delivers approximately 3x more usage compared to spending the same amount on pay-as-you-go API calls. The math is simple: you're buying credits in bulk at a substantial discount.
Here's what makes the Token Plan uniquely powerful: off-peak rates (22:00–08:00 Beijing Time) stack directly with your Token Plan credits. During off-peak hours, your credits stretch even further:
Model | Standard Credit Consumption | Off-Peak Credit Consumption | Effective Multiplier |
Qwen3.7-Max | 1x | 0.2x | 5x more usage |
Qwen3.7-Plus | 1x | 0.4x | 2.5x more usage |
What this means in practice: A Standard plan user (10,000 credits/week) running Qwen3.7-Max exclusively during off-peak hours effectively gets 50,000 credits worth of usage per week — that's the equivalent of 18/month.
The stacking effect is the key differentiator. No other AI platform offers a subscription discount that compounds with time-of-day pricing. For US and EU developers, off-peak hours cover most of the business day — so this isn't a niche optimization, it's your default operating mode.
Consider a developer on the Standard plan ($18/mo) who runs 70% of their workload during off-peak hours using Qwen3.7-Max:
That's nearly 4x the Pro tier's weekly allocation at less than a third of the price — simply by timing your workloads smartly.
700 credits per 5 hours / 2,500 credits per week is enough for regular AI-assisted coding, occasional image generation, and light experimentation. Perfect if you're:
3,000 credits per 5 hours / 10,000 credits per week covers serious daily usage across multiple modalities. Built for developers who:
12,000 credits per 5 hours / 40,000 credits per week for high-volume production workloads. Designed for:
Start with Lite, move to Standard when you hit your stride, upgrade to Pro when you scale. No penalty for upgrading mid-cycle — your new allocation kicks in immediately. You don't need to commit big upfront.
The Token Plan's rolling reset model keeps your credits fresh — here's what each tier provides:
Tier | Refreshed Every 5 Hours | Refreshed Every 7 Days | Concurrent Agents |
Lite | 700 credits | 2,500 credits | 1–2 |
Standard | 3,000 credits | 10,000 credits | 3–4 |
Pro | 12,000 credits | 40,000 credits | 6–8 |
Unlike a flat monthly quota that encourages hoarding or binge-spending, the 5-hour + 7-day rolling reset ensures:
This is the same cadence used by Codex and other professional AI subscription products — it's purpose-built for developers who use AI tools every day, not once a month.
Concurrent agent limits define how many AI tools or agents can use your API key simultaneously — critical if you're running multiple coding assistants or agent pipelines in parallel.
Credits are the unified usage unit across all models on the Token Plan. Here's what you need to know:
Think of credits like a renewable daily energy bar, not a savings account. The system is designed so you always have credits available for your next session — use them actively, because they refresh on schedule regardless.
The Token Plan's multimodal coverage makes it uniquely suited for cross-capability workflows:
Models/Capabilities | Industries | Scenarios |
Qwen3.8-Max-Preview (flagship text + agentic) | Enterprise, Finance, Consulting | Report generation, code assistant, real-time analysis, competitive intelligence |
HappyHorse 1.1 (video generation) | Entertainment/MCN, E-commerce, Education | Short video production, course animation, brand ads |
Qwen-Audio-3.0-TTS-Plus, Fun-ASR | Customer Service, Education, Smart Devices | Voice assistants, language tutoring, podcast narration, meeting transcription |
Wan2.7-Image (image generation) | E-commerce, Design, Social Media | Product images, poster design, social media content |
A typical e-commerce AI workflow on the Standard plan:
All four steps draw from the same 40,000 credit pool — no separate subscriptions, no juggling API keys.
Visit the Model Studio Token Plan page and select your tier. Early-bird pricing is applied automatically during the launch period.
After subscribing, you'll receive a dedicated Token Plan API key in the format sk-sp-xxxxx. This key is separate from regular Model Studio API keys.
Important: Your Token Plan API key (sk-sp-xxxxx) and regular Model Studio API key are not interchangeable. Use the dedicated key to ensure calls are billed against your credit allocation.
The Token Plan works with any OpenAI-compatible or Anthropic-compatible tool:
OpenAI-compatible endpoint:
https://coding.dashscope.aliyuncs.com/v1Anthropic-compatible endpoint:
https://coding.dashscope.aliyuncs.com/apps/anthropicCompatible tools include:
One API key works across all your tools simultaneously — no need to generate separate keys for each application.
That's it. Your credits are live, off-peak stacking is automatic, and every model on the platform is available through your single API key. Start building.
The Token Plan for Individual is a monthly credit-based subscription that gives you access to all Model Studio models — text, video, image, and audio — at a significantly lower effective cost than pay-as-you-go pricing. One subscription, one API key, every modality.
On average, Token Plan users get approximately 3x more usage for the same dollar amount. When combined with off-peak discounts (which stack with Token Plan credits), effective savings can reach 5x or more on supported models.
Yes. During off-peak hours (22:00–08:00 Beijing Time), your credit consumption is reduced — Qwen3.7-Max uses only 0.2x credits and Qwen3.7-Plus uses only 0.4x credits. This is in addition to the baseline savings of the Token Plan itself. Learn more about off-peak rates.
All current Model Studio models: Qwen3.8-Max-Preview, Qwen3.7-Max, Qwen3.7-Plus (text/reasoning), HappyHorse 1.1 (video), Wan2.7-Image (image), Qwen-Audio-3.0-TTS-Plus, Fun-ASR (audio), plus built-in harness tools like web search, code interpreter, and web scraping.
No. The Token Plan uses a dedicated API key format (sk-sp-xxxxx) that is separate from regular Model Studio API keys. You'll receive your dedicated key after subscribing.
No. Credits refresh on a rolling basis — every 5 hours and every 7 days — and unused credits within a window do not carry forward to the next reset. The system is designed for consistent daily use, not accumulation. Off-peak stacking helps you maximize every credit during each window.
Yes. You can upgrade from Lite to Standard or Standard to Pro at any time. Your new credit allocation takes effect immediately.
No. The Token Plan is designed for interactive use with AI coding and agent tools — not for automated scripts or backend service integrations. Use standard pay-as-you-go API keys for programmatic production workloads.
No. One subscription per identity. You can upgrade your tier but cannot hold multiple active plans simultaneously.
Any OpenAI-compatible or Anthropic-compatible tool works — including Qwen Code, Claude Code, OpenClaw, OpenCode, and any client that supports the OpenAI or Anthropic API format. One API key works across all tools simultaneously.
Credits operate on two rolling windows: a 5-hour allocation and a 7-day allocation. If you exhaust either, service pauses until that window resets — the 5-hour window refreshes every 5 hours, and the weekly window resets every 7 days. If you consistently run out, consider upgrading to a higher tier for more headroom.
No refunds are available for Token Plan subscriptions. We recommend starting with the Lite tier to evaluate fit before committing to a higher plan or longer billing cycle.
The Token Plan is the most cost-effective way to access every AI capability on Model Studio — text, video, image, and audio — through a single subscription. Early-bird pricing won't last forever.
sk-sp-xxxxx API key from the consoleBuild more, spend less — one plan, every modality. Get promo coupon and subscribe now →
For more on the AI tools and models available on Model Studio, explore our guides on OpenAI API alternatives, AI agents, and HappyHorse video generation.