AI model subscription service for teams with Credits-based billing across text, image, and video generation models
Token Plan Team Edition is an AI model subscription service from Qwen Cloud, with Credits-based billing across text, image, and video generation models. Compatible with mainstream AI programming and agent tools, with data security guarantees and stable performance.
Go to the Token Plan pricing page to subscribe.
A seat is the smallest subscription unit, representing one team member's usage quota. Admins assign seats to members in the Team Management Console, which generates a dedicated API Key per member.
A flexible shared usage package for all seats in your team. When a seat's quota is exhausted, overage is deducted from this package. Each package is valid for one month, and unused Credits expire at the end of the period. If you have multiple packages, the one with the earliest expiration date is used first.
Credits consumed per request are not fixed -- they depend on model type, token usage, thinking mode, and tool calls. Token usage grows with accumulated context (conversation history, code, tool returns, etc.), and some models use tiered pricing based on context length.
Estimated Credits for a single request using
Go to the Token Plan page to view subscription status, expiry date, remaining Credits, and usage details. Admins can also view member-level consumption in the Management Console.
Plan details
- Flexible model switching: Switch between various models on demand, with usage uniformly deducted from your Credits.
- Broad tool compatibility: Compatible with various popular programming and agent tools. See Set up your AI tool.
- Multiple plan tiers: Choose from Standard, Pro, and Max seats to match different usage levels.
- Team management: Provides an admin console with seat assignment, reclamation, and member usage analytics.
- Predictable costs: Monthly subscription model keeps your budget predictable.
- Data security: Your conversation data is never used for model training.
- Stable performance: Multi-tenant isolation architecture ensures smooth service without queuing.
Supported models
Team Edition supports the qwen3.8-max-preview preview model with the following limited-time benefits:
- Preview: qwen3.8-max-preview is currently a preview model. Its capabilities may be iteratively upgraded during the preview period. After the preview ends, this model may be taken offline or replaced with the production version.
- Limited-time 10x usage: During the promotional period, Credits consumption is as low as 10% of the standard rate, effectively providing 10x the usage.
| Brand | Model | Capability |
|---|---|---|
| Qwen | qwen3.8-max-preview | Reasoning, vision understanding, text generation |
| Qwen | qwen3.7-max | Reasoning, text generation |
| Qwen | qwen3.7-plus | Reasoning, vision understanding, text generation |
| Qwen | qwen3.6-plus | Reasoning, vision understanding, text generation |
| Qwen | qwen3.6-flash | Reasoning, vision understanding, text generation |
| Qwen | qwen-image-2.0 | Image generation |
| Qwen | qwen-image-2.0-pro | Image generation |
| Wan | wan2.7-image | Image generation |
| Wan | wan2.7-image-pro | Image generation |
| DeepSeek | deepseek-v4-pro | Reasoning, text generation |
| DeepSeek | deepseek-v4-flash | Reasoning, text generation |
| DeepSeek | deepseek-v3.2 | Reasoning, text generation |
| Moonshot AI | kimi-k2.7-code | Reasoning, vision understanding, text generation |
| Moonshot AI | kimi-k2.6 | Reasoning, vision understanding, text generation |
| Moonshot AI | kimi-k2.5 | Reasoning, vision understanding, text generation |
| Zhipu AI | glm-5.2 | Reasoning, text generation |
| Zhipu AI | glm-5.1 | Reasoning, text generation |
| Zhipu AI | glm-5 | Reasoning, text generation |
| MiniMax | MiniMax-M2.5 | Reasoning, text generation |
| HappyHorse | happyhorse-1.1-t2v | Video generation |
| HappyHorse | happyhorse-1.1-i2v | Video generation |
| HappyHorse | happyhorse-1.1-r2v | Video generation |
Pricing
Go to the Token Plan pricing page to subscribe.
Seat plans
A seat is the smallest subscription unit, representing one team member's usage quota. Admins assign seats to members in the Team Management Console, which generates a dedicated API Key per member.
| Seat type | Price | Quota | Use cases |
|---|---|---|---|
| Standard Seat | Limited-time $20/seat/month | 25,000 Credits/seat/month | For team members with light AI usage |
| Pro Seat | Limited-time $75/seat/month | 100,000 Credits/seat/month | For team members who frequently use AI for coding |
| Max Seat | $200/seat/month | 250,000 Credits/seat/month | For core developers who heavily rely on AI for coding |
Shared usage package
A flexible shared usage package for all seats in your team. When a seat's quota is exhausted, overage is deducted from this package. Each package is valid for one month, and unused Credits expire at the end of the period. If you have multiple packages, the one with the earliest expiration date is used first.
| Tier | Price | Quota |
|---|---|---|
| Shared usage package | $700/package | 625,000 Credits/package |
Credits billing
How it works
Credits consumed per request are not fixed -- they depend on model type, token usage, thinking mode, and tool calls. Token usage grows with accumulated context (conversation history, code, tool returns, etc.), and some models use tiered pricing based on context length.
Example
Estimated Credits for a single request using qwen3.6-plus (actual consumption varies by model):
| Token type | Quantity | Credits consumed |
|---|---|---|
| Input Tokens | 8,349 | 1.67 |
| Cached Tokens | 40,794 | 0.82 |
| Output Tokens | 573 | 0.69 |
| Total | Approx. 3.18 Credits |
The above is a single-request example and does not represent a fixed cost per request. In multi-turn conversations (AI coding, agents, etc.), each request carries accumulated context, increasing input tokens and Credits consumption over time. To control consumption, start new sessions when switching tasks, and clear irrelevant history to keep context short.
Deduction order
- Credits are first deducted from the seat's monthly quota.
- After the seat quota is exhausted, Credits are deducted from the shared usage package. If you have multiple packages, the one expiring soonest is used first.
- When all quotas are depleted, the service is suspended until the next billing cycle or until you purchase a shared usage package.
View usage
Go to the Token Plan page to view subscription status, expiry date, remaining Credits, and usage details. Admins can also view member-level consumption in the Management Console.
Terms of use
- Scope of use: For interactive use within compatible AI tools only. Not for automated scripts or application backends. Violations may result in suspension.
- Data security: Token Plan Team Edition does not use your conversation data for model training.
- Account rules: API Keys are for the assigned member's use only. Do not share or expose them publicly.
- Service region and cross-border data transfer: The only available region is Singapore with Global deployment mode. Model inference is performed globally, involving cross-border data transfer. You are responsible for compliance with applicable laws.