One flat rate. Then nothing.
Every tier is genuinely unlimited: no token meters, no rate limits, no speed caps. GLM tiers differ by concurrent streams over the full 262K window. Qwen tiers ladder the context window from 64K to 262K. Concurrency scales with tier; the token count never changes.
GLM-5.3-Flash · the flagship - unquantized · 262K context · vision · tools · ~100 tok/s
Standard
$109/mo
GLM-5.3-Flash · unquantized
- Unlimited tokens & messages
- Full 262K context
- 1 concurrent stream
- ~100 tok/s, vision, tool calls
Most popular
Pro
$199/mo
GLM-5.3-Flash · unquantized
- Unlimited tokens & messages
- Full 262K context
- 2 concurrent streams
- Priority capacity at peak
Unlimited
$399/mo
GLM-5.3-Flash · unquantized · top of stack
- Unlimited tokens & messages
- Full 262K context
- 4 concurrent streams, top priority
- First access to new models
Founding · 75% off forever
GLM Founding
$99/mo · locked for life
The Unlimited tier ($399) at $99, 75% below the forever price, and your rate never rises as long as you stay subscribed. 10 seats, then gone.
- The Unlimited tier at $99, forever
- Full 262K context
- 4 concurrent streams, top priority
- Your rate never rises as long as you stay subscribed
Dedicated
$5,999/mo
your own Blackwell backend · zero neighbors
- A whole serving node, nobody else on it
- Full 262K context, all 8 streams yours
- No fair-share, no throttle, ever
- Sleep/wake control of your box
Qwen 3.8 27B · uncensored & fast - the community's favorite uncensored model · ~100+ tok/s · 64K → 262K window ladder
Cheapest uncensored
Roleplay
$34.99/mo
Qwen 3.8 27B · uncensored · roleplay-tuned
- Unlimited roleplay & chat
- 64K context window
- 1 stream
- Vision, tool calling & thinking included
Sweet spot
Qwen Plus
$49/mo
Qwen 3.8 27B · uncensored
- Unlimited usage
- 128K context window
- 1 stream
- Vision, tool calling & thinking included
Qwen Pro
$79/mo
Qwen 3.8 27B · uncensored
- Unlimited usage
- Full 262K context window
- 2 streams
- Vision, tool calling & thinking
Qwen Unlimited
$109/mo
Qwen 3.8 27B · uncensored · top Qwen tier
- Unlimited usage
- Full 262K context window
- 4 streams, top priority
- Vision, tool calling & thinking, agents, coding & research
Founding · 46% off forever
Qwen Founding
$59/mo · locked for life
The Qwen Unlimited tier ($109) at $59, 46% below the forever price, locked in writing as long as you stay subscribed. 10 seats, then gone.
- Qwen Unlimited tier at $59, forever
- Full 262K context window
- 4 streams, top priority
- Vision, tool calling & thinking
Qwen Dedicated
$549/mo
your own RTX 5090 backend · zero neighbors
- A whole serving node, nobody else on it
- Full 262K context, all 8 streams yours
- No fair-share, no throttle, ever
- Great for studios & heavy agents
What "unlimited" includes, on every tier
| Daily token cap | none |
| Monthly token cap | none |
| Per-request context limit | full 262,144 (model maximum) |
| Message length ceiling | none |
| Precision served | Unquantized, full precision |
| Rate limits | none |
| Speed caps / throttling | none, ever |
| Fair-share queues | none |
| Concurrent streams | 1 to 4 by tier · 8 on Dedicated |
| Clients | any OpenAI-compatible SDK |
Prices localise to your region automatically (you see pounds if you're in the UK). Founding seats appear with live availability from the queue.
Claim your position
Applications are processed in order. Founding seats are allocated to the first qualifying applicants.