{T}
TokenMath.net
Official Provider HubHangzhou, China

Alibaba Cloud / Qwen API Pricing & Token Economics

Alibaba Cloud's Qwen 2.5 series (including Qwen 2.5 72B Instruct and Qwen 2.5 Coder 32B) ranks among the highest-performing open foundation models globally. Featuring a 152k vocabulary BPE tokenizer, Qwen models offer remarkable token efficiency for multilingual text, Chinese, mathematics, and agentic coding tasks.

Prompt Caching Policy

Batch / Asynchronous Rules

Standard API execution rules. Refer to individual endpoint rate limits for asynchronous processing.

Special Billing Conditions

Zero write fees on standard prefix evaluation; automatic cache invalidation.

Active Alibaba Cloud / Qwen Models(2 models verified)

All 30 Models →
ModelContextInput / 1MCached / 1MOutput / 1MBatch DiscountActions
Qwen 2.5 72B Instruct — Together AI
qwen/qwen-2.5-72b-instruct
128k$0.35$0.175$0.40N/ASpecs →
Qwen 2.5 Coder 32B — Together AI
qwen/qwen-2.5-coder-32b-instruct
128k$0.08$0.16N/ASpecs →

Calculate your monthly Alibaba Cloud / Qwen bill

Simulate prompt caching hit rates, batch discounts, and output token costs with zero server logging.

Open Cost Calculator →