Official Provider HubHangzhou, China
Alibaba Cloud / Qwen API Pricing & Token Economics
Alibaba Cloud's Qwen 2.5 series (including Qwen 2.5 72B Instruct and Qwen 2.5 Coder 32B) ranks among the highest-performing open foundation models globally. Featuring a 152k vocabulary BPE tokenizer, Qwen models offer remarkable token efficiency for multilingual text, Chinese, mathematics, and agentic coding tasks.
Prompt Caching Policy
Batch / Asynchronous Rules
Standard API execution rules. Refer to individual endpoint rate limits for asynchronous processing.
Special Billing Conditions
Zero write fees on standard prefix evaluation; automatic cache invalidation.
Active Alibaba Cloud / Qwen Models(2 models verified)
All 30 Models →Calculate your monthly Alibaba Cloud / Qwen bill
Simulate prompt caching hit rates, batch discounts, and output token costs with zero server logging.