Qwen3.8-Max-0902 Launches: 2.4T Parameters, 1M Context, and New API Pricing

Qwen ·

Key Info

Qwen3.8-Max has been upgraded to Qwen3.8-Max-0902, a 2.4T-parameter model with a 1M-token context window designed for complex, real-world workloads. It is now available via the QwenCloud API at $2 per 1M input tokens and $6 per 1M output tokens.

Highlights

  • Stronger on real-world tasks: Further post-training on Coding & Cowork improves performance on complex enterprise tasks, scientific research, and long-horizon workflows.
  • Large scale and context: 2.4T parameters and 1M context tokens support demanding production scenarios.
  • API pricing: $2 input / $6 output per 1M tokens; explicit cache hits cost $0.17 and implicit cache hits cost $0.25 per 1M tokens.
  • Immediate availability: The upgrade is live via API on QwenCloud.
Loading...