Qwen builds the Qwen3 Max Thinking model, a chat-based AI capable of extended thinking and structured output, with notable strengths in intelligence and general problem-solving as evidenced by its top 25% scores in the Intelligence Index, GPQA, and HLE benchmarks. Its context window of 256,000 tokens allows for complex and nuanced conversations.
Input
Output
Context
256K
Max Output
66K
Parameters
-
Input Modalities
Output Modalities
Data sourced from official provider APIs and documentation
Last updated: Aug 6, 2026
Every month new models become cheaper, faster, and more capable. Inferbase ensures your application automatically benefits without changing a single API call.