Qwen3-Max-Thinking

ThinkingNews Desk · how this was written
Qwen has released its flagship reasoning model, Qwen3-Max-Thinking, which it says performs comparably to GPT-5.2 Thinking and Opus 4.5. Independent testing shows the model outperforms Gemini 3 Pro and GPT-5.2 on the Humanity’s Last Exam benchmark, using a test-time scaling technique that delivers efficient, high-scoring reasoning across multiple evaluations.
Written from all 3 reports below, not from any single one.
How it was reported
- Hacker News·Qwen3-Max-Thinking
The article likely covers Qwen3-Max-Thinking, possibly a technology or concept. It may discuss its features, applications, or implications. Key details are not provided, suggesting the article may introduce or explore this topic.
- TechMeme·Qwen releases Qwen3-Max-Thinking, its flagship reasoning model that it says demonstrates performance comparable to models such as GPT-5.2 Thinking and Opus 4.5 (Qwen)
Qwen releases Qwen3-Max-Thinking, its flagship reasoning model. It demonstrates performance comparable to GPT-5.2 Thinking and Opus 4.5. Qwen3-Max-Thinking is the company's latest model.
- VentureBeat·Qwen3-Max Thinking beats Gemini 3 Pro and GPT-5.2 on Humanity's Last Exam (with search)
Alibaba Cloud's Qwen3-Max-Thinking outperforms Gemini 3 Pro and GPT-5.2. It uses a "test-time scaling" technique for efficient reasoning. Qwen3-Max-Thinking achieves high scores on benchmarks like Humanity's Last Exam.