Top Stories
3 outlets·3 reports·over 8h

Qwen3-Max-Thinking

ThinkingNews Desk · how this was written

Qwen has released its flagship reasoning model, Qwen3-Max-Thinking, which it says performs comparably to GPT-5.2 Thinking and Opus 4.5. Independent testing shows the model outperforms Gemini 3 Pro and GPT-5.2 on the Humanity’s Last Exam benchmark, using a test-time scaling technique that delivers efficient, high-scoring reasoning across multiple evaluations.

Written from all 3 reports below, not from any single one.

How it was reported

  1. Hacker News·
    Qwen3-Max-Thinking

    The article likely covers Qwen3-Max-Thinking, possibly a technology or concept. It may discuss its features, applications, or implications. Key details are not provided, suggesting the article may introduce or explore this topic.

  2. TechMeme·
    Qwen releases Qwen3-Max-Thinking, its flagship reasoning model that it says demonstrates performance comparable to models such as GPT-5.2 Thinking and Opus 4.5 (Qwen)

    Qwen releases Qwen3-Max-Thinking, its flagship reasoning model. It demonstrates performance comparable to GPT-5.2 Thinking and Opus 4.5. Qwen3-Max-Thinking is the company's latest model.

  3. VentureBeat·
    Qwen3-Max Thinking beats Gemini 3 Pro and GPT-5.2 on Humanity's Last Exam (with search)

    Alibaba Cloud's Qwen3-Max-Thinking outperforms Gemini 3 Pro and GPT-5.2. It uses a "test-time scaling" technique for efficient reasoning. Qwen3-Max-Thinking achieves high scores on benchmarks like Humanity's Last Exam.

Related stories

Share: