头条
3 家媒体·3 篇报道·历时 8h

Qwen3-Max-Thinking

Qwen has released its flagship reasoning model, Qwen3-Max-Thinking, which it says performs comparably to GPT-5.2 Thinking and Opus 4.5. Independent testing shows the model outperforms Gemini 3 Pro and GPT-5.2 on the Humanity’s Last Exam benchmark, using a test-time scaling technique that delivers efficient, high-scoring reasoning across multiple evaluations.

本文综合以下全部报道写成,而非依据任何单一来源。

各家媒体如何报道

  1. Hacker News·
    Qwen3-Max-Thinking
  2. TechMeme·
    Qwen releases Qwen3-Max-Thinking, its flagship reasoning model that it says demonstrates performance comparable to models such as GPT-5.2 Thinking and Opus 4.5 (Qwen)
  3. VentureBeat·
    Qwen3-Max Thinking beats Gemini 3 Pro and GPT-5.2 on Humanity's Last Exam (with search)