3 媒体·3 本の記事·期間 8h
Qwen3-Max-Thinking
Qwen has released its flagship reasoning model, Qwen3-Max-Thinking, which it says performs comparably to GPT-5.2 Thinking and Opus 4.5. Independent testing shows the model outperforms Gemini 3 Pro and GPT-5.2 on the Humanity’s Last Exam benchmark, using a test-time scaling technique that delivers efficient, high-scoring reasoning across multiple evaluations.
以下のすべての記事をもとに作成しています。特定の一社によるものではありません。
各社の報じ方
- Hacker News·Qwen3-Max-Thinking
- TechMeme·Qwen releases Qwen3-Max-Thinking, its flagship reasoning model that it says demonstrates performance comparable to models such as GPT-5.2 Thinking and Opus 4.5 (Qwen)
- VentureBeat·Qwen3-Max Thinking beats Gemini 3 Pro and GPT-5.2 on Humanity's Last Exam (with search)