The new Qwen3-Max model, with over one trillion parameters and training on 36 trillion tokens, shows significant improvements in reasoning, programming, and tool usage, according to independent evaluations.
Alibaba has launched Qwen3-Max, its largest language model to date. It features over one trillion parameters and was trained on a dataset of 36 trillion tokens. Its architecture is based on a mixture of experts, an approach that distributes tasks among specialized subcomponents, contributing to stable and efficient training. Throughout the entire process, the learning curve remained uniform, without interruptions or the need to restart or adjust the data.
Thanks to improvements in distributed computing management, the model achieves 30% more efficiency in resource usage than its predecessor. Additionally, it can handle contexts of up to one million tokens, allowing it to process extremely long documents or interactions without performance loss.
The instructional variant, Qwen3-Max-Instruct, ranks third on LMArena's Text Arena leaderboard. On SWE-Bench Verified, a test that evaluates the ability to solve real-world programming problems extracted from public repositories, it achieves 69.6%, placing it among the most competent models globally. On Tau2-Bench, designed to measure precision in tool usage by AI agents, it scores 74.8%, surpassing systems like Claude Opus 4 and DeepSeek V3.1.
Alibaba is also developing Qwen3-Max-Thinking, a version specialized in complex reasoning. Although still in training, it has already achieved perfect results on demanding mathematical tests such as AIME 25 and HMMT, by combining code execution and advanced inference strategies. The company plans to publicly release this variant in the coming months.
Qwen3-Max-Instruct is now available on the Qwen Chat platform and through the API on Alibaba Cloud. Its compatibility with the OpenAI API format facilitates its integration into existing applications. To access it, users must register on Alibaba Cloud, activate the Model Studio service, and generate an API key. The launch reinforces Alibaba's commitment to offering scalable and open artificial intelligence infrastructure to developers and researchers.
Set of AI models integrating natural language processing, vision, and audio, with some models available as open source. Provides multimodal content analysis and generation, with specialized models ...
16/07/2026
SpaceXAI has introduced Grok 4.5, its most advanced model to date, trained alongside Cursor and built for coding, agentic tasks and knowledge work, ...
15/07/2026
Thinking Machines Lab launches Inkling, its first open-weight AI model: it natively processes text, images and audio, and is designed for other ...
13/07/2026
The Stanford Digital Economy Lab has presented the manifesto We Must Act Now, backed by 16 Nobel laureates and more than 200 economists and ...
09/07/2026
The GPT-5.6 family, with the Sol, Terra and Luna models, is now available in ChatGPT, Codex and the API after its limited preview, with better ...