Qwen3小升级即SOTA，开源大模型王座快变中国内部赛了

闻乐
2025-07-22
13:05:00

来源：量子位

官方：大招还在后面

闻乐发自凹非寺

量子位 | 公众号 QbitAI

开源大模型正在进入中国时间。

Kimi K2风头正盛，然而不到一周，Qwen3就迎来最新升级，235B总参数量仅占Kimi K2 1T规模的四分之一。

基准测试性能上却超越了Kimi K2。

Qwen官方还宣布不再使用混合思维模式，而是分别训练Instruct和Thinking模型。

所以，此次发布的新模型仅支持非思考模式，现在网页版已经可以上线使用了，但通义APP还未见更新。

Qwen官方还透露：这次只是一个小更新！大招很快就来了！

但总归就是，再见Qwen3-235B-A22B，你好Qwen3-235B-A22B-2507了。

By the way，这个名字怎么取得越来越复杂了。

先来看看这次的“小更新”都有哪些～

增强了对256K长上下文的理解能力

新模型是一款因果语言模型，采用MoE架构，总参数量达235B，其中非嵌入参数为234B，推理时激活参数为22B。

在官方介绍中显示，模型共包含94层，采用分组查询注意力（GQA）机制，配备64个查询头和4个键值头，并设置128个专家，每次推理时激活8个专家。

该模型原生支持262144的上下文长度。

这次改进主要有以下几个方面：

显著提升了通用能力，包括指令遵循、逻辑推理、文本理解、数学、科学、编码和工具使用。
大幅增加了多语言长尾知识的覆盖范围。
更好地符合用户在主观和开放式任务中的偏好，能够提供更有帮助的响应和更高质量的文本生成。
增强了对256K长上下文的理解能力。

在官方发布的基准测试中可以看到，相较于上一版本，新模型在AIME25上准确率从24.7%上升到70.3%，表现出良好的数学推理能力。

而且对比Kimi K2、DeepSeek-V3，Qwen3新模型的能力也都略胜一筹。

为了提高使用体验，官方还推荐了最佳设置：

Qwen3新版本深夜发布就立刻收获了一众好评：Qwen在中等规模的语言模型中已经领先。

也有网友感慨Qwen在开启新的架构范式：

One More Thing

有趣的是，就在Qwen3新模型发布的前两天，NVIDIA也宣称发布了新的SOTA开源模型OpenReasoning-Nemotron。

该模型提供四个规模：1.5B、7B、14B和32B，并且可以实现100%本地运行。

但实际上，这只是基于Qwen-2.5在Deepseek R1数据上微调的模型。

而现在Qwen3已经更新，大招已经被预告。

随着Llama转向闭源的消息传出，OpenAI迟迟不见Open，开源基础大模型的竞争，现在正在进入中国时间。

DeepSeek丢了王座，Kimi K2补上，Kimi K2坐稳没几天，Qwen的挑战就来了。

体验链接：https://chat.qwen.ai/

参考链接：
[1]https://x.com/Alibaba_Qwen/status/1947344511988076547
[2]https://x.com/giffmana/status/1947362393983529005

— 完 —

2025 年 7 月
一	二	三	四	五	六	日
	1	2	3	4	5	6
7	8	9	10	11	12	13
14	15	16	17	18	19	20
21	22	23	24	25	26	27
28	29	30	31

ufabet มีเกมให้เลือกเล่นมากมาย: เกมเดิมพันหลากหลาย ครบทุกค่ายดัง

tornado crypto mixer Discover the power of privacy with TornadoCash! Learn how this decentralized mixer ensures your transactions remain confidential.

ดูบอลสด Very well presented. Every quote was awesome and thanks for sharing the content. Keep sharing and keep motivating others.

ดูบอลสด Pretty! This has been a really wonderful post. Many thanks for providing these details.

ดูบอลสด Hi there to all, for the reason that I am genuinely keen of reading this website’s post to be updated on a regular basis. It carries pleasant stuff.

Obrazy Sztuka Nowoczesna Thank you for this wonderful contribution to the topic. Your ability to explain complex ideas simply is admirable.

ufabet Hi there to all, for the reason that I am genuinely keen of reading this website’s post to be updated on a regular basis. It carries pleasant stuff.

ufabet You’re so awesome! I don’t believe I have read a single thing like that before. So great to find someone with some original thoughts on this topic. Really.. thank you for starting this up. This website is something that is needed on the internet, someone with a little originality!

ufabet Very well presented. Every quote was awesome and thanks for sharing the content. Keep sharing and keep motivating others.

Qwen3小升级即SOTA，开源大模型王座快变中国内部赛了

Qwen3小升级即SOTA，开源大模型王座快变中国内部赛了

增强了对256K长上下文的理解能力

One More Thing

test

test

文心AIGC

test

test