- 开始
- 2026年8月3日可信度 90%来自来源
Qwen3.8-Max 发布
由原文自动翻译
原标题: Qwen3.8-Max Released
- 阿里巴巴于 2026 年 8 月 3 日推出 Qwen3.8-Max,称其为“Qwen 系列迄今最强大的模型”,面向全球开发者上线其 Model Studio API,并登陆其智能体平台 QwenWork;此前 7 月已进行过预览[1][3]
- 该模型在 Text Arena 排名第五,Vision Arena 排名第二,Frontend Code Arena 排名第四;在一次内部测试中,它自主编程 16 天,构建并开源了名为“oh-my-cli”的智能体框架[1][6]
- 当天,阿里巴巴在纽约上市的股票上涨 4.5%,在香港上市的股票上涨 7%[6]
- 8 月 12 日,该模型以 Qwen3.8-2.4T-A95B 之名开放了权重,是首款公开发布权重的 Max 级 Qwen 模型;其许可条款要求年收入超过 5000 万美元的服务提供商签署商业许可协议;这一开放模型不支持图像输入和非思考模式,而采用 Apache 2.0 许可、于 8 月 14 日发布的 Qwen3.8-27B 则保留了这两项功能[3][5]
- 9 月 2 日,
qwen3.8-maxAPI 切换到了新的快照版本 qwen3.8-max-0902;9 月 23 日,阿里巴巴新增了 Qwen3.8 Max Prime,这是同一模型的更高吞吐量档位,价格为标准版的两倍[2][4]
值得关注的功能
- 该模型是一款采用混合注意力机制的稀疏混合专家模型,基于 Qwen 3.5 构建:总参数 2.4 万亿,激活参数 950 亿,共 92 层、512 个专家(每个 token 路由到 10 个专家,外加 1 个共享专家)[1][5]
- 规格:云端 API 上下文窗口为 1M token,最多支持 131,072 个输出 token(开放权重版本原生支持 262,144 token,可扩展至 1,010,000);支持文本、图像和视频输入,支持函数调用、结构化输出、网络搜索和上下文缓存[2][5]
- 原生多模态:阿里巴巴表示,该模型可以将上百页的文档、完整的电视剧集或长达 100 小时的直播内容转化为可检索的知识库,还能仅凭一张截图重建前端项目[1]
- 每百万 token 价格:在北京地域的 Model Studio 上,输入 12 元人民币、输出 36 元人民币;Prime 版本输入 4 美元、输出 12 美元[2][4]
- 阿里巴巴公布的基准测试表显示,该模型在 Terminal Bench 2.1 上得分 86.6,而 Fable 5 为 84.6、GPT-5.6 Sol 为 88.8;在 SWE-bench Pro 上得分 67.7,Fable 5 为 80.0[5]
参考资料 7可信度 81%总体可信度: 81%该图钉的来源和参考资料对其日期的支持程度。包括来源在内的 7 份资料对图钉开始和结束时间支持程度的加权平均;资料每比最新的一份旧 180 天,权重减半显示所有可信度不低于 75% 的图钉
第一项始终是图钉的来源。总体可信度是各份资料对上方所用开始和结束时间支持程度的加权平均;资料每比最新的一份旧 180 天,权重减半。
- [1]90%alibabagroup.com/en-US/document-2021044032125272064alibabagroup.com· 发布于 2026年9月28日· 开始 2026年8月3日 ✓· 占评分 16%
阿里巴巴集团的发布公告日期为“2026 年 8 月 3 日”,文中写道:“该模型现已可通过阿里云 Model Studio 的 API 面向全球开发者使用,模型权重计划于下周发布”;CNBC[6] 报道称阿里巴巴于“周一”(8 月 3 日)“发布”了该模型。
- [2]90%qwen3.8-max Model Info - Alibaba Cloud Model Studiohelp.aliyun.com· 添加于 2026年9月28日· 占评分 16%
Alibaba Cloud's model page: model ID qwen3.8-max, now the snapshot qwen3.8-max-0902 (alias qwen3.8-max-2026-09-02); image, text and video input; a 1,000,000-token context window with 131,072 max output; function calling, structured outputs, web search and context caching; CNY 12 input and CNY 36 output per million tokens in Beijing.[1]
- [3]75%Qwen - Wikipediaen.wikipedia.org· 添加于 2026年9月28日· 占评分 16%
Dates the cloud release to 3 August 2026, the open weights (Qwen3.8-2.4T-A95B, without image input or non-thinking mode) to 12 August, with a licence requiring providers earning more than US$50 million to get a commercial licence, and the Apache 2.0 Qwen3.8-27B to 14 August.[1]
- [4]70%Qwen3.8 Max Prime: Is it actually better? - DataNorth AIdatanorth.ai· 发表于 2026年9月24日· 占评分 15%
Reports that Alibaba's Qwen team released Qwen3.8 Max Prime on 23 September 2026 as a higher-throughput version of Qwen3.8 Max at $4 per million input and $12 per million output tokens, "exactly twice the price of the standard model", and that it was not measurably faster in OpenRouter's first measurements.[1]
- [5]90%Qwen/Qwen3.8-2.4T-A95B · Hugging Facehuggingface.co· 发表于 2026年8月12日· 占评分 13%
The open-weights model card: "For the first time, Qwen3.8 brings a Qwen-Max-class model to open release"; 2.4T total and 95B activated parameters, 92 layers, 512 experts (10 routed + 1 shared), 262,144 tokens natively extensible to 1,010,000, and benchmarks against Opus 4.8, Fable 5 and GPT-5.6 Sol.[1]
建议更正
有遗漏或错误吗?用你自己的话说明:能佐证此图钉的链接、不同的开始或结束日期及理由,或缺失、有误的信息。AI 会对照此图钉的来源进行核实,搜索更好的来源,并添加任何支持你说法的页面。图钉自身的来源仍然最重要。AI 也会查看图片:显示的是别的东西或显示效果差的图片会被移到后面或替换。