摘要
腾讯发布并开源了下一代大模型 Hy4 预览版:770B 参数 MoE,每 token 激活 49B,上下文超过 1M token;以标准与 FP8 两种权重上架,许可证为 Apache 2.0(以 HF 标签为准,公告正文未写许可)。主打场景:软件工程(规划、调试、前端质量)、办公(金融与数据分析、跨文档协作)、游戏开发(一句自然语言生成可玩原型)与科学研究。公开 benchmark 证据很薄:仅一项腾讯内部盲测——163 名专家、203 个工程任务,Hy4 预览版 2.99/4.00,略高于 GLM-5.3(2.92)与 Kimi K3(2.94)。官方称模型参与了自己的训练流水线,并自优化推理系统使端到端吞吐提升 31.8%。完整版 Hy4 尚未发布——官方采用 preview-first 策略,下一批「很快」。API 经腾讯云 TokenHub 与 OpenRouter 提供:输入 $0.834/M、输出 $2.501/M、缓存 $0.042/M。
为什么重要
这是一周内第三家在 Hugging Face 放出重要开源权重的中国实验室(Z.ai 8/25、Qwen 8/24-26、腾讯 8/27-28),770B + Apache 2.0 的组合也最接近「同级别、不同组织」的开源前沿发布——趋势 #1 的直接证据。但能力声称未经验证:没有公开 benchmark 分数、只有一项厂商自跑的盲测、且带 preview 标。服务团队有 day-0 的 vLLM recipe 可用;对模型的真实判断要等公开测试框架(SWE-bench、Terminal-Bench)跑出结果。
技术细节
| 架构 | 770B 总参数 / 49B 激活 MoE;上下文超过 1M token |
|---|---|
| 许可证 | 以 Hugging Face 标签为准为 Apache 2.0(公告正文未写许可) |
| 产物 | HF tencent/Hy4-preview + FP8(仓库 2026-08-27T08:52Z 创建);ModelScope 镜像;day-0 vLLM recipe;截至 2026-08-30 主仓约 1.4k 下载、FP8 约 1.3k、283 likes |
| 评测 | 仅内部盲测:163 名专家、203 个工程任务;Hy4 预览版 2.99/4.00,GLM-5.3 2.92、Kimi K3 2.94;无公开 benchmark 分数(未上 SWE-bench/Terminal-Bench) |
| 自我改进 | 官方称模型参与自身训练流水线,并自优化推理系统使端到端吞吐 +31.8%(厂商声称) |
| API | 腾讯云 TokenHub + OpenRouter;输入 $0.834/M、输出 $2.501/M、缓存 $0.042/M;WorkBuddy/CodeBuddy 免费两周;Hy3 免费延长至 2026-09-30 |
| 节奏 | preview-first;完整版 Hy4 权重未发布;下一批 Hy4 系列「预计很快」 |
后续更新
2026-08-31 Adoption & follow-up check @2026-08-31T00:00Z: downloads 2,123 main (+~700 vs the 8/30 sample) + 319 likes; FP8 1,469 — modest velocity compared with the GLM-5.3-Flash cohort. No full Hy4 release yet; the announcement still only says the next batch is expected 'soon' (community speculates ~2.5 months to GA from the Hy3 precedent). Verification note: aggregators (datalearner etc.) list public benchmark scores (Terminal Bench 2.1 85.4, SWE-Bench Pro 65.7) that do NOT appear in the official announcement or the HF model card — unverified, not counted; the official page still shows only the internal blind test. Community GGUF compression to ~200GB claimed at ~98% performance retention (unverified). Recommendation stays WATCH.
2026-09-02 adoption check @2026-09-02T00:0xZ: 3,516 downloads (+66% vs 2,123 @8/31) + 383 likes — continued but still modest vs the GLM-5.3-Flash cohort; full Hy4 not shipped ('soon' unchanged), no official public benchmarks. Trend #1 criterion (b) candidate status unchanged. Stays WATCH
2026-09-05 Adoption check @2026-09-05: 5,684 downloads (+62% vs 9/2) + 430 likes; FP8 2,666. Full Hy4 still not shipped (tencent org's latest repos remain Hy4-preview/Hy4-preview-FP8 from 8/27; flagship line still Hy3 from July). Third-party quantization ecosystem growing despite the preview: AngelSlim/Hy4-preview-GGUF 109,240 downloads; mlx-community 4bit; inferencerlabs MLX Q4i (9/2); anemll FlashMoE-STQ1_0 (9/3). Stays WATCH
标签
open-weightsmoelong-contextapache-2codingtencent