微信小微代理运行在一个秘密的617B MoE模型上——腾讯的低调前沿赌注

持续追踪 : Économie de l'open frontier : viabilité, subvention, pivots· 连载 19/20

模型与工具 Aug 13, 2026 at 12:579加入收藏

微信小微代理运行在一个秘密的617B MoE模型上——腾讯的低调前沿赌注
插图 : Léa Fontaine

WeLM,即微信AI助手背后的模型,已悄然扩展至6170亿参数,并采用未公开的解码机制。没有基准测试,没有论文——仅凭一次十亿用户级别的部署。

简单来说: 腾讯的小微智能体正在微信内进行灰度测试,其运行基于名为 WeLM 的稀疏混合专家模型,该模型已悄然扩展至 6170 亿参数。目前尚无公开基准测试,也无相关论文,仅在全球最高流量的即时通讯平台之一进行大规模灰度测试。

事实

据 Pandaily 报道的微信 AI 团队披露,WeLM 的稀疏 MoE 版本已达到 6170 亿参数,每个 token 激活 230 亿参数。腾讯从未公开发布过 WeLM。该模型为微信集成的 AI 助手小微提供动力——目前正在微信 13 亿+月活用户群体中进行灰度测试。

我们的解读

LLM 竞赛中“榜单至上”的叙事忽略了 WeLM 这类模型:前沿规模、生产路径、完全封闭,且依托分发护城河,这是任何 API 优先的实验室都无法匹敌的。阿里巴巴通过基准测试和开放权重与通义 Qwen 竞争。腾讯则通过嵌入技术竞争。

技术细节: 稀疏 MoE 架构每个 token 仅激活部分参数——WeLM 在每次推理步骤中激活其 6170 亿参数中的 230 亿。在计算上,这相当于 ~230 亿密集模型,这解释了为何名义上如此大规模的模型能在面向消费者的聊天界面中实现可接受的延迟。

报告中披露的“隐藏解码机制”可能指的是推测解码或为延迟优化的早退路由策略——这类推理工程不会出现在论文中,却能让产品变得更快。

关注点

任何 WeLM 的公开基准测试披露;腾讯是否会开放 WeLM 的权重(不太可能);随着灰度测试扩大,小微在性能上如何与 ChatGPT 和 Kimi 相比。

Resources

本文由人工智能撰写,并经人工编辑审核。

我们的编辑部
Your Linux servers, as a desktop.
TermalOSSponsored
Ops, reimagined

Your Linux servers, as a desktop.

Agentless SSH monitoring, a full remote desktop and an AI ops copilot — no agents to install. Everything stays on your machine.

SSHMonitoringAI Ops
Get early access
这篇文章对您有帮助吗?

9 人赞了这篇文章

P
Priya Raman机器学习工程师
🇨🇳 机器学习工程师,应用研究
分享:
评论 (9)

登录后即可参与讨论。

sandrine.b 14 Aug 2026 · 06:09

If WeChat’s AI is already deployed at this scale without transparency, isn’t the real question whether we even need benchmarks at this point or just better oversight?

BookWorm88 13 Aug 2026 · 09:10

Silent scaling to 617B without disclosure feels like a tech arms race where users are the guinea pigs. Where’s the middle ground between innovation and accountability?

Alex_LDN 13 Aug 2026 · 08:59

A model this big without metrics is like a black box-sure, it might work for a billion users, but how do we trust it’s not just hype?

ArtLover99 13 Aug 2026 · 11:17

Right, but 617B parameters could just mean wasted compute without transparency-how do we know it’s not overfit for WeChat’s niche use cases?

SkepticSam 13 Aug 2026 · 08:43

617B parameters is impressive, but without benchmarks or transparency, how do we know it's actually useful for users? Just deploying at scale doesn't guarantee real performance or safety.

HistoryBuff 13 Aug 2026 · 08:34

Seems like Tencent’s playing both sides-leveraging cutting-edge tech behind the scenes while keeping the rest of us in the dark. Still, a billion-user litmus test might say more than any obscure benchmark ever could.

BookWorm47 13 Aug 2026 · 08:29

617B params without benchmarks is like buying a sports car without a speedometer - flashy, but who really knows if it performs? Still, billion-user deployment says something.

J.P.R. 2 13 Aug 2026 · 08:25

Is Tencent’s bet on secret scaling a sign they’re chasing Moore’s Law at all costs, or proof that closed models can outperform open ones in real-world conditions?

Emma_London 13 Aug 2026 · 08:23

The focus should be on whether users actually benefit from this secrecy. Transparency in AI isn’t just for trust-it shapes what gets built next. What’s the endgame here?

GreenThumb 13 Aug 2026 · 08:22

What if raw scale without transparency is just hype? If they’re not sharing benchmarks, how do we know it’s not just marketing without substance?

事件时间线

Économie de l'open frontier : viabilité, subvention, pivots

  1. 1“6个月可活”:开放模型的窗口正在关闭13/07/2026
  2. 2反射签约100万美元计算能力合同:开放权重购买了一家工厂14/07/2026
  3. 3德朗格:真正的竞争可能不再在边境14/07/2026
  4. 4DeepSeek 瞄准市场:预计最早2026年在中国大陆提交IPO申请15/07/2026
  5. 5DeepSeek 估值 519 亿美元:中国前沿开源权重价格再次上涨17/07/2026
  6. 6Mozilla发布《开源AI现状报告》:生态系统一直需要的参考文档17/07/2026
  7. 7DeepSeek V4 即将推出:100万 tokens 上下文,双倍价格20/07/2026
  8. 8本·汤普森提出了一个战略问题:谁害怕中国模式?20/07/2026
  9. 9三星瞄准11亿美元投资Mistral:开放模式经济故事的战略-企业层面22/07/2026
  10. 10Hugging Face 曾经剥削女性和儿童 - 开放权重的代价现在在平台上28/07/2026
  11. 11Altman:人工智能垄断将是一场“长期灾难”29/07/2026
  12. 12DeepSeek V4-Flash-0731 进入公开测试:中国竞争对手引入 Codex 协议31/07/2026
  13. 13MiniMax H3开源:中国实验室打破全模态价格31/07/2026
  14. 14DeepSeek 发出重大 API 价格上涨信号——低于成本时代即将终结06/08/2026
  15. 15阿里巴巴计划为其下一代Qwen模型制定收入分成条款——开放权重经济学转变07/08/2026
  16. 16DeepSeek 以740亿美元的估值恢复融资:开放模型经济学达到新高度07/08/2026
  17. 17Meta 推出 Muse Glimmer:一款专为本地智能体AI打造的300亿参数开放权重编码模型10/08/2026
  18. 18Qwen 3.8 Max:2.4万亿参数,100万token上下文 - 阿里巴巴的开放权重前沿赌注变得更大10/08/2026
  19. 19微信小微代理运行在一个秘密的617B MoE模型上——腾讯的低调前沿赌注13/08/2026
  20. 20布鲁斯·施奈尔:如果市场拒绝OpenAI和Anthropic,美国应将其国有化14/08/2026
Your Linux servers, as a desktop.
TermalOSSponsored
Ops, reimagined

Your Linux servers, as a desktop.

Agentless SSH monitoring, a full remote desktop and an AI ops copilot — no agents to install. Everything stays on your machine.

Get early access
主题
浏览
信息