The500Feed.Live

Everything going on in AI - updated daily from 500+ sources

← Back to The 500 Feed
🌐 NewsAugust 13, 2026

WeChat AI Team Details WeLM Models Scaling to 617B Parameters

Tencent’s WeChat AI team has detailed a new scaling approach for its WeLM model family. The team trained WeLM-HD4-80B and WeLM-HD4-617B models using a method called Hidden Decoding, which expands each token into multiple internal computation streams without increasing the main Transformer backbone. The 80B model activates 3 billion parameters, while the 617B model activates […]

Read Original Article →

Source

https://technode.com/2026/08/13/wechat-ai-team-details-welm-models-scaling-to-617b-parameters/