According to MarkTechPost, Z.ai and Qwen independently released models with similar architecture. The description mentions the ratio of linear components 3:1, compressed indexers, governed residual connections, and training with Muon.

They refer to GLM-5.3-Flash and Qwen3.8-Flash-Next. The source presents them as an example of independent convergence of architectural decisions between two Chinese AI labs.

The practical significance of such convergence cannot yet be assessed from available data: the package contains no test results, product details, or confirmations from Z.ai and Qwen themselves. Therefore, an open question remains whether these solutions offer comparable advantages in speed, quality, or cost.