GLM-5.3-Flash vs Qwen3.8-Flash-Next: Two Chinese AI Labs Independently Converge on the Same Model Architecture

MarkTechPost
Read full post
Chinese AI labs Z.ai and Alibaba independently developed large multimodal models, GLM-5.3-Flash and Qwen3.8-Flash-Next, with remarkably similar architectures featuring a 3:1 ratio of linear to full attention layers and large context windows.

More on this story


More in Machine Learning

Machine Learning3 min read

Nvidia and Palantir fine-tune a 30B Nemotron model for Nvidia’s supply chain. It beats a model 18 times its size.

Covered by 3 sources
Machine Learning6 min read

CoreWeave Puts Field Engineers Inside Customer Teams for Physical AI

Covered by 2 sources
Machine Learning2 min read

Weatherwatch: AI model beats standard methods at predicting cyclones

The Guardian