News Score: Score the News, Sort the News, Rewrite the Headlines

GLM-5.3: Frontier Coding with Emergent Cyber Capabilities

Scaling post-training is all we did for GLM-5.3. With GLM-5.2 we built the stack: IndexShare for efficient long-context processing, SAO for RL on long-horizon tasks, and slime for large-scale asynchronous training — all running on the long-horizon task environments we have been accumulating. Over the past month we kept scaling on this stack: more environments, more diverse tasks, and more compute spent training on them. Today we are releasing GLM-5.3. It uses the same base model as GLM-5.2 — eve...

Read more at z.ai

© News Score  score the news, sort the news, rewrite the headlines