6 ms·
Now GLM 5.3! And they explicitly confirms my theory: "Scaling post-training is all we did for GLM-5.3. With GLM-5.2 we built the stack: IndexShare for efficient
by sinuhe69 1mo ago
Now GLM 5.3! And they explicitly confirms my theory:
"Scaling post-training is all we did for GLM-5.3. With GLM-5.2 we built the stack: IndexShare for efficient long-context processing, SAO for RL on long-horizon tasks, and slime for large-scale asynchronous training — all running on the long-horizon task environments we have been accumulating. Over the past month we kept scaling on this stack: more environments, more diverse tasks, and more compute spent training on them."