THE BEST NEWS OF THE WEB
Lien affiliéAIRTOK Purificateur d'Air, H13 HEPAVoir l'offre Amazon fr(s'ouvre dans un nouvel onglet)
← Toutes les actus
AI

🇬🇧 Prime Intellect Releases prime-rl 0.6.0 to Train Trillion-Parameter MoE Models on Agentic RL Workloads

Prime Intellect has released prime-rl 0.6.0, an open framework for asynchronous reinforcement learning on trillion-parameter Mixture-of-Experts models. It trained GLM-5 on SWE tasks at up to 131k sequence length, with sub-5-minute step times and 256 rollouts, on 28 H200 nodes. This breakdown covers the inference and training optimizations behind those numbers — FP8 inference, …

1 vote / personne · anonyme

1 source