RLinf v0.3 closes the robot training loop
RLinf v0.3 turns embodied AI training into a fuller pipeline, linking data collection, SFT, reinforcement learning, evaluation and real-robot deployment.
7 verified stories covering Reinforcement learning, product updates and industry developments.
RLinf v0.3 turns embodied AI training into a fuller pipeline, linking data collection, SFT, reinforcement learning, evaluation and real-robot deployment.
Oak Lab sets a 20-watt goal for continuously learning agents. The piece reviews the verified facts and why the signal matters beyond one announcement.
The reinforcement learning pioneer argues that next-token models lack causality, experimentation and self-generated experience.
David Silver's Ineffable Intelligence has raised $1.1 billion in a seed round at a $5.1 billion valuation, marking the largest seed round in European history. The former DeepMind researcher plans to build a "superlearner" AI through reinforcement learning, diverging from the dominant large language model approach.
Thinking Machines Lab, founded by former OpenAI CTO Mira Murati, has secured a multibillion-dollar deal with Google Cloud for GB300 systems and reached a $12 billion valuation just 14 months after launch.
Sony AI's Ace robot has defeated elite amateur and professional table tennis players in real matches, marking the first time a machine has reached elite-level performance in a competitive sport.
Moonshot AI's K1.5 model, released before K2, demonstrates that pure reinforcement learning can enhance reasoning capabilities, validating the approach for scaling to larger models.