Topic
Everything filed under Reinforcement Learning, newest first.
RSS · JSON · All topics
Towards Universal Post-Training for Robotics
Perry Dong and Chelsea Finn on why robotics RL differs from LLM RL, what EXPO-FT gets right and wrong, and what a universal post-training recipe for real-world robots still needs.
15 min · 3,484 words