Why robotics RL is a different problem than LLM RL, what EXPO-FT gets right and wrong, and what a universal post-training recipe for robotics needs. By Perry Dong, PhD student in Computer Science at Stanford University.

1 points•gmays•2 days ago•0 comments•

0 comments

No comments yet.

Related stories