- Ph.D. Student in Mechanical Engineering at Hybrid Robotics Lab, BAIR, UC Berkeley
- Working on Reinforcement Learning, World Models, and Reasoning
Decision-Centric World Models, Safe and Robust RL, Scalable RL
- What is intelligence?
- How can agents achieve safety, robustness, stability, efficiency, and continual learning?
- Can agents performance a.s. monotone increase using any data stream? (= Is RL scalable?)
World Models, Offline RL, Off2On RL, Off-Policy RL, Stochastic Control, Reasoning