robotics
critique
neutral
Existing reinforcement-learning methods for robot crowd navigation are limited because they output a single reactive action at each timestep, which constrains their ability to represent diverse short-term avoidance strategies
Existing reinforcement-learning methods typically output a single reactive action at each timestep, which limits their ability to represent diverse short-term avoidance strategies.
Machine Learning30 Aug 2026