Sort by
Refine Your Search
-
, offline RL, policy distillation and human-to-robot transfer. Possible directions include data-efficient learning, RL for Physical AI, continual adaptation, compact foundation models, edge-deployable VLAs
-
constraints. The project will explore orchestrating reasoning models, vision-language models, world models and efficient control policies for reliable robot behaviour. A key focus is deciding when to deliberate
Searches related to policy
Enter an email to receive alerts for policy positions