Motor-Skill-Learning

Whole-Body Teleoperation and Data-Collection Infrastructure for Humanoids featured image

Whole-Body Teleoperation and Data-Collection Infrastructure for Humanoids

To support foundation-model learning on humanoids and quadruped robots, we develop whole-body teleoperation interfaces and data-collection environments. From affordable …

Multimodal Instruction Foundation Models with Mask Images featured image

Multimodal Instruction Foundation Models with Mask Images

We extend vision-language-action (VLA) foundation models with attention-guided mask images, making the correspondence between natural-language and visual instructions explicit so …

Foundation Models for Humanoid Motion Learning featured image

Foundation Models for Humanoid Motion Learning

We apply foundation models such as GPT to humanoid motor control, building unified policies that handle diverse motion tasks including locomotion and whole-body actions. By …

Embodiment Selection and Distance-Aware World Models for Mobile Manipulation featured image

Embodiment Selection and Distance-Aware World Models for Mobile Manipulation

We study world-model-based reinforcement learning for mobile manipulation that jointly addresses embodiment selection (when to move the base vs. use the arm) and motion planning. …

Autonomous Motion Learning via Large Language Models featured image

Autonomous Motion Learning via Large Language Models

Using large language models (LLMs) and vision-language models (VLMs), the robot autonomously generates demonstrations from task instructions, enabling motion-skill learning without …