๐Ÿค– Awesome Loco-Manipulation [](https://awesome.re)

June 25, 2026 ยท View on GitHub

A curated list of resources on loco-manipulation โ€” the joint problem of locomotion and manipulation for legged, humanoid, wheeled-legged, and mobile robots.

This list is organized by methodology (model-based control & planning โ†’ learning-based control โ†’ foundation models & high-level reasoning), with thematic sub-sections inside each. Entries are tagged by robot embodiment and annotated with the research group / institution. Contributions welcome โ€” see Contributing.

๐Ÿ“– Two surveys anchor this taxonomy: Humanoid Locomotion and Manipulation (2501.02116, pipeline view) and A Survey of Behavior Foundation Models (2506.20487, foundation-model view).


๐Ÿงญ Legend

Embodiment tags: ๐Ÿง Humanoid ยท ๐Ÿพ Quadruped (+arm/leg) ยท ๐Ÿ›ž Wheeled-legged ยท ๐Ÿš— Mobile manipulator ยท ๐Ÿฆฟ General legged

Other: โญ = seminal / highly-cited ยท ๐Ÿ›๏ธ research group in italics


๐Ÿ“‘ Table of Contents


๐Ÿ“Š Surveys & Benchmarks

  • โญ [2025 arXiv] Humanoid Locomotion and Manipulation: Current Progress and Challenges in Control, Planning, and Learning [paper] ๐Ÿง Georgia Tech, USC, Stanford, NVIDIA, CMU, ...
  • โญ [2025 TPAMI] A Survey of Behavior Foundation Model: Next-Generation Whole-Body Control System of Humanoid Robots [paper] [repo] ๐Ÿง EIT Ningbo
  • [2023 Front. Mech. Eng.] Legged Robots for Object Manipulation: A Review [paper] ๐Ÿฆฟ

1. โš™๏ธ Model-Based Control & Planning

Whole-Body MPC for Loco-Manipulation

  • โญ [2021 RA-L/ICRA] A Unified MPC Framework for Whole-Body Dynamic Locomotion and Manipulation [paper] ๐Ÿพ ETH Zurich, RSL โ€” canonical unified whole-body MPC on ALMA (ANYmal + arm).
  • โญ [2019 RA-L] Whole-Body MPC for a Dynamically Stable Mobile Manipulator [paper] ๐Ÿฆฟ ETH Zurich, RSL
  • [2022 ICRA] A Collision-Free MPC for Whole-Body Dynamic Locomotion and Manipulation [paper] ๐Ÿพ ETH Zurich, RSL
  • [2022 IROS] Articulated Object Interaction in Unknown Scenes with Whole-Body Mobile Manipulation [paper] ๐Ÿพ ETH Zurich RSL + NVIDIA
  • [2024 RA-L] Tailoring Solution Accuracy for Fast Whole-Body MPC of Legged Robots [paper] ๐Ÿง MIT Biomimetics โ€” full-dynamics NMPC at 90 Hz on the MIT Humanoid.
  • [2024 arXiv] Real-Time Whole-Body Control of Legged Robots with MPPI Control [paper] [project] ๐Ÿฆฟ CMU, Robotic Exploration Lab
  • [2025 RA-L] Whole-Body Inverse Dynamics MPC for Legged Loco-Manipulation [paper] ๐Ÿพ ETH Zurich, CRL โ€” torque-level WB-MPC, 80 Hz on B2 + Z1.
  • [2024 ICRA] Hierarchical Optimization-Based Control for Whole-Body Loco-Manipulation of Heavy Objects [paper] ๐Ÿฆฟ

Multi-Contact Planning & Control

  • โญ [2023 Science Robotics] Versatile Multi-Contact Planning and Control for Legged Loco-Manipulation [paper] [journal] ๐Ÿพ ETH Zurich, RSL โ€” lightly-guided TAMP auto-generating whole-body trajectories + contact schedules.
  • [2024 CoRL] Guided Reinforcement Learning for Robust Multi-Contact Loco-Manipulation [paper] ๐Ÿพ ETH Zurich, RSL (Oral)
  • [2022 arXiv] Multi-Contact MPC for Dynamic Loco-Manipulation on Humanoid Robots [paper] ๐Ÿง USC, DRCL
  • [2023 Humanoids] Whole-Body MPC for Highly Redundant Legged Manipulators (37-DoF CENTAURO) [paper] ๐Ÿ›ž IIT, HHCM

Trajectory Optimization & Contact-Implicit Methods

  • โญ [2014 IJRR] A Direct Method for Trajectory Optimization of Rigid Bodies Through Contact [paper] ๐Ÿฆฟ MIT โ€” foundational contact-implicit trajectory optimization.
  • โญ [2024 IJRR] Contact-Implicit MPC: Controlling Diverse Quadruped Motions Without Pre-Planned Contact Modes [paper] ๐Ÿพ KAIST
  • [2021 IROS] Contact-Implicit Trajectory Optimization for Dynamic Object Manipulation [paper] ๐Ÿพ ETH Zurich, RSL
  • [2025 T-RO] Inverse Dynamics Trajectory Optimization for Contact-Implicit MPC [paper] [project] [code] ๐Ÿฆฟ Notre Dame / Toyota Research

Hierarchical Optimization for Heavy Objects

  • โญ [2024 ICRA] Hierarchical Optimization-Based Control for Whole-Body Loco-Manipulation of Heavy Objects [paper] ๐Ÿง USC, DRCL
  • [2024 arXiv] Safety-Critical Motion Planning for Collaborative Legged Loco-Manipulation over Discrete Terrain [paper] ๐Ÿฆฟ USC
  • [2026 arXiv] Sumo: Dynamic and Generalizable Whole-Body Loco-Manipulation [paper] [project] [code] ๐Ÿพ RAI Institute + CMU โ€” sample-based planner steering a whole-body policy (hybrid model-based).

Force / Impedance / Interaction Control

  • โญ [2021 ICRA] Model Predictive Robot-Environment Interaction Control for Mobile Manipulation Tasks [paper] ๐Ÿพ ETH Zurich, RSL โ€” force regulation against unknown environments (door opening).
  • [2024 IROS] Physically Consistent Online Inertial Adaptation for Humanoid Loco-Manipulation [paper] ๐Ÿง
  • [2023 IROS] Kinematically-Decoupled Impedance Control for Fast Visual Servoing and Grasping on Quadruped Manipulators [paper] ๐Ÿพ
  • [2023 Auton. Robots] RoLoMa: Robust Loco-Manipulation for Quadruped Robots with Arms [paper] ๐Ÿพ

Pattern Generation & Stabilization (Humanoid)

  • โญ [2003 ICRA] Biped Walking Pattern Generation by Using Preview Control of Zero-Moment Point [paper] ๐Ÿง AIST โ€” the seminal ZMP-preview walking pattern generator.
  • โญ [2015 T-RO] Three-Dimensional Bipedal Walking Control Based on Divergent Component of Motion [paper] ๐Ÿง DLR / TU Munich โ€” seminal DCM formulation.
  • โญ [2021 RA-L] Humanoid Loco-Manipulations Pattern Generation and Stabilization Control [paper] ๐Ÿง AIST / CNRS-AIST JRL
  • [2023 ICRA] Online Nonlinear Centroidal MPC for Humanoid Payload Carrying [paper] ๐Ÿง IIT (iCub)
  • [2023 arXiv] Dynamic Loco-Manipulation on HECTOR: Humanoid for Enhanced ConTrol and Open-source Research [paper] [code] ๐Ÿง USC, DRCL
  • [2025 arXiv] Whole-Body Control Framework for Humanoid Robots with Heavy Limbs [paper] ๐Ÿง CUHK
  • [2024 IROS] Physically Consistent Online Inertial Adaptation for Humanoid Loco-Manipulation [paper] ๐Ÿง

Wheeled-Legged & Quadruped-with-Arm

  • โญ [2019 ICRA] ALMA โ€” Articulated Locomotion and Manipulation for a Torque-Controllable Robot [paper] ๐Ÿพ ETH Zurich, RSL
  • โญ [2021 IROS] Whole-Body MPC and Online Gait Sequence Generation for Wheeled-Legged Robots [paper] ๐Ÿ›ž ETH Zurich, RSL
  • โญ [2019 RA-L] Keep Rollin' โ€” Whole-Body Motion Control and Planning for Wheeled Quadrupedal Robots [paper] ๐Ÿ›ž ETH Zurich, RSL
  • [2019 RA-L] Rolling in the Deep โ€” Hybrid Locomotion for Wheeled-Legged Robots [paper] ๐Ÿ›ž ETH Zurich, RSL
  • [2023 arXiv] Versatile Telescopic-Wheeled-Legged Locomotion of Tachyon 3 via Full-Centroidal NMPC [paper] ๐Ÿ›ž Sony Group
  • [2024 IROS] Arm-Constrained Curriculum Learning for Loco-Manipulation of the Wheel-Legged Robot [paper] [code] ๐Ÿ›ž
  • [2022 RA-L] Combining Learning-Based Locomotion Policy With Model-Based Manipulation for Legged Mobile Manipulators [paper] ๐Ÿพ
  • [2020 TRO] A Motion Planning Approach for Nonprehensile Manipulation and Locomotion Tasks of a Legged Robot [paper] ๐Ÿฆฟ

2. ๐Ÿง  Learning-Based Control

Whole-Body RL Control

  • โญ [2024 CoRL] Visual Whole-Body Control for Legged Loco-Manipulation (VBC) [paper] [code] ๐Ÿพ UC San Diego (Xiaolong Wang) (Oral)
  • โญ [2025 ICRA] HOVER: Versatile Neural Whole-Body Controller for Humanoid Robots [paper] [project] [code] ๐Ÿง NVIDIA GEAR + CMU + UC San Diego + UT Austin
  • [2024 RA-L] RoboDuet: Learning a Cooperative Policy for Whole-Body Legged Loco-Manipulation [paper] [project] [code] ๐Ÿพ Tsinghua / Shanghai AI Lab
  • [2025 RSS] HOMIE: Humanoid Loco-Manipulation with Isomorphic Exoskeleton Cockpit [paper] [project] [code] ๐Ÿง Shanghai AI Lab / OpenRobotLab + CUHK
  • [2025 arXiv] FALCON: Learning Force-Adaptive Humanoid Loco-Manipulation [paper] [project] [code] ๐Ÿง CMU LeCAR + NVIDIA
  • [2025 CoRL] Learning Unified Force and Position Control for Legged Loco-Manipulation [paper] [project] [code] ๐Ÿฆฟ
  • [2025 arXiv] Efficient Learning of a Unified Policy for Whole-Body Manipulation and Locomotion Skills [paper] ๐Ÿฆฟ
  • [2025 arXiv] Versatile Loco-Manipulation through Flexible Interlimb Coordination (RELIC) [paper] [code] ๐Ÿพ RAI Institute
  • [2025 arXiv] Embracing Bulky Objects with Humanoid Robots: Whole-Body Manipulation with RL [paper] ๐Ÿง HKUST(GZ)
  • [2024 CoRL] WoCoCo: Learning Whole-Body Humanoid Control with Sequential Contacts [project] [code] ๐Ÿง CMU LeCAR
  • [2024 IROS] HiLMa-Res: A Hierarchical Framework via Residual RL for Quadrupedal Locomotion and Manipulation [paper] ๐Ÿพ UC Berkeley + SFU
  • [2024 arXiv] Pedipulate: Enabling Manipulation Skills using a Quadruped Robot's Leg [paper] [project] ๐Ÿพ ETH Zurich, RSL
  • โญ [2022 CoRL] Deep Whole-Body Control: Learning a Unified Policy for Manipulation and Locomotion [paper] [code] ๐Ÿพ

Motion Tracking & Retargeting from Human Motion

  • โญ [2024 IROS] H2O: Learning Human-to-Humanoid Real-Time Whole-Body Teleoperation [paper] [project] ๐Ÿง CMU
  • โญ [2024 CoRL] OmniH2O: Universal and Dexterous Human-to-Humanoid Whole-Body Teleoperation and Learning [paper] [project] ๐Ÿง CMU + SJTU + Shanghai AI Lab
  • โญ [2024 RSS] ExBody: Expressive Whole-Body Control for Humanoid Robots [paper] [project] [code] ๐Ÿง UC San Diego (Xiaolong Wang) / MIT
  • [2024 arXiv] ExBody2: Advanced Expressive Humanoid Whole-Body Control [paper] [project] ๐Ÿง UC San Diego + UC Berkeley + MIT
  • โญ [2025 arXiv] ASAP: Aligning Simulation and Real-World Physics for Agile Humanoid Whole-Body Skills [paper] [project] [code] ๐Ÿง CMU + NVIDIA + UC Berkeley + UT Austin
  • โญ [2025 arXiv] SONIC: Supersizing Motion Tracking for Natural Humanoid Whole-Body Control [paper] [project] ๐Ÿง NVIDIA GEAR โ€” 42M-param, 700h-mocap motion-tracking foundation controller.
  • [2025 arXiv] GMT: General Motion Tracking for Humanoid Whole-Body Control [paper] [project] ๐Ÿง UC San Diego + SFU
  • [2025 ICRA] Mobile-TeleVision: Predictive Motion Priors for Humanoid Whole-Body Control [paper] [project] [code] ๐Ÿง UC San Diego + MIT
  • [2025 arXiv] BeyondMimic: From Motion Tracking to Versatile Humanoid Control via Guided Diffusion [paper] [project] ๐Ÿง UC Berkeley + Stanford
  • [2025 arXiv] UniTracker: Learning a Universal Whole-Body Motion Tracker for Humanoid Robots [paper] [project] [code] ๐Ÿง Shanghai AI Lab + ShanghaiTech + SJTU
  • [2025 arXiv] ResMimic: From General Motion Tracking to Whole-Body Loco-Manipulation via Residual Learning [paper] [project] [code] ๐Ÿง Amazon FAR + Stanford + UC Berkeley + CMU
  • [2025 arXiv] OmniRetarget: Interaction-Preserving Data Generation for Humanoid Loco-Manipulation [paper] [project] [code] ๐Ÿง Amazon FAR + Stanford + UC Berkeley + CMU
  • [2025 CVPR] TokenHSI: Unified Synthesis of Physical Human-Scene Interactions through Task Tokenization [project] ๐Ÿง
  • [2025 CVPR] Learning Physics-Based Full-Body Human Reaching and Grasping from Brief Walking References [project] ๐Ÿง
  • [2025 CoRL] MaskedManipulator: Versatile Whole-Body Control for Loco-Manipulation [paper] [project] [code] ๐Ÿง NVIDIA + SFU

Visual / Perceptive Loco-Manipulation

  • โญ [2025 CVPR] WildLMa: Long-Horizon Loco-Manipulation in the Wild [paper] [project] ๐Ÿพ UC San Diego (Xiaolong Wang) + MIT
  • [2024 CoRL] Visual Manipulation with Legs [paper] [project] ๐Ÿพ UC San Diego + CMU
  • [2025 arXiv] VisualMimic: Visual Humanoid Loco-Manipulation via Motion Tracking and Generation [paper] [project] [code] ๐Ÿง Stanford
  • [2025 arXiv] VIRAL: Visual Sim-to-Real at Scale for Humanoid Loco-Manipulation [paper] ๐Ÿง NVIDIA + CMU + UC Berkeley
  • [2025 arXiv] Opening the Sim-to-Real Door for Humanoid Pixel-to-Action Policy Transfer [paper] ๐Ÿง NVIDIA + CMU + UC Berkeley
  • [2024 CoRL] Learning to Manipulate Anywhere: A Visual Generalizable Framework for Reinforcement Learning [paper] ๐Ÿฆฟ
  • [2024 IROS] Learning Visual Quadrupedal Loco-Manipulation from Demonstrations [paper] [project] ๐Ÿพ UC Berkeley + Tsinghua
  • [2024 arXiv] Generalizable Humanoid Manipulation with 3D Diffusion Policies [paper] ๐Ÿง
  • [2024 arXiv] Learning Generalizable Feature Fields for Mobile Manipulation [paper] ๐Ÿš—
  • [2024 Humanoids] Perceptive Pedipulation with Local Obstacle Avoidance [paper] ๐Ÿพ ETH Zurich, RSL
  • [2024 CoRL] Learning to Open and Traverse Doors with a Legged Manipulator [paper] ๐Ÿพ ETH Zurich, RSL

Skill Blending / Hierarchical / Unified Policies

  • [2025 arXiv] SkillBlender: Towards Versatile Humanoid Whole-Body Loco-Manipulation via Skill Blending [paper] [project] [code] ๐Ÿง USC + UC Berkeley + Stanford/NVIDIA
  • [2025 arXiv] DemoHLM: From One Demonstration to Generalizable Humanoid Loco-Manipulation [paper] [project] [code] ๐Ÿง Peking University
  • [2024 CoRL] UMI on Legs: Making Manipulation Policies Mobile with Whole-Body Controllers [project] ๐Ÿพ
  • [2024 IROS] HYPERmotion: Learning Hybrid Behavior Planning for Autonomous Loco-Manipulation [paper] [project] ๐Ÿพ IIT
  • [2025 ICRA] Catch It! Learning to Catch in Flight with Mobile Dexterous Hands [project] ๐Ÿพ
  • [2024 arXiv] Helpful DoggyBot: Open-World Object Fetching using Legged Robots and VLMs [project] ๐Ÿพ UC San Diego + Stanford
  • [2024 arXiv] Playful DoggyBot: Learning Agile and Precise Quadrupedal Locomotion [project] ๐Ÿพ
  • [2024 arXiv] LocoMan: Advancing Versatile Quadrupedal Dexterity with Lightweight Loco-Manipulators [paper] [project] [code] ๐Ÿพ
  • [2024 IROS] On Learning Scene-Aware Generative State Abstractions for Task-Level Mobile Manipulation Planning [paper] [code] ๐Ÿš— ETH Zurich
  • [2024 IROS] BASENET: A Learning-Based Mobile Manipulator Base Pose Sequence Planning for Pickup Tasks [paper] ๐Ÿš—

Teleoperation & Data Collection

  • โญ [2024 CoRL] Mobile ALOHA: Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation [paper] [project] ๐Ÿš— Stanford
  • โญ [2024 CoRL] HumanPlus: Humanoid Shadowing and Imitation from Humans [paper] [project] ๐Ÿง Stanford
  • โญ [2025 CoRL] TWIST: Teleoperated Whole-Body Imitation System [paper] [project] ๐Ÿง Stanford + SFU
  • [2025 arXiv] TWIST2: Scalable, Portable, and Holistic Humanoid Data Collection System [paper] [code] ๐Ÿง Amazon FAR + Stanford + UC Berkeley + CMU
  • โญ [2024 CoRL] Open-TeleVision: Teleoperation with Immersive Active Visual Feedback [paper] [project] [code] ๐Ÿง UC San Diego + MIT
  • โญ [2023 RSS] AnyTeleop: A General Vision-Based Dexterous Arm-Hand Teleoperation System [paper] [project] [code] ๐Ÿฆพ NVIDIA + UC San Diego
  • [2024 CoRL] ACE: A Cross-Platform Visual-Exoskeletons System for Low-Cost Dexterous Teleoperation [paper] [project] [code] ๐Ÿฆพ UC San Diego
  • [2024 arXiv] Bunny-VisionPro: Real-Time Bimanual Dexterous Teleoperation for Imitation Learning [paper] [project] [code] ๐Ÿฆพ UC San Diego + HKU
  • [2024 IROS] GELLO: A General, Low-Cost, Intuitive Teleoperation Framework for Robot Manipulators [paper] [project] [code] ๐Ÿฆพ UC Berkeley
  • [2024 RSS] DexCap: Scalable Portable Mocap Data Collection for Dexterous Manipulation [paper] [project] [code] ๐Ÿฆพ Stanford
  • โญ [2024 RSS] DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset [paper] [project] [code] ๐Ÿฆพ multi-institution (Levine, Finn)
  • [2022 arXiv] TeLeMan: Teleoperation for Legged Robot Loco-Manipulation using Wearable IMU-based Motion Capture [paper] ๐Ÿพ

Diffusion Policies for Mobile / Legged Manipulation

  • โญ [2023 RSS] Diffusion Policy: Visuomotor Policy Learning via Action Diffusion [paper] [project] [code] ๐Ÿฆพ Columbia + Stanford + TRI + MIT
  • โญ [2024 RSS] 3D Diffusion Policy (DP3): Generalizable Visuomotor Policy Learning via Simple 3D Representations [paper] [project] [code] ๐Ÿฆพ Tsinghua / Shanghai Qi Zhi
  • [2024 ICRA] NoMaD: Goal Masked Diffusion Policies for Navigation and Exploration [paper] [project] [code] ๐Ÿš— UC Berkeley
  • [2024 ICML] DiffuseLoco: Real-Time Legged Locomotion Control with Diffusion from Offline Datasets [paper] [project] [code] ๐Ÿพ UC Berkeley
  • [2025 TPAMI] M2Diffuser: Diffusion-Based Trajectory Optimization for Mobile Manipulation in 3D Scenes [paper] [project] [code] ๐Ÿš— BIGAI + HUST
  • [2024 arXiv] Combining Planning and Diffusion for Mobility with Unknown Dynamics [paper] ๐Ÿš—

3. ๐ŸŒ Foundation Models & High-Level Reasoning

Behavioral Foundation Models (BFM)

  • โญ [2021 NeurIPS] Learning One Representation to Optimize All Rewards [paper] Meta FAIR โ€” original Forward-Backward representation.
  • โญ [2023 ICLR] Does Zero-Shot Reinforcement Learning Exist? [paper] Meta FAIR
  • โญ [2025 ICLR] Meta Motivo: Zero-Shot Whole-Body Humanoid Control via Behavioral Foundation Models [paper] [project] [code] ๐Ÿง Meta FAIR โ€” first humanoid BFM (FB-CPR).
  • [2025 arXiv] BFM-Zero: A Promptable Behavioral Foundation Model for Humanoid Control via Unsupervised RL [paper] [project] ๐Ÿง CMU + Meta FAIR โ€” sim-to-real on Unitree G1.
  • [2025 arXiv] Behavior Foundation Model for Humanoid Robots [paper] ๐Ÿง Shanghai AI Lab / InternRobotics
  • [2024 NeurIPS] Finer Behavioral Foundation Models via Auto-Regressive Features and Advantage Weighting [paper] Meta FAIR

Vision-Language-Action Models (VLA)

  • โญ [2025 arXiv] NVIDIA GR00T N1: An Open Foundation Model for Generalist Humanoid Robots [paper] [project] [code] ๐Ÿง NVIDIA GEAR โ€” open humanoid VLA (dual-system).
  • [2025 release] GR00T N1.5 (Isaac GR00T) [blog] ๐Ÿง NVIDIA GEAR
  • [2025 blog] Helix: A Vision-Language-Action Model for Generalist Humanoid Control [blog] ๐Ÿง Figure AI
  • โญ [2023 CoRL] RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control [paper] ๐Ÿฆพ Google DeepMind
  • โญ [2024 CoRL] OpenVLA: An Open-Source Vision-Language-Action Model [paper] [project] [code] ๐Ÿฆพ Stanford + UC Berkeley + Google DeepMind + TRI + MIT
  • [2024 RSS] Octo: An Open-Source Generalist Robot Policy [paper] [project] [code] ๐Ÿฆพ UC Berkeley + Stanford + CMU + DeepMind
  • โญ [2025 RSS] ฯ€โ‚€ (Pi0): A Vision-Language-Action Flow Model for General Robot Control [paper] [project] [code] ๐Ÿฆพ Physical Intelligence
  • [2025 arXiv] ฯ€โ‚€.โ‚… (Pi0.5): Open-World Generalization for Mobile Manipulation [paper] [project] ๐Ÿš— Physical Intelligence
  • [2025 ICLR] RDT-1B: A Diffusion Foundation Model for Bimanual Manipulation [paper] [project] [code] ๐Ÿฆพ Tsinghua TSAIL
  • [2025 arXiv] Gemini Robotics: Bringing AI into the Physical World [paper] ๐Ÿฆพ Google DeepMind
  • [2025 arXiv] LeVERB: Humanoid Whole-Body Control with Latent Vision-Language Instruction [paper] [project] ๐Ÿง UC Berkeley + CMU + SFU
  • [2025 RSS] NaVILA: Legged Robot Vision-Language-Action Model for Navigation [paper] [project] [code] ๐Ÿฆฟ UC San Diego + NVIDIA + USC
  • [2025 arXiv] Humanoid-VLA: Towards Universal Humanoid Control with Visual Integration [paper] ๐Ÿง Westlake MiLAB
  • [2024 ECCV] QUAR-VLA: Vision-Language-Action Model for Quadruped Robots [paper] [project] ๐Ÿพ Westlake MiLAB
  • [2025 arXiv] Galaxea G0: A Dual-System Vision-Language-Action Model [paper] ๐Ÿš— Galaxea AI + Tsinghua

VLM / LLM as High-Level Planner

  • โญ [2022 CoRL] SayCan: Do As I Can, Not As I Say โ€” Grounding Language in Robotic Affordances [paper] ๐Ÿš— Google / Everyday Robots
  • โญ [2023 ICML] PaLM-E: An Embodied Multimodal Language Model [paper] ๐Ÿฆพ Google + TU Berlin
  • [2023 CoRL] VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models [paper] [project] [code] ๐Ÿฆพ Stanford SVL
  • [2024 CoRL] ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints [paper] [project] [code] ๐Ÿš— Stanford SVL
  • [2023 CoRL] SayTap: Language to Quadrupedal Locomotion [paper] ๐Ÿพ Google DeepMind
  • โญ [2024 ICLR] Eureka: Human-Level Reward Design via Coding LLMs [paper] [project] [code] ๐Ÿฆฟ NVIDIA + UPenn + Caltech + UT Austin
  • [2024 RSS] DrEureka: Language Model Guided Sim-to-Real Transfer [paper] [project] [code] ๐Ÿพ UPenn + NVIDIA
  • [2024 IROS] HYPERmotion: Learning Hybrid Behavior Planning for Autonomous Loco-Manipulation [paper] [project] ๐Ÿพ IIT (also listed under Skill Blending)

World Models / World-Action Models (WAM)

  • โญ [2025 arXiv] Cosmos World Foundation Model Platform for Physical AI [paper] [code] NVIDIA
  • [2025 arXiv] Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control [paper] NVIDIA
  • [2025 arXiv] Cosmos-Reason1: From Physical Common Sense to Embodied Reasoning [paper] NVIDIA
  • [2026 arXiv] MotionWAM: Foundation World-Action Models for Real-Time Humanoid Loco-Manipulation [paper] ๐Ÿง Mondo Robotics + HKUST(GZ)
  • [2024 tech report] Genie 2: A Large-Scale Foundation World Model [blog] Google DeepMind
  • [2024 tech report] 1X World Model [blog] ๐Ÿง 1X Technologies

๐Ÿ›๏ธ Research Groups & Lineages

A few major threads run through the field โ€” useful for tracing how ideas evolved:

GroupFocusRepresentative lineage
NVIDIA GEAR / NVLabsHumanoid WBC + foundation modelsHOVER โ†’ ASAP โ†’ SONIC โ†’ GR00T N1/N1.5; ProtoMotions, MaskedMimic
CMU LeCAR Lab (Guanya Shi)Whole-body RL, sim-to-realH2O โ†’ OmniH2O โ†’ WoCoCo โ†’ FALCON โ†’ ASAP โ†’ BFM-Zero
UC San Diego (Xiaolong Wang)Expressive & visual loco-manipExBody โ†’ ExBody2 โ†’ GMT; VBC, WildLMa, Mobile-TeleVision
Stanford (C. Karen Liu, Chelsea Finn, Fei-Fei Li)Teleop, data, imitationMobile ALOHA, HumanPlus, TWIST, DexCap, VisualMimic, VoxPoser
Amazon FAR (Abbeel, Duan)Humanoid data & teleop at scaleTWIST2, OmniRetarget
ETH Zurich RSL/CRL (Marco Hutter)Model-based WB-MPC, multi-contactALMA โ†’ Unified MPC โ†’ Science Robotics multi-contact; wheeled-legged line
USC DRCL (Quan Nguyen)Model-based humanoid loco-manipHECTOR, hierarchical opt. for heavy objects
MIT (Biomimetics; Improbable AI)WB-MPC; learned force controlFast WB-MPC; Learning Force Control; SoftMimic
Meta FAIRBehavioral foundation modelsFB representation โ†’ Meta Motivo โ†’ BFM-Zero
Shanghai AI Lab / OpenRobotLabHumanoid systems & controlHOMIE, UniTracker, RoboDuet, BFM for Humanoids
Physical IntelligenceGeneralist VLAฯ€โ‚€ โ†’ ฯ€โ‚€.โ‚…

๐Ÿ’ก Note on "GRILITE": there is no NVIDIA model by that name โ€” it most likely refers to GR00T (NVIDIA's humanoid VLA), possibly conflated with the Fourier GR-1 humanoid that GR00T is demonstrated on.


๐Ÿค– Robot Models

๐Ÿพ Quadruped robots with arm

NameRobot BaseArmFormatsLicenseMeshesInertiasCollisions
Go2-ArxGo2Arx-L5-ProURDFBSD-3-Clauseโœ”๏ธโœ”๏ธโœ”๏ธ
B1-Z1Unitree B1Unitree Z1URDFBSD-3-Clauseโœ”๏ธโœ”๏ธโœ”๏ธ
Aliengo-Z1Unitree AliengoUnitree Z1URDFBSD-3-Clauseโœ”๏ธโœ”๏ธโœ”๏ธ

๐Ÿ›ž Wheel-legged robots with arm

NameRobot BaseArmFormatsLicenseMeshesInertiasCollisions
AirbotAirbotAirbot ArmURDF-โœ”๏ธโœ”๏ธโœ”๏ธ

๐Ÿš— Mobile robots with arm

NameRobot BaseArmFormatsLicenseMeshesInertiasCollisions
PR2PR2 robotPR2 robotURDFBSD-3โœ”๏ธโœ”๏ธโœ”๏ธ
Ridgeback-UR5ClearPath Ridgeback baseUR-5URDFApache 2.0โœ”๏ธโœ”๏ธโœ”๏ธ
Mabi-Mobile--URDF-โœ”๏ธโœ”๏ธโœ”๏ธ

โœจ Contributing

If you find any interesting papers or resources related to loco-manipulation, feel free to open an issue or submit a pull request!

๐Ÿ“ Format Guidelines

  • Section: Place the entry under the matching methodology sub-section (model-based / learning-based / foundation model).
  • Entry format: - **[Year Venue]** Title [[paper](url)] [[project](url)] [[code](url)] <embodiment emoji> *Group/Institution*
  • Venue abbreviations: Use standard ones (CoRL, ICRA, IROS, RSS, CVPR, arXiv, ...).
  • Embodiment: Tag with ๐Ÿง humanoid ยท ๐Ÿพ quadruped ยท ๐Ÿ›ž wheeled-legged ยท ๐Ÿš— mobile ยท ๐Ÿฆฟ general legged ยท ๐Ÿฆพ arm/manipulation-only.
  • Seminal works: Prefix with โญ if foundational / highly-cited.
  • Links: Include both paper and project/code links when available.

Last updated: June 2026 โ€” all arXiv IDs verified to resolve; project/code links HTTP-checked.