Dual‐View Representation Learning for Deployable Humanoid Whole‐Body Control

Humanoid whole‐body control policies receive limited proprioception at execution time, while simulation exposes richer privileged information during training. Such information can improve policy learning, but using it for deployable control requires transferring training‐only cues without changing the execution‐time sensing interface. While existing methods often rely on teacher distillation or explicit prediction of physical quantities, this work treats the transfer as a representation‐learning problem. To address this, DVP is introduced as a Dual‐View Proprioceptive (DVP) framework that turns privileged simulation states into training‐only latent supervision. For each simulated state, DVP pairs the privileged state with a view containing only the inputs available to the deployed actor. A dual‐cosine objective then aligns the privileged and deployable latent features while PPO optimizes the control policy. At execution time, the actor uses only deployment‐available inputs, and the privileged branch is removed. On LimX Oli velocity tracking and whole‐body motion tracking, DVP improves learning progress, control quality, action smoothness, and tracking accuracy over representative representation‐learning and privileged‐learning baselines. Cross‐platform motion‐tracking simulations on Unitree G1, Booster K1, and Noetix E1 test generality across morphology scales, while MuJoCo sim‐to‐sim evaluation and limited LimX Oli hardware demonstrations show execution through the deployable interface. Source code is available at DVP_code .

Authors

Institutions

Publication Details

Journal
Advanced Intelligent Systems
Published
2026-09-30
DOI
https://doi.org/10.1002/aisy.70563
Primary Topic
Prosthetics and Rehabilitation Robotics
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Dual‐View Representation Learning for Deployable Humanoid Whole‐Body Control

Wengang Zhou, Houqiang Li, Haolin Song, Mingxiao Feng
Advanced Intelligent Systems
Prosthetics and Rehabilitation Robotics
article

Dual‐View Representation Learning for Deployable Humanoid Whole‐Body Control

Wengang Zhou, Houqiang Li, Haolin Song, Mingxiao Feng
article en

Abstract

Humanoid whole‐body control policies receive limited proprioception at execution time, while simulation exposes richer privileged information during training. Such information can improve policy learning, but using it for deployable control requires transferring training‐only cues without changing the execution‐time sensing interface. While existing methods often rely on teacher distillation or explicit prediction of physical quantities, this work treats the transfer as a representation‐learning problem. To address this, DVP is introduced as a Dual‐View Proprioceptive (DVP) framework that turns privileged simulation states into training‐only latent supervision. For each simulated state, DVP pairs the privileged state with a view containing only the inputs available to the deployed actor. A dual‐cosine objective then aligns the privileged and deployable latent features while PPO optimizes the control policy. At execution time, the actor uses only deployment‐available inputs, and the privileged branch is removed. On LimX Oli velocity tracking and whole‐body motion tracking, DVP improves learning progress, control quality, action smoothness, and tracking accuracy over representative representation‐learning and privileged‐learning baselines. Cross‐platform motion‐tracking simulations on Unitree G1, Booster K1, and Noetix E1 test generality across morphology scales, while MuJoCo sim‐to‐sim evaluation and limited LimX Oli hardware demonstrations show execution through the deployable interface. Source code is available at DVP_code .

Advanced Intelligent Systems
University of Science and Technology of China (CN)
Openalex Percentile: Top 22%
Prosthetics and Rehabilitation Robotics
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.

Dual‐View Representation Learning for Deployable Humanoid Whole‐Body Control — Wengang Zhou, Houqiang Li, et al. · Advanced Intelligent Systems (2026) | TGRS Research Map | TGRS