Latest Research in Embodied Vision-Language Control

794 research papers · 2026 median publication year

Top Research Topics in Embodied Vision-Language Control

Highest-Cited Papers

  1. Civil: causal and intuitive visual imitation learning
  2. Explicit Language Memory for Long-Horizon Planning in Vision-Language-Action Models
  3. Teach and Grow: An Agent-Centered Architecture for General Robot Learning
  4. PonderPounce: A Pretrained MLLM as an Episode Context Engine for Robot Control
  5. M2Tok: Multi-head Multi-codebook Discrete Action Tokenization for Vision-Language-Action Models
  6. MAGMA-GEN: Validated Recovery Supervision from Ambiguous Failures via Counterfactual Re-Execution
  7. A survey on large model-driven embodied grasping technologies
  8. Skeleton-guided geometry-aware auto-labeling for robotic grasping: From 4-DoF learning to 6-DoF execution
  9. Runtime Safety Filtering for Two-Terminal Hazards in Robotic Battery Recycling
  10. TraceFlow: Guiding Frozen Flow-Matching Robot Policies with Success and Failure Traces
  11. Robotic Video World Models: A Survey of Applications, Research Challenges, Future Directions
  12. CoRef-GS: Cooperative Referring Gaussian Splatting for Multi-Agent Scene Understanding
  13. Coding Agents with an Obstacle-Aware Harness for Safe Robot Manipulation
  14. HEROIC: Heterogeneous Evidential Reasoning for Open-Vocabulary Identification and Cross-Robot Collaboration
  15. ForwardDLO: Model-Based Bimanual Shape Matching of Unconstrained Deformable Linear Objects
  16. Recovering Aggressively Pruned Vision-Language-Action Models with Offline Hidden-State Distillation
  17. SceneTeract: Probing and Improving Agent-Aware Activity Reasoning in 3D Indoor Scenes
  18. HIL-UMI: Bringing Human-in-the-Loop Post-Training of Vision-Language-Action Models to Universal Manipulation Interface
  19. Graph-Aware Group Testing with Locally Clustered Infections
  20. Imagine-TAMP: Imagination-Guided Task and Motion Planning in Partial Observability

Sub-Regions

Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
L2 Region - - 2026 Sep Q3

Embodied Vision-Language Control

794 papers

Top Topics (10)

Robotics559
Computer Vision and Pattern Recognition107
Artificial Intelligence38
Machine Learning21
Robot Manipulation and Learning20
Multimodal Machine Learning Applications9
Reinforcement Learning in Robotics8
Computation and Language6
Systems and Control3
Human Pose and Action Recognition3

Top Publications (20)

1.Civil: causal and intuitive visual imitation learning2.Explicit Language Memory for Long-Horizon Planning in Vision-Language-Action Models3.Teach and Grow: An Agent-Centered Architecture for General Robot Learning4.PonderPounce: A Pretrained MLLM as an Episode Context Engine for Robot Control5.M2Tok: Multi-head Multi-codebook Discrete Action Tokenization for Vision-Language-Action Models6.MAGMA-GEN: Validated Recovery Supervision from Ambiguous Failures via Counterfactual Re-Execution7.A survey on large model-driven embodied grasping technologies8.Skeleton-guided geometry-aware auto-labeling for robotic grasping: From 4-DoF learning to 6-DoF execution9.Runtime Safety Filtering for Two-Terminal Hazards in Robotic Battery Recycling10.TraceFlow: Guiding Frozen Flow-Matching Robot Policies with Success and Failure Traces11.Robotic Video World Models: A Survey of Applications, Research Challenges, Future Directions12.CoRef-GS: Cooperative Referring Gaussian Splatting for Multi-Agent Scene Understanding13.Coding Agents with an Obstacle-Aware Harness for Safe Robot Manipulation14.HEROIC: Heterogeneous Evidential Reasoning for Open-Vocabulary Identification and Cross-Robot Collaboration15.ForwardDLO: Model-Based Bimanual Shape Matching of Unconstrained Deformable Linear Objects16.Recovering Aggressively Pruned Vision-Language-Action Models with Offline Hidden-State Distillation17.SceneTeract: Probing and Improving Agent-Aware Activity Reasoning in 3D Indoor Scenes18.HIL-UMI: Bringing Human-in-the-Loop Post-Training of Vision-Language-Action Models to Universal Manipulation Interface19.Graph-Aware Group Testing with Locally Clustered Infections20.Imagine-TAMP: Imagination-Guided Task and Motion Planning in Partial Observability

Sub-Regions (3)

AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.