Latest Research in Visual Foundation Grounding

39 research papers · 2026 median publication year

Top Research Topics in Visual Foundation Grounding

Highest-Cited Papers

  1. Can 4D Foundation Models Remember?
  2. CitySTAR: Structured and Topology-Aware Reasoning for Open-Vocabulary Urban 3D Grounding
  3. SpaRRTa: A Synthetic Benchmark for Evaluating Spatial Intelligence in Visual Foundation Models
  4. FAMOS: Feed-Forward 3D Articulation Modeling from Sparse Observations
  5. VLEM: Real-Time 3D Vision-Language Embedding Mapping
  6. GeoCond: A Conditioning-Aware Reliability Adapter for Feed-Forward 3D Reconstruction
  7. NormLift: From Lifted Features To Semantic Reliability In 3D Gaussian Splatting
  8. ParticleSplat: Self-supervised Object-centric Latent Particle Splatting
  9. DGSG-Mind: Dynamic 3D Gaussian Scene Graphs for Long-Term Scene Understanding and Grounding
  10. Human-aware Design Generation: Adding 3D Humans into Graphic Designs
  11. OCH3R: Object-Centric Holistic 3D Reconstruction
  12. P-POSEMEM: Projective Semantic Memory for Consistent Language Grounding under Pose-Graph Rewrites
  13. SceneBench: A Hierarchical Benchmark for Vision-Language Understanding of 3D Scenes
  14. ProClosure: Hierarchical Room-Object Assignment using Progressive Boundary Closure from Monocular Video
  15. LangStreet: Persistent Language Fields for Anchor-Decoded Street Gaussians
  16. SAMV-DUSt3R: Instance-Centric 3D Scene Decoupling from Sparse Multi-Views
  17. GoDeep: Annotation-Free Open-Vocabulary 3D Scene Understanding via Language-Space Lifting
  18. CoVeR: Coverage-Based Token Pruning for Multi-View 3D Reasoning in VLMs
  19. Point Cloud-Based Weld Seam Recognition and Localization for Robotic Welding
  20. Quality gated multimodal fusion improves scene construction and interaction experience in augmented reality digital media

Sub-Regions

Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
L2 Region - - 2026 Sep Q3

Visual Foundation Grounding

39 papers

Top Topics (5)

Computer Vision and Pattern Recognition29
Robotics7
Graphics1
Welding Techniques and Residual Stresses1
Augmented Reality Applications1

Top Publications (20)

1.Can 4D Foundation Models Remember?2.CitySTAR: Structured and Topology-Aware Reasoning for Open-Vocabulary Urban 3D Grounding3.SpaRRTa: A Synthetic Benchmark for Evaluating Spatial Intelligence in Visual Foundation Models4.FAMOS: Feed-Forward 3D Articulation Modeling from Sparse Observations5.VLEM: Real-Time 3D Vision-Language Embedding Mapping6.GeoCond: A Conditioning-Aware Reliability Adapter for Feed-Forward 3D Reconstruction7.NormLift: From Lifted Features To Semantic Reliability In 3D Gaussian Splatting8.ParticleSplat: Self-supervised Object-centric Latent Particle Splatting9.DGSG-Mind: Dynamic 3D Gaussian Scene Graphs for Long-Term Scene Understanding and Grounding10.Human-aware Design Generation: Adding 3D Humans into Graphic Designs11.OCH3R: Object-Centric Holistic 3D Reconstruction12.P-POSEMEM: Projective Semantic Memory for Consistent Language Grounding under Pose-Graph Rewrites13.SceneBench: A Hierarchical Benchmark for Vision-Language Understanding of 3D Scenes14.ProClosure: Hierarchical Room-Object Assignment using Progressive Boundary Closure from Monocular Video15.LangStreet: Persistent Language Fields for Anchor-Decoded Street Gaussians16.SAMV-DUSt3R: Instance-Centric 3D Scene Decoupling from Sparse Multi-Views17.GoDeep: Annotation-Free Open-Vocabulary 3D Scene Understanding via Language-Space Lifting18.CoVeR: Coverage-Based Token Pruning for Multi-View 3D Reasoning in VLMs19.Point Cloud-Based Weld Seam Recognition and Localization for Robotic Welding20.Quality gated multimodal fusion improves scene construction and interaction experience in augmented reality digital media

Sub-Regions (6)

AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.