Latest Research in Computer Vision and Pattern Recognition

69 research papers · 2026 median publication year

Top Research Topics in Computer Vision and Pattern Recognition

Highest-Cited Papers

  1. Search-to-World: Evaluation of 3D World Delivery from User Request through Web Search
  2. Clueing up LLMs with Tool-Augmented Deductive Reasoning
  3. When Faster VLA Deployment Changes Closed-Loop Behavior: Task Success-Latency Analysis of SmolVLA Across PyTorch and ONNX Variants
  4. What Do Hallucinations Reveal About Multimodal Reasoning? Diagnosing Visual Grounding Failures via Contrastive Decoding Probes
  5. Counterfactual Reasoning for Robust Visual Question Answering
  6. Layers, Sinks, and Scaling: Adaptive Evidence Selection for Multimodal Large Language Models
  7. Semantic-Spatial Agreement Verification for Mitigating Object Hallucination in Multimodal Large Language Models
  8. VisInteract: Towards Dynamic Interactive Text-to-Visualization under Imperfect Queries
  9. Spatial Reasoning via Modality Switching Between Language and Symbolic Representations
  10. Understanding the Effects of Distractors on Reasoning Vision-Language Models
  11. Learning Steerable Clarification Policies with Collaborative Self-play
  12. A Cohort-Level Floor with Limits to Fixed-Detector Transfer: A Pre-Registered Study of Calibrated Early-Response Geometry for Input-Label Discrimination Across Ten Language Models, with a Registered Six-Task Extension
  13. LifeMem: Enabling Lifelong Experience Reuse for LLM Agents
  14. KuaiRP Series Role-playing Models Technical Report
  15. SocialRL: Refining LLMs' Social Intelligence through Multi-turn Reinforcement Learning and Reward Design
  16. Beyond Surface Imitation: Contrastive Modeling for Reasoning Path Alignment in Multimodal In-Context Learning
  17. SRPO: Setwise Relative Policy Optimization for Multi-Agent LLMs
  18. Evolution of Multimodal Question Answering: From Modality-Adaptive Extraction to Unified Language Representation
  19. DSAEval: Evaluating Data Science Agents on a Wide Range of Real-World Data Science Problems
  20. Unlocking Multimodal Document Intelligence: From Current Triumphs to Future Frontiers of Visual Document Retrieval
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
L3 Region - - 2026 Sep Q3

Computer Vision and Pattern Recognition

69 papers

Top Topics (10)

Computation and Language25
Artificial Intelligence19
Computer Vision and Pattern Recognition13
Information Retrieval3
Robotics2
Machine Learning2
Neurobiology of Language and Bilingualism1
Multimodal Machine Learning Applications1
Multisensory perception and integration1
Computational Engineering, Finance, and Science1

Top Publications (20)

1.Search-to-World: Evaluation of 3D World Delivery from User Request through Web Search2.Clueing up LLMs with Tool-Augmented Deductive Reasoning3.When Faster VLA Deployment Changes Closed-Loop Behavior: Task Success-Latency Analysis of SmolVLA Across PyTorch and ONNX Variants4.What Do Hallucinations Reveal About Multimodal Reasoning? Diagnosing Visual Grounding Failures via Contrastive Decoding Probes5.Counterfactual Reasoning for Robust Visual Question Answering6.Layers, Sinks, and Scaling: Adaptive Evidence Selection for Multimodal Large Language Models7.Semantic-Spatial Agreement Verification for Mitigating Object Hallucination in Multimodal Large Language Models8.VisInteract: Towards Dynamic Interactive Text-to-Visualization under Imperfect Queries9.Spatial Reasoning via Modality Switching Between Language and Symbolic Representations10.Understanding the Effects of Distractors on Reasoning Vision-Language Models11.Learning Steerable Clarification Policies with Collaborative Self-play12.A Cohort-Level Floor with Limits to Fixed-Detector Transfer: A Pre-Registered Study of Calibrated Early-Response Geometry for Input-Label Discrimination Across Ten Language Models, with a Registered Six-Task Extension13.LifeMem: Enabling Lifelong Experience Reuse for LLM Agents14.KuaiRP Series Role-playing Models Technical Report15.SocialRL: Refining LLMs' Social Intelligence through Multi-turn Reinforcement Learning and Reward Design16.Beyond Surface Imitation: Contrastive Modeling for Reasoning Path Alignment in Multimodal In-Context Learning17.SRPO: Setwise Relative Policy Optimization for Multi-Agent LLMs18.Evolution of Multimodal Question Answering: From Modality-Adaptive Extraction to Unified Language Representation19.DSAEval: Evaluating Data Science Agents on a Wide Range of Real-World Data Science Problems20.Unlocking Multimodal Document Intelligence: From Current Triumphs to Future Frontiers of Visual Document Retrieval
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.