Latest Research in Computer Vision and Pattern Recognition
69 research papers · 2026 median publication year
Top Research Topics in Computer Vision and Pattern Recognition
- Computation and Language — 25 papers
- Artificial Intelligence — 19 papers
- Computer Vision and Pattern Recognition — 13 papers
- Information Retrieval — 3 papers
- Robotics — 2 papers
- Machine Learning — 2 papers
- Neurobiology of Language and Bilingualism — 1 papers
- Multimodal Machine Learning Applications — 1 papers
- Multisensory perception and integration — 1 papers
- Computational Engineering, Finance, and Science — 1 papers
Highest-Cited Papers
- Search-to-World: Evaluation of 3D World Delivery from User Request through Web Search
- Clueing up LLMs with Tool-Augmented Deductive Reasoning
- When Faster VLA Deployment Changes Closed-Loop Behavior: Task Success-Latency Analysis of SmolVLA Across PyTorch and ONNX Variants
- What Do Hallucinations Reveal About Multimodal Reasoning? Diagnosing Visual Grounding Failures via Contrastive Decoding Probes
- Counterfactual Reasoning for Robust Visual Question Answering
- Layers, Sinks, and Scaling: Adaptive Evidence Selection for Multimodal Large Language Models
- Semantic-Spatial Agreement Verification for Mitigating Object Hallucination in Multimodal Large Language Models
- VisInteract: Towards Dynamic Interactive Text-to-Visualization under Imperfect Queries
- Spatial Reasoning via Modality Switching Between Language and Symbolic Representations
- Understanding the Effects of Distractors on Reasoning Vision-Language Models
- Learning Steerable Clarification Policies with Collaborative Self-play
- A Cohort-Level Floor with Limits to Fixed-Detector Transfer: A Pre-Registered Study of Calibrated Early-Response Geometry for Input-Label Discrimination Across Ten Language Models, with a Registered Six-Task Extension
- LifeMem: Enabling Lifelong Experience Reuse for LLM Agents
- KuaiRP Series Role-playing Models Technical Report
- SocialRL: Refining LLMs' Social Intelligence through Multi-turn Reinforcement Learning and Reward Design
- Beyond Surface Imitation: Contrastive Modeling for Reasoning Path Alignment in Multimodal In-Context Learning
- SRPO: Setwise Relative Policy Optimization for Multi-Agent LLMs
- Evolution of Multimodal Question Answering: From Modality-Adaptive Extraction to Unified Language Representation
- DSAEval: Evaluating Data Science Agents on a Wide Range of Real-World Data Science Problems
- Unlocking Multimodal Document Intelligence: From Current Triumphs to Future Frontiers of Visual Document Retrieval