Latest Research in Computer Vision and Pattern Recognition
159 research papers · 2026 median publication year
Top Research Topics in Computer Vision and Pattern Recognition
- Computer Vision and Pattern Recognition — 121 papers
- Generative Adversarial Networks and Image Synthesis — 9 papers
- Advanced Neural Network Applications — 8 papers
- Aesthetic Perception and Analysis — 5 papers
- Machine Learning — 3 papers
- Graphics — 3 papers
- Human Motion and Animation — 1 papers
- Robotics and Sensor-Based Localization — 1 papers
- Artificial Intelligence — 1 papers
- Human-Computer Interaction — 1 papers
Highest-Cited Papers
- Multimodal content generation and enhancement for smart imaging
- MudraGen: Geometrically Supervised Generation of Interacting Two-Hand Mudras for preserving Indian Classical Dance Heritage
- Paint-Anything: Unified Any-Color Control for Image Generation and Editing
- CapMap-MS-TTA: 3rd Place Solution for the MUMU Track of the 8th LSVOS Challenge at ECCV 2026
- DailyBench: A Unified Benchmark for AI-Generated and Manipulated Images from Modern Generative Models
- CompArt: Operationalizing Aesthetic Alignment in Text-to-Image Generation via Principles of Art
- JigSync: Gauge-Resolved Synchronization for Jigsaw Reassembly under Unknown Piece Orientation
- MDN-Control: Mask-Depth-Noise Guided Region Control for Multi-Subject Video Editing
- Structure-Detail Decoupled Autoregressive Generation for Fast and High-Fidelity Virtual Try-On
- A Generative AI Framework for Structural Analysis and DCGAN-Based Synthesis of Traditional Chinese Papercutting Patterns
- Text-Driven Artistic Staging: 3D Posing, Lighting, and Camera References from Paintings
- FSANet: Frequency-Spatial Aware Network for Image Segmentation
- FRPSS: Feature Rearrangement in Pre-Shape Space for Single-Image Generation
- VOR-Bench: A Human Perception-Driven Benchmark for Video Object Removal
- FROD: Feature Matching Residual Denoising Oracle Bone Decipher
- Opt-In Art: Learning Art Styles Only from Few Examples
- LynnReal-Omni: Native multi-modal Video Generation for Agentic Visual Workflows
- DG-SegNet: A depth-guided RGB-D semantic segmentation framework for emergency escape ramps toward traffic accident prevention
- PaintCopilot: Modeling Painting as Autonomous Artistic Continuation
- To Blend In, First Decouple: Rethinking Camouflage Image Generation via Context-Decoupled Representations