Learning Adaptive Search with Reinforcement Learning for Small and Fast Object Tracking

The persistent challenge in visual object tracking, particularly for small and fast-moving targets, lies in the trade-off between effective resolution and contextual information. Fixed search regions cannot adapt to variations in target scale and motion, often resulting in degraded target representation and tracking failures. In this paper, we introduce AdaSAM2, a framework that formulates adaptive search-region selection as a sequential decision-making problem. Unlike conventional per-frame heuristics, our approach employs an event-driven reinforcement learning policy that selects the cropping scale only during initialization and unreliable tracking states, while reusing the previous configuration during stable tracking. To handle target disappearance, we further introduce a lost-aware recovery mechanism that combines progressive search-region enlargement with constrained policy re-selection. Extensive experiments on TSFMO, LaTOT, UAV123, and UAVDT demonstrate consistent improvements over the SAMITE baseline. For example, AdaSAM2 improves Success and Precision on TSFMO from 42.6% and 74.0% to 43.9% and 75.3%, respectively, while improving UAV123 Success and Precision from 69.6% and 92.7% to 71.1% and 94.8%. Moreover, the RL policy is activated on only 0.74% of processed frames on TSFMO, resulting in an average overhead of only 0.0085 ms per frame. These results demonstrate that adaptive input-space optimization can improve tracking accuracy while introducing negligible computational overhead.

Authors

Institutions

Publication Details

Journal
Sensors
Published
2026-09-14
DOI
https://doi.org/10.3390/s26185819
Primary Topic
Video Surveillance and Tracking Methods
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Learning Adaptive Search with Reinforcement Learning for Small and Fast Object Tracking

Xinyi Bo, Binrui Liu, Jingqi Wang, Haolun Li et al.
Sensors
Video Surveillance and Tracking Methods
article

Learning Adaptive Search with Reinforcement Learning for Small and Fast Object Tracking

Xinyi Bo, Binrui Liu, Jingqi Wang, Haolun Li, Wenbin Luo, Shuiwang Li, Ge Zheng
article en

Abstract

The persistent challenge in visual object tracking, particularly for small and fast-moving targets, lies in the trade-off between effective resolution and contextual information. Fixed search regions cannot adapt to variations in target scale and motion, often resulting in degraded target representation and tracking failures. In this paper, we introduce AdaSAM2, a framework that formulates adaptive search-region selection as a sequential decision-making problem. Unlike conventional per-frame heuristics, our approach employs an event-driven reinforcement learning policy that selects the cropping scale only during initialization and unreliable tracking states, while reusing the previous configuration during stable tracking. To handle target disappearance, we further introduce a lost-aware recovery mechanism that combines progressive search-region enlargement with constrained policy re-selection. Extensive experiments on TSFMO, LaTOT, UAV123, and UAVDT demonstrate consistent improvements over the SAMITE baseline. For example, AdaSAM2 improves Success and Precision on TSFMO from 42.6% and 74.0% to 43.9% and 75.3%, respectively, while improving UAV123 Success and Precision from 69.6% and 92.7% to 71.1% and 94.8%. Moreover, the RL policy is activated on only 0.74% of processed frames on TSFMO, resulting in an average overhead of only 0.0085 ms per frame. These results demonstrate that adaptive input-space optimization can improve tracking accuracy while introducing negligible computational overhead.

SensorsVol. 26(18)
Guilin University of Technology (CN), Guangxi Science and Technology Department (CN)
Peace, Justice and strong institutions
Openalex Percentile: Top 13%
Video Surveillance and Tracking Methods
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.