Relay dueling network for visual tracking with broad field-of-view

  • Jiang, Yifan
  • Han, David K.
  • Ko, Hanseok
Citations

WEB OF SCIENCE

8
Citations

SCOPUS

9

초록

A deep reinforcement-learning-based method is presented for visual object tracking tasks. The key objective is to generate a sequence of actions which can move or scale the bounding box in the previous frame to track the target in the current frame. Two intelligent agents are trained to accomplish the above task with a special dueling deep Q-learning network (Dueling DQN), referred to as a relay dueling network. The proposed model is divided into two agents: the movement agent and the scaling agent. The former performs horizontal or vertical movements and the latter generates scaling actions to change the size of the bounding box. The model has multiple inputs that cover both the bounding box region and the enlarged search region to improve the agents' perception of the surroundings. The proposed method has a broader field of vision than other similar trackers and its distribution of actions makes it easy to train and improve its tracking performance. The proposed network is tested on popular standard tracker benchmarks and its performance is compared with state-of-the-art trackers. The proposed network is found to be competitive in tracking accuracy and execution effectiveness when compared to conventional methods.

키워드

object trackingobject detectiontrackinglearning (artificial intelligence)computer visionrelay dueling networkvisual trackingfield-of-viewdeep reinforcement-learning-based methodvisual object tracking tasksprevious frameintelligent agentsspecial dueling deep Q-learning networkDueling DQNmovement agentscaling agentvertical movementsbounding box regionenlarged search regiontracking performance
제목
Relay dueling network for visual tracking with broad field-of-view
저자
Jiang, YifanHan, David K.Ko, Hanseok
DOI
10.1049/iet-cvi.2018.5546
발행일
2019-10
유형
Article
저널명
IET Computer Vision
13
7
페이지
615 ~ 622