Official implementation of deep-multi-trajectory-based single object tracking (IEEE T-CSVT 2021).

Last update: Dec 03, 2022

Overview

DeepMTA_PyTorch

Officical PyTorch Implementation of "Dynamic Attention-guided Multi-TrajectoryAnalysis for Single Object Tracking", Xiao Wang, Zhe Chen, Jin Tang, Bin Luo, Yaowei Wang, Yonghong Tian, Feng Wu, IEEE Transactions on Circuits and Systems for Video Technology (T-CSVT 2021) [Paper] [Project]

Abstract:

Most of the existing single object trackers track the target in a unitary local search window, making them particularly vulnerable to challenging factors such as heavy occlusions and out-of-view movements. Despite the attempts to further incorporate global search, prevailing mechanisms that cooperate local and global search are relatively static, thus are still sub-optimal for improving tracking performance. By further studying the local and global search results, we raise a question: can we allow more dynamics for cooperating both results? In this paper, we propose to introduce more dynamics by devising a dynamic attention-guided multi-trajectory tracking strategy. In particular, we construct dynamic appearance model that contains multiple target templates, each of which provides its own attention for locating the target in the new frame. Guided by different attention, we maintain diversified tracking results for the target to build multi-trajectory tracking history, allowing more candidates to represent the true target trajectory. After spanning the whole sequence, we introduce a multi-trajectory selection network to find the best trajectory that deliver improved tracking performance. Extensive experimental results show that our proposed tracking strategy achieves compelling performance on various large-scale tracking benchmarks.

Our Proposed Approach:

Install:

git clone https://github.com/wangxiao5791509/DeepMTA_PyTorch
cd DeepMTA_TCSVT_project

# create the conda environment
conda env create -f environment.yml
conda activate deepmta

# build the vot toolkits
bash benchmark/make_toolkits.sh

Download Dataset and Model:

download pre-trained Traj-Evaluation-Network [Onedrive] and Dynamic-TANet-Model [Onedrive]

get the dataset OTB2015, GOT-10k, LaSOT, UAV123, UAV20L, OxUvA from [List].

Download TNL2K dataset (published on CVPR 2021, 1300/700 for train and test subset) from: https://sites.google.com/view/langtrackbenchmark/

Train:

you can directly use the pre-trained tracking model of THOR [github];
train Dynamic Target-aware Attention:

cd ~/DeepMTA_TCSVT_project/trackers/dcynet_modules_adaptis/ 
python train.py

train Trajectory Evaluation Network:

python train_traj_measure_net.py

Tracking:

take got-10k and LaSOT dataset as the examples:

python testing.py -d GOT10k -t SiamRPN --lb_type ensemble

python testing.py -d LaSOT -t SiamRPN --lb_type ensemble

Benchmark Results:

Experimental results on the compared tracking benchmarks

[OTB2015] [LaSOT] [OxUvA] [GOT-10k] [UAV123] [TNL2K]

Tracking Results:

Tracking results on LaSOT dataset.

Tracking results on TNL2K dataset.

Attention prediciton and Tracking Results.

Acknowledgement:

Our tracker is developed based on THOR which is published on BMVC-2019 [Paper] [Code]

Other related works:

MTP: Multi-hypothesis Tracking and Prediction for Reduced Error Propagation, Xinshuo Weng, Boris Ivanovic, and Marco Pavone [Paper] [Code]
D.-Y. Lee, J.-Y. Sim, and C.-S. Kim, “Multihypothesis trajectory analysis for robust visual tracking,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 2015, pp. 5088–5096. [Paper]
C. Kim, F. Li, A. Ciptadi, and J. M. Rehg, “Multiple hypothesis tracking revisited,” in Proceedings of the IEEE International Conference on Computer Vision, 2015, pp. 4696–4704. [Paper]

Citation:

If you find this paper useful for your research, please consider to cite our paper:

@inproceedings{wang2021deepmta,
 title={Dynamic Attention guided Multi-Trajectory Analysis for Single Object Tracking},
 author={Xiao, Wang and Zhe, Chen and Jin, Tang and Bin, Luo and Yaowei, Wang and Yonghong, Tian and Feng, Wu},
 booktitle={IEEE Transactions on Circuits and Systems for Video Technology},
 doi={10.1109/TCSVT.2021.3056684}, 
 year={2021}
}

If you have any questions about this work, please contact with me via: [email protected] or [email protected]

Official implementation of deep-multi-trajectory-based single object tracking (IEEE T-CSVT 2021).

Related tags

Overview

DeepMTA_PyTorch

Officical PyTorch Implementation of "Dynamic Attention-guided Multi-TrajectoryAnalysis for Single Object Tracking", Xiao Wang, Zhe Chen, Jin Tang, Bin Luo, Yaowei Wang, Yonghong Tian, Feng Wu, IEEE Transactions on Circuits and Systems for Video Technology (T-CSVT 2021) [Paper] [Project]

Abstract:

Our Proposed Approach:

Install:

Download Dataset and Model:

Train:

Tracking:

Benchmark Results:

Tracking Results:

Tracking results on LaSOT dataset.

Tracking results on TNL2K dataset.

Attention prediciton and Tracking Results.

Acknowledgement:

Other related works:

Citation:

Owner

Xiao Wang（王逍）

ZEBRA: Zero Evidence Biometric Recognition Assessment

RepVGG: Making VGG-style ConvNets Great Again

C3D is a modified version of BVLC caffe to support 3D ConvNets.

[ICLR2021oral] Rethinking Architecture Selection in Differentiable NAS

AoT is a system for automatically generating off-target test harness by using build information.

Title: Heart-Failure-Classification

[ICCV 2021] Learning A Single Network for Scale-Arbitrary Super-Resolution

Awesome Graph Classification - A collection of important graph embedding, classification and representation learning papers with implementations.

GluonMM is a library of transformer models for computer vision and multi-modality research

UV matrix decompostion using movielens dataset

The official implementation of the paper, "SubTab: Subsetting Features of Tabular Data for Self-Supervised Representation Learning"

TSP: Temporally-Sensitive Pretraining of Video Encoders for Localization Tasks

Image-generation-baseline - MUGE Text To Image Generation Baseline

A package, and script, to perform imaging transcriptomics on a neuroimaging scan.

Simple and Robust Loss Design for Multi-Label Learning with Missing Labels

MWPToolkit is a PyTorch-based toolkit for Math Word Problem (MWP) solving.

CS50's Introduction to Artificial Intelligence Test Scripts

[ICLR 2022] Contact Points Discovery for Soft-Body Manipulations with Differentiable Physics

[ICCV'21] Neural Radiance Flow for 4D View Synthesis and Video Processing

Recurrent Scale Approximation (RSA) for Object Detection