Multi-Scale Aligned Distillation for Low-Resolution Detection (CVPR2021)

Last update: Dec 23, 2022

Related tags

Deep Learning MSAD

Overview

MSAD

Multi-Scale Aligned Distillation for Low-Resolution Detection

Lu Qi*, Jason Kuen*, Jiuxiang Gu, Zhe Lin, Yi Wang, Yukang Chen, Yanwei Li, Jiaya Jia

This project provides an implementation for the CVPR 2021 paper "Multi-Scale Aligned Distillation for Low-Resolution Detection" based on Detectron2. MSAD targets to detect objects using low-resolution instead of high-resolution image. MSAD could obtain comparable performance in high-resolution image size. Our paper use Slimmable Neural Networks as our pretrained weight.

Installation

This project is based on Detectron2, which can be constructed as follows.

Install Detectron2 following the instructions.
Setup the dataset following the structure.
Copy this project to /path/to/detectron2/projects/MSAD
Download the slimmable networks in the github. The slimmable resnet50 pretrained weight link is here.

Pretrained Weight

Move the pretrained weight to your target path
Modify the weight path in configs/Base-SLRESNET-FCOS.yaml

Teacher Training

To train teacher model with 8 GPUs, run:

cd /path/to/detectron2
python3 projects/MSAD/train_net_T.py --config-file <projects/MSAD/configs/config.yaml> --num-gpus 8

For example, to launch MSAD teacher training (1x schedule) with Slimmable-ResNet-50 backbone in 0.25 width on 8 GPUs and save the model in the path "/data/SLR025-50-T". one should execute:

cd /path/to/detectron2
python3 projects/MSAD/train_net_T.py --config-file projects/MSAD/configs/SLR025-50-T.yaml --num-gpus 8 OUTPUT_DIR /data/SLR025-50-T

Student Training

To train student model with 8 GPUs, run:

cd /path/to/detectron2
python3 projects/MSAD/train_net_S.py --config-file <projects/MSAD/configs/config.yaml> --num-gpus 8

For example, to launch MSAD student training (1x schedule) with Slimmable-ResNet-50 backbone in 0.25 width on 8 GPUs and save the model in the path "/data/SLR025-50-S". We assume the teacher weight is saved in the path "/data/SLR025-50-T/model_final.pth" one should execute:

cd /path/to/detectron2
python3 projects/MSAD/train_net_S.py --config-file projects/MSAD/configs/MSAD-R50-S025-1x.yaml --num-gpus 8 MODEL.WEIGHTS /data/SLR025-50-T/model_final.pth OUTPUT_DIR MSAD-R50-S025-1x

Evaluation

To evaluate a teacher or student pre-trained model with 8 GPUs, run:

cd /path/to/detectron2
python3 projects/MSAD/train_net_T.py --config-file <config.yaml> --num-gpus 8 --eval-only MODEL.WEIGHTS model_checkpoint

cd /path/to/detectron2
python3 projects/MSAD/train_net_S.py --config-file <config.yaml> --num-gpus 8 --eval-only MODEL.WEIGHTS model_checkpoint

Results

We provide the results on COCO val set with pretrained models. In the following table, we define the backbone FLOPs as capacity. For brevity, we regard the FLOPs of Slimmable Resnet50 in width 1.0 and high resolution input (800,1333) as 1x.

Method	Backbone	Capacity	Sched	Width	Role	Resolution	BoxAP	download
FCOS	Slimmable-R50	1.25x	1x	1.00	Teacher	H & L	42.8	model \| metrics
FCOS	Slimmable-R50	0.25x	1x	1.00	Student	L	39.9	model \| metrics
FCOS	Slimmable-R50	0.70x	1x	0.75	Teacher	H & L	41.2	model \| metrics
FCOS	Slimmable-R50	0.14x	1x	0.75	Student	L	38.8	model \| metrics
FCOS	Slimmable-R50	0.31x	1x	0.50	Teacher	H & L	38.4	model \| metrics
FCOS	Slimmable-R50	0.06x	1x	0.50	Student	L	35.7	model \| metrics
FCOS	Slimmable-R50	0.08x	1x	0.25	Teacher	H & L	33.2	model \| metrics
FCOS	Slimmable-R50	0.02x	1x	0.25	Student	L	30.3	model \| metrics

Citing MSAD

Consider cite MSAD in your publications if it helps your research.

@article{qi2021msad,
  title={Multi-Scale Aligned Distillation for Low-Resolution Detection},
  author={Lu Qi, Jason Kuen, Jiuxiang Gu, Zhe Lin, Yi Wang, Yukang Chen, Yanwei Li, Jiaya Jia},
  journal={IEEE Conference on Computer Vision and Pattern Recognition (CVPR)},
  year={2021}
}

Multi-Scale Aligned Distillation for Low-Resolution Detection (CVPR2021)

Related tags

Overview

MSAD

Installation

Pretrained Weight

Teacher Training

Student Training

Evaluation

Results

Citing MSAD

Owner

Jia Research Lab

Official repository for the paper "Going Beyond Linear Transformers with Recurrent Fast Weight Programmers"

这个开源项目主要是对经典的时间序列预测算法论文进行复现，模型主要参考自GluonTS，框架主要参考自Informer

Awesome Long-Tailed Learning

Repository for "Toward Practical Monocular Indoor Depth Estimation" (CVPR 2022)

PyTorch code for our ECCV 2018 paper "Image Super-Resolution Using Very Deep Residual Channel Attention Networks"

A pytorch reproduction of { Co-occurrence Feature Learning from Skeleton Data for Action Recognition and Detection with Hierarchical Aggregation }.

Quickly comparing your image classification models with the state-of-the-art models (such as DenseNet, ResNet, ...)

PyTorch-centric library for evaluating and enhancing the robustness of AI technologies

An implementation of Geoffrey Hinton's paper "How to represent part-whole hierarchies in a neural network" in Pytorch.

Codes to calculate solar-sensor zenith and azimuth angles directly from hyperspectral images collected by UAV. Works only for UAVs that have high resolution GNSS/IMU unit.

Official implement of Paper：A deeply supervised image fusion network for change detection in high resolution bi-temporal remote sening images

Neural network for digit classification powered by cuda

1st place solution to the Satellite Image Change Detection Challenge hosted by SenseTime

Video Matting via Consistency-Regularized Graph Neural Networks

Pytorch Implementation of "Desigining Network Design Spaces", Radosavovic et al. CVPR 2020.

A hue shift helper for OBS

Learning Neural Painters Fast! using PyTorch and Fast.ai

Numerai tournament example scripts using NN and optuna

Spatial Single-Cell Analysis Toolkit

The trained model and denoising example for paper : Cardiopulmonary Auscultation Enhancement with a Two-Stage Noise Cancellation Approach