Official code of the paper "ReDet: A Rotation-equivariant Detector for Aerial Object Detection" (CVPR 2021)

Last update: Dec 23, 2022

Overview

ReDet: A Rotation-equivariant Detector for Aerial Object Detection

ReDet: A Rotation-equivariant Detector for Aerial Object Detection (CVPR2021),
Jiaming Han^*, Jian Ding^*, Nan Xue, Gui-Song Xia^†,
arXiv preprint (arXiv:2103.07733).

The repo is based on AerialDetection and mmdetection. AerialDetection is a powerful framework for object detection in aerial images, which contains a lot of useful algorithms and tools.

Introduction

Recently, object detection in aerial images has gained much attention in computer vision. Different from objects in natural images, aerial objects are often distributed with arbitrary orientation. Therefore, the detector requires more parameters to encode the orientation information, which are often highly redundant and inefficient. Moreover, as ordinary CNNs do not explicitly model the orientation variation, large amounts of rotation augmented data is needed to train an accurate object detector. In this paper, we propose a Rotation-equivariant Detector (ReDet) to address these issues, which explicitly encodes rotation equivariance and rotation invariance. More precisely, we incorporate rotation-equivariant networks into the detector to extract rotation-equivariant features, which can accurately predict the orientation and lead to a huge reduction of model size. Based on the rotation-equivariant features, we also present Rotation-invariant RoI Align (RiRoI Align), which adaptively extracts rotation-invariant features from equivariant features according to the orientation of RoI. Extensive experiments on several challenging aerial image datasets DOTA-v1.0, DOTA-v1.5 and HRSC2016, show that our method can achieve state-of-the-art performance on the task of aerial object detection. Compared with previous best results, our ReDet gains 1.2, 3.5 and 2.6 mAP on DOTA-v1.0, DOTA-v1.5 and HRSC2016 respectively while reducing the number of parameters by 60% (313 Mb vs. 121 Mb).

Changelog

2021-03-09. Code released.

Benchmark and model zoo

ImageNet pretrain

We pretrain our ReResNet on the ImageNet-1K. Related codes can be found at the ReDet_mmcls branch. Here we provide our pretrained ReResNet-50 model for convenience. If you want to train and use ReResNet in your own project, please check out ReDet_mmcls for the installation and basic usage.

Model	Group	Top-1 (%)	Top-5 (%)	Download
ReR50	C₈	71.20	90.28	model \| log

Object Detection

Model	Data	Backbone	MS	Rotate	Lr schd	box AP	Download
ReDet	DOTA-v1.0	ReR50-FPN	-	-	1x	76.25	cfg model log
ReDet	DOTA-v1.0	ReR50-FPN	✓	✓	1x	80.10	cfg model log
ReDet	DOTA-v1.5	ReR50-FPN	-	-	1x	66.86	cfg model log
ReDet	DOTA-v1.5	ReR50-FPN	✓	✓	1x	76.80	cfg model log
ReDet	HRSC2016	ReR50-FPN	-	-	3x	90.46	cfg model log

If you cannot get access to Google Drive, BaiduYun download link can be found here with extracting code ABCD.

Installation

Please refer to INSTALL.md for installation and dataset preparation.

Getting Started

Please see GETTING_STARTED.md for the basic usage.

Citation

@inproceedings{han2021ReDet,
  author = {Han, Jiaming and Ding, Jian and Xue, Nan and Xia, Gui-Song},
  title = {ReDet: A Rotation-equivariant Detector for Aerial Object Detection},
  booktitle = {Proc. IEEE Conf. Computer Vision and Pattern Recognition (CVPR)},
  year = {2021}
}

Official code of the paper "ReDet: A Rotation-equivariant Detector for Aerial Object Detection" (CVPR 2021)

Related tags

Overview

ReDet: A Rotation-equivariant Detector for Aerial Object Detection

Introduction

Changelog

Benchmark and model zoo

Installation

Getting Started

Citation

Owner

csuhan

Sharing of contents on mitochondrial encounter networks

A Lighting Pytorch Framework for Recommendation System, Easy-to-use and Easy-to-extend.

PyTorch implementation of SMODICE: Versatile Offline Imitation Learning via State Occupancy Matching

RM Operation can equivalently convert ResNet to VGG, which is better for pruning; and can help RepVGG perform better when the depth is large.

PointCNN: Convolution On X-Transformed Points (NeurIPS 2018)

GLM (General Language Model)

《LightXML: Transformer with dynamic negative sampling for High-Performance Extreme Multi-label Text Classiﬁcation》(AAAI 2021) GitHub:

Official implementation for "QS-Attn: Query-Selected Attention for Contrastive Learning in I2I Translation" (CVPR 2022)

Efficient electromagnetic solver based on rigorous coupled-wave analysis for 3D and 2D multi-layered structures with in-plane periodicity

Generative Models as a Data Source for Multiview Representation Learning

Code and dataset for AAAI 2021 paper FixMyPose: Pose Correctional Describing and Retrieval Hyounghun Kim, Abhay Zala, Graham Burri, Mohit Bansal.

Automatic self-diagnosis program (python required)Automatic self-diagnosis program (python required)

Create and implement a deep learning library from scratch.

Malware Env for OpenAI Gym

AutoDeeplab / auto-deeplab / AutoML for semantic segmentation, implemented in Pytorch

Code for ACL'2021 paper WARP 🌀 Word-level Adversarial ReProgramming

PassAPI is a password generator in hash format and fully developed in Python, with the aim of teaching how to handle and build

✅ How Robust are Fact Checking Systems on Colloquial Claims?. In NAACL-HLT, 2021.

With this package, you can generate mixed-integer linear programming (MIP) models of trained artificial neural networks (ANNs) using the rectified linear unit (ReLU) activation function

Computer Vision and Pattern Recognition, NUS CS4243, 2022