This repository provides the official implementation of 'Learning to ignore: rethinking attention in CNNs' accepted in BMVC 2021.

Last update: Jul 08, 2022

Overview

inverse_attention

This repository provides the official implementation of 'Learning to ignore: rethinking attention in CNNs' accepted in BMVC 2021.

Learning to ignore: rethinking attention in CNNs

Abstract:

Recently, there has been an increasing interest in applying attention mechanisms in Convolutional Neural Networks (CNNs) to solve computer vision tasks. Most of these methods learn to explicitly identify and highlight relevant parts of the scene and pass the attended image to further layers of the network. In this paper, we argue that such an approach might not be optimal. Arguably, explicitly learning which parts of the image are relevant is typically harder than learning which parts of the image are less relevant and, thus, should be ignored. In fact, in vision domain, there are many easy-to-identify patterns of irrelevant features. For example, image regions close to the borders are less likely to contain useful information for a classification task. Based on this idea, we propose to reformulate the attention mechanism in CNNs to learn to ignore instead of learning to attend. Specifically, we propose to explicitly learn irrelevant information in the scene and suppress it in the produced representation, keeping only important attributes. This implicit attention scheme can be incorporated into any existing attention mechanism. In this work, we validate this idea using two recent attention methods Squeeze and Excitation (SE) block and Convolutional Block Attention Module (CBAM). Experimental results on different datasets and model architectures show that learning to ignore, i.e., implicit attention, yields superior performance compared to the standard approaches.

Dependencies

The project was tested in Python 3 and Tensorflow 2. Run pip install -r requirements.txt to install dependent packages. Parts of the code are based on 'CBAM-keras'.

Running the code:

To test our approach on ImageNet, run main_imagenet.py. You need to: 1/ specify dataset_dir the TF-record directory of the dataset. 2/ choose the attention model to use, i.e., attention_module.

To test our approach on CIFAR10 or CIFAR100, run main_CIFAR.py. You need to: 1/ specify dataset and num_classes 2/ choose the attention model to use, i.e., attention_module.

Cite This Work

@article{laakom2021learning,
  title={Learning to ignore: rethinking attention in CNNs},
  author={Laakom, Firas and Chumachenko, Kateryna and Raitoharju, Jenni and Iosifidis, Alexandros and Gabbouj, Moncef},
  journal={arXiv preprint arXiv:2111.05684},
  year={2021}
}

This repository provides the official implementation of 'Learning to ignore: rethinking attention in CNNs' accepted in BMVC 2021.

Related tags

Overview

inverse_attention

Learning to ignore: rethinking attention in CNNs

Dependencies

Running the code:

Cite This Work

Owner

Firas Laakom

Python Implementation of the CoronaWarnApp (CWA) Event Registration

Cervix ROI Segmentation Using U-NET

Official code for "InfoGraph: Unsupervised and Semi-supervised Graph-Level Representation Learning via Mutual Information Maximization" (ICLR 2020, spotlight)

The source code of the paper "SHGNN: Structure-Aware Heterogeneous Graph Neural Network"

Two types of Recommender System : Content-based Recommender System and Colaborating filtering based recommender system

[CVPR 2020] GAN Compression: Efficient Architectures for Interactive Conditional GANs

Official repository for the CVPR 2021 paper "Learning Feature Aggregation for Deep 3D Morphable Models"

[BMVC'21] Official PyTorch Implementation of Grounded Situation Recognition with Transformers

Ground truth data for the Optical Character Recognition of Historical Classical Commentaries.

Learning Correspondence from the Cycle-consistency of Time (CVPR 2019)

Open-Set Recognition: A Good Closed-Set Classifier is All You Need

MLP-Numpy - A simple modular implementation of Multi Layer Perceptron in pure Numpy.

Codebase for arXiv preprint "NeRF++: Analyzing and Improving Neural Radiance Fields"

Official implementation for paper: A Latent Transformer for Disentangled Face Editing in Images and Videos.

MINERVA: An out-of-the-box GUI tool for offline deep reinforcement learning

Code for "Unsupervised Source Separation via Bayesian inference in the latent domain"

Affine / perspective transformation in Pose Estimation with Tensorflow 2

EM-POSE 3D Human Pose Estimation from Sparse Electromagnetic Trackers.

Pytorch version of SfmLearner from Tinghui Zhou et al.

An open source bike computer based on Raspberry Pi Zero (W, WH) with GPS and ANT+. Including offline map and navigation.