Leveraging Human Selective Attention for Medical Image Analysis with Limited Training Data

12/02/2021
by   Yifei Huang, et al.
15

The human gaze is a cost-efficient physiological data that reveals human underlying attentional patterns. The selective attention mechanism helps the cognition system focus on task-relevant visual clues by ignoring the presence of distractors. Thanks to this ability, human beings can efficiently learn from a very limited number of training samples. Inspired by this mechanism, we aim to leverage gaze for medical image analysis tasks with small training data. Our proposed framework includes a backbone encoder and a Selective Attention Network (SAN) that simulates the underlying attention. The SAN implicitly encodes information such as suspicious regions that is relevant to the medical diagnose tasks by estimating the actual human gaze. Then we design a novel Auxiliary Attention Block (AAB) to allow information from SAN to be utilized by the backbone encoder to focus on selective areas. Specifically, this block uses a modified version of a multi-head attention layer to simulate the human visual search procedure. Note that the SAN and AAB can be plugged into different backbones, and the framework can be used for multiple medical image analysis tasks when equipped with task-specific heads. Our method is demonstrated to achieve superior performance on both 3D tumor segmentation and 2D chest X-ray classification tasks. We also show that the estimated gaze probability map of the SAN is consistent with an actual gaze fixation map obtained by board-certified doctors.

READ FULL TEXT

page 1

page 4

page 5

page 6

page 8

research
06/19/2020

Unified Representation Learning for Efficient Medical Image Analysis

Medical image analysis typically includes several tasks such as image en...
research
05/05/2021

CUAB: Convolutional Uncertainty Attention Block Enhanced the Chest X-ray Image Analysis

In recent years, convolutional neural networks (CNNs) have been successf...
research
07/31/2013

A Prototyping Environment for Integrated Artificial Attention Systems

Artificial visual attention systems aim to support technical systems in ...
research
10/12/2021

MEDUSA: Multi-scale Encoder-Decoder Self-Attention Deep Neural Network Architecture for Medical Image Analysis

Medical image analysis continues to hold interesting challenges given th...
research
09/28/2022

Deeply Supervised Layer Selective Attention Network: Towards Label-Efficient Learning for Medical Image Classification

Labeling medical images depends on professional knowledge, making it dif...
research
04/06/2022

Follow My Eye: Using Gaze to Supervise Computer-Aided Diagnosis

When deep neural network (DNN) was first introduced to the medical image...
research
07/13/2020

Learning and Exploiting Interclass Visual Correlations for Medical Image Classification

Deep neural network-based medical image classifications often use "hard"...

Please sign up or login with your details

Forgot password? Click here to reset