vox2vec: A Framework for Self-supervised Contrastive Learning of Voxel-level Representations in Medical Images

07/27/2023
by   Mikhail Goncharov, et al.
0

This paper introduces vox2vec - a contrastive method for self-supervised learning (SSL) of voxel-level representations. vox2vec representations are modeled by a Feature Pyramid Network (FPN): a voxel representation is a concatenation of the corresponding feature vectors from different pyramid levels. The FPN is pre-trained to produce similar representations for the same voxel in different augmented contexts and distinctive representations for different voxels. This results in unified multi-scale representations that capture both global semantics (e.g., body part) and local semantics (e.g., different small organs or healthy versus tumor tissue). We use vox2vec to pre-train a FPN on more than 6500 publicly available computed tomography images. We evaluate the pre-trained representations by attaching simple heads on top of them and training the resulting models for 22 segmentation tasks. We show that vox2vec outperforms existing medical imaging SSL techniques in three evaluation setups: linear and non-linear probing and end-to-end fine-tuning. Moreover, a non-linear head trained on top of the frozen vox2vec representations achieves competitive performance with the FPN trained from scratch while having 50 times fewer trainable parameters. The code is available at https://github.com/mishgon/vox2vec .

READ FULL TEXT

page 4

page 12

research
07/14/2020

Learning Semantics-enriched Representation via Self-discovery, Self-classification, and Self-restoration

Medical images are naturally associated with rich semantics about the hu...
research
05/17/2023

Self-Supervised Learning for Physiologically-Based Pharmacokinetic Modeling in Dynamic PET

Dynamic positron emission tomography imaging (dPET) provides temporally ...
research
01/05/2022

Advancing 3D Medical Image Analysis with Variable Dimension Transform based Supervised 3D Pre-training

The difficulties in both data acquisition and annotation substantially r...
research
10/14/2021

Inverse Problems Leveraging Pre-trained Contrastive Representations

We study a new family of inverse problems for recovering representations...
research
09/07/2021

Self-Supervised Representation Learning using Visual Field Expansion on Digital Pathology

The examination of histopathology images is considered to be the gold st...
research
11/25/2020

PGL: Prior-Guided Local Self-supervised Learning for 3D Medical Image Segmentation

It has been widely recognized that the success of deep learning in image...
research
03/23/2023

MV-JAR: Masked Voxel Jigsaw and Reconstruction for LiDAR-Based Self-Supervised Pre-Training

This paper introduces the Masked Voxel Jigsaw and Reconstruction (MV-JAR...

Please sign up or login with your details

Forgot password? Click here to reset