vox2vec: A Framework for Self-supervised Contrastive Learning of Voxel-level Representations in Medical Images

Goncharov, Mikhail; Soboleva, Vera; Kurmukov, Anvar; Pisov, Maxim; Belyaev, Mikhail

Computer Science > Computer Vision and Pattern Recognition

arXiv:2307.14725 (cs)

[Submitted on 27 Jul 2023]

Title:vox2vec: A Framework for Self-supervised Contrastive Learning of Voxel-level Representations in Medical Images

Authors:Mikhail Goncharov, Vera Soboleva, Anvar Kurmukov, Maxim Pisov, Mikhail Belyaev

View PDF

Abstract:This paper introduces vox2vec - a contrastive method for self-supervised learning (SSL) of voxel-level representations. vox2vec representations are modeled by a Feature Pyramid Network (FPN): a voxel representation is a concatenation of the corresponding feature vectors from different pyramid levels. The FPN is pre-trained to produce similar representations for the same voxel in different augmented contexts and distinctive representations for different voxels. This results in unified multi-scale representations that capture both global semantics (e.g., body part) and local semantics (e.g., different small organs or healthy versus tumor tissue). We use vox2vec to pre-train a FPN on more than 6500 publicly available computed tomography images. We evaluate the pre-trained representations by attaching simple heads on top of them and training the resulting models for 22 segmentation tasks. We show that vox2vec outperforms existing medical imaging SSL techniques in three evaluation setups: linear and non-linear probing and end-to-end fine-tuning. Moreover, a non-linear head trained on top of the frozen vox2vec representations achieves competitive performance with the FPN trained from scratch while having 50 times fewer trainable parameters. The code is available at this https URL .

Comments:	MICCAI 2023
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2307.14725 [cs.CV]
	(or arXiv:2307.14725v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2307.14725

Submission history

From: Mikhail Goncharov [view email]
[v1] Thu, 27 Jul 2023 09:30:22 UTC (7,102 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:vox2vec: A Framework for Self-supervised Contrastive Learning of Voxel-level Representations in Medical Images

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:vox2vec: A Framework for Self-supervised Contrastive Learning of Voxel-level Representations in Medical Images

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators