Vision Transformer

发布日期: 2022-10-12

2022-10-12 更新

Lightweight Transformer Backbone for Medical Object Detection

Authors:Yifan Zhang, Haoyu Dong, Nicholas Konz, Hanxue Gu, Maciej A. Mazurowski

Lesion detection in digital breast tomosynthesis (DBT) is an important and a challenging problem characterized by a low prevalence of images containing tumors. Due to the label scarcity problem, large deep learning models and computationally intensive algorithms are likely to fail when applied to this task. In this paper, we present a practical yet lightweight backbone to improve the accuracy of tumor detection. Specifically, we propose a novel modification of visual transformer (ViT) on image feature patches to connect the feature patches of a tumor with healthy backgrounds of breast images and form a more robust backbone for tumor detection. To the best of our knowledge, our model is the first work of Transformer backbone object detection for medical imaging. Our experiments show that this model can considerably improve the accuracy of lesion detection and reduce the amount of labeled data required in typical ViT. We further show that with additional augmented tumor data, our model significantly outperforms the Faster R-CNN model and state-of-the-art SWIN transformer model.
PDF

点此查看论文截图

木子已

https://ipaper.today/2022/10/12/2022-10-12-vision-transformer/

本博客所有文章除特別声明外，均采用 CC BY 4.0 许可协议。转载请注明来源木子已 !

Vision Transformer

检测/分割/跟踪

2022-10-12 检测/分割/跟踪

检测分割跟踪

Diffusion Models

2022-10-12 Diffusion Models

Diffusion Models

Vision Transformer

2022-10-12 更新

Lightweight Transformer Backbone for Medical Object Detection

打赏用于支持本站流量费