Learning Multi-axis Representation in Frequency Domain for Medical Image Segmentation

December 28, 2023 · Declared Dead · 🏛 Machine-mediated learning

"No code URL or promise found in abstract"

Evidence collected by the PWNC Scanner

Authors Jiacheng Ruan, Jingsheng Gao, Mingye Xie, Suncheng Xiang arXiv ID 2312.17030 Category eess.IV: Image & Video Processing Cross-listed cs.CV Citations 15 Venue Machine-mediated learning Last Checked 4 months ago

Abstract

Recently, Visual Transformer (ViT) has been extensively used in medical image segmentation (MIS) due to applying self-attention mechanism in the spatial domain to modeling global knowledge. However, many studies have focused on improving models in the spatial domain while neglecting the importance of frequency domain information. Therefore, we propose Multi-axis External Weights UNet (MEW-UNet) based on the U-shape architecture by replacing self-attention in ViT with our Multi-axis External Weights block. Specifically, our block performs a Fourier transform on the three axes of the input features and assigns the external weight in the frequency domain, which is generated by our External Weights Generator. Then, an inverse Fourier transform is performed to change the features back to the spatial domain. We evaluate our model on four datasets, including Synapse, ACDC, ISIC17 and ISIC18 datasets, and our approach demonstrates competitive performance, owing to its effective utilization of frequency domain information.

📄 View on arXiv 🌐 View on ar5iv 📑 PDF 🎉 Report Code Found

Community Contributions

Found the code? Know the venue? Think something is wrong? Let us know!

📜 Similar Papers

In the same crypt — Image & Video Processing

R.I.P. 👻 Ghosted

Variational image compression with a scale hyperprior

Johannes Ballé, David Minnen, ... (+3 more)

eess.IV 🏛 ICLR 📚 2.2K cites 8 years ago

📚 📚 The Cartographer

Deep Learning for Hyperspectral Image Classification: An Overview

Shutao Li, Weiwei Song, ... (+4 more)

eess.IV 🏛 IEEE TGRS 📚 1.5K cites 6 years ago

R.I.P. 👻 Ghosted

U-Net and its variants for medical image segmentation: theory and applications

Nahian Siddique, Paheding Sidike, ... (+2 more)

eess.IV 🏛 IEEE Access 📚 1.4K cites 5 years ago

R.I.P. 👻 Ghosted

Algorithm Unrolling: Interpretable, Efficient Deep Learning for Signal and Image Processing

Vishal Monga, Yuelong Li, Yonina C. Eldar

eess.IV 🏛 IEEE Signal Processing Magazine 📚 1.3K cites 6 years ago

R.I.P. 💀 404 Not Found

Lightweight Image Super-Resolution with Information Multi-distillation Network

Zheng Hui, Xinbo Gao, ... (+2 more)

eess.IV 🏛 ACM MM 📚 1.1K cites 6 years ago

R.I.P. 👻 Ghosted

Deep Learning on Image Denoising: An overview

Chunwei Tian, Lunke Fei, ... (+4 more)

eess.IV 🏛 Neural Networks 📚 941 cites 6 years ago

Died the same way — 👻 Ghosted

R.I.P. 👻 Ghosted

Federated Learning: Strategies for Improving Communication Efficiency

Jakub Konečný, H. Brendan McMahan, ... (+4 more)

cs.LG 🏛 arXiv 📚 5.2K cites 9 years ago

R.I.P. 👻 Ghosted

In-Datacenter Performance Analysis of a Tensor Processing Unit

Norman P. Jouppi, Cliff Young, ... (+73 more)

cs.AR 🏛 ISCA 📚 5.1K cites 9 years ago

R.I.P. 👻 Ghosted

Deep Convolutional Neural Networks for Computer-Aided Detection: CNN Architectures, Dataset Characteristics and Transfer Learning

Hoo-Chang Shin, Holger R. Roth, ... (+7 more)

cs.CV 🏛 IEEE TMI 📚 4.9K cites 10 years ago

R.I.P. 👻 Ghosted

Explanation in Artificial Intelligence: Insights from the Social Sciences

Tim Miller

cs.AI 🏛 AI 📚 4.9K cites 9 years ago