article · Procedia Computer Science
Several challenges in computer vision prompted the research community to propose innovative approaches and unravel new perspectives to optimize deep learning models, thus enhancing the efficiency of vision tasks. Due to its growing applications and promising results in the NLP domain, Transformers models have inspired researchers to adapt this technology to computer vision problems by introducing the Vision Transformers networks. For this purpose, this paper provides a detailed comparative study of the properties of internal representations of Vision Transformers and Convolutional Neural Networks and presents some recent medical hybrid applications. On the other hand, the segmentation task poses various challenges due to the diversity of object position and size. Hence, development of new approaches and techniques is required. In this vein, we adapted the SegFormer model in order to segment the fetal head circumference. Our hybrid model provides a competitive and notable result with a Dice Coefficient of 93.54%, relying only on a small training dataset.
This page summarises published work. The authoritative version sits with the publisher.
DOI: 10.1016/j.procs.2024.11.120
Is something wrong with this record? Report it or request removal.
Discussion
Have you built on this work, tried to replicate it, or seen it applied in practice? Share what you know. Verified researchers and MARATTO™ domain experts can open a discussion, and any member can reply. Contributions are reviewed before they appear.
No discussion yet. Open the first thread.
New to MARATTO™? Create a free account.