Toward Advanced Deep Learning Techniques for Medical Image Analysis
Loading...
Date
Journal Title
Journal ISSN
Volume Title
Publisher
Université Sétif 1 - Ferhat ABBAS , Faculté des Sciences
Abstract
Deep learning (DL) has revolutionized medical diagnosis by enabling the analysis of diverse data modalities, particularly medical images, from a wide range of sources. This transformative capability has made DL approaches an increasingly vital tool in modern healthcare. Indeed, the potential of DL in medical image analysis is undeniably promising, with significant advancements already being made. However, only a few DL-based approaches have successfully transitioned into clinical practice. This may be due to various factors such as overfitted models, selection bias, and the extensive preprocessing of datasets, which fail to accurately represent clinical diversity and local variations.
This thesis aims to explore the transformative potential of DL in medical image analysis, focusing on the development of novel models that enhance diagnostic accuracy, efficiency, and personalization.
In this thesis, we address several key medical imaging modalities—X-rays, computed tomography (CT) scans, optical coherence tomography (OCT), and dermatoscopic imaging—to tackle critical diagnostic challenges in diseases such as COVID-19, retinal disorders, and skin lesions. This work offers a thorough exploration of advanced DL architectures and hybrid methodologies tailored for medical diagnostics. The first contribution underscores the effectiveness of attention mechanisms in achieving high-accuracy diagnoses for COVID-19 and other pulmonary diseases. Specifically, the integrated attention mechanisms resulted in notable improvements, with quantified performance metrics indicating high sensitivity and precision as high as 99% in COVID-19 diagnosis. The second introduces HTC-Retina, a cutting-edge hybrid approach that combines vision transformers (ViTs) with convolutional neural networks (CNNs). This approach, applied to OCT image classification, showed enhanced performance in identifying retinal disorders, with significant improvements in accuracy over traditional methods, achieving classification rates ranging from 97% to 99%. The third presents JILDYA-Net, a novel lightweight model optimized for skin lesion classification. This model employs an efficient architecture designed for resource-constrained clinical settings. It was evaluated on dermatoscopic images and demonstrated superior classification accuracy. The final contribution is the development of DA-UNet-Plus, a novel hybrid segmentation model. Initially applied to detect intraretinal fluid in OCT images with notable performance, DA-UNet-Plus was later adapted for skin lesion localization in dermatoscopic images, achieving improved lesion boundary delineation.
Overall, these contributions demonstrate the powerful potential of DL in automating complex diagnostic tasks and fostering personalized healthcare while paving the way for future innovations in medical image analysis.
Description
لقد أحدث التعلم العميق ثورة في تشخيص الأمراض من خلال تمكين تحليل أنواع متعددة من البيانات، وخاصة الصور الطبية، من مجموعة واسعة من المصادر. وقد جعلت هذه القدرة التحويلية أساليب التعلم العميق أداة حيوية بشكل متزايد في الرعاية الصحية الحديثة. إن الإمكانات التي يقدمها التعلم العميق في تحليل الصور الطبية واعدة بلا شك، مع تحقيق تقدم كبير بالفعل. ومع ذلك، فإن القليل من الأساليب المعتمدة على التعلم العميق قد انتقلت بنجاح إلى الممارسة السريرية. قد يرجع ذلك إلى عوامل متعددة مثل النماذج المفرطة التكيف، والانحياز في الاختيار، والإعداد المفرط لمجموعات البيانات، والتي لا تعكس التنوع السريري و المحلي بدقة.
تسعى هذه الأطروحة إلى استكشاف الإمكانات التحويلية للتعلم العميق في تحليل الصور الطبية، مع التركيز على تطوير نماذج جديدة تعزز دقة التشخيص وكفاءته وتخصيصه. في هذه الأطروحة، نتناول العديد من طرق التصوير الطبي الرئيسية: الأشعة السينية، التصوير المقطعي المحوسب، التصوير المقطعي البصري، والتصوير الجلدي، لمعالجة التحديات التشخيصية الحاسمة في الأمراض مثل، فيروس كورونا 2019، واضطرابات الشبكية، والآفات الجلدية. يقدم هذا العمل استكشافاً شاملاً لهياكل التعلم العميق والمنهجيات الهجينة المصممة خصيصاً للتشخيص الطبي. تسلط المساهمة الأولى الضوء على فعالية آليات الانتباه في تحقيق تشخيصات عالية الدقة لـكوفيد-19 وأمراض الرئة الأخرى. على وجه التحديد، أدت آليات الانتباه المدمجة إلى تحسينات ملحوظة، مع مؤشرات أداء تشير إلى حساسية ودقة تصل إلى 99% في تشخيص كوفيد-19. المساهمة الثانية تقدم HTC-Retina، وهو نهج هجيني متقدم يجمع بين المحولات البصرية والشبكات العصبية التلافيفية. أظهر هذا النهج، عند تطبيقه على تصنيف صورOCT، أداءً معززاً في تحديد اضطرابات الشبكية، مع تحسينات كبيرة في الدقة مقارنة بالطرق التقليدية، محققاً معدلات تصنيف تتراوح بين 97% و99%. المساهمة الثالثة تقدم JILDYA-Net، وهو نموذج خفيف الوزن جديد مُحسن لتصنيف الآفات الجلدية. يستخدم هذا النموذج بنية فعالة مصممة للأوضاع السريرية المحدودة الموارد. تم تقييمه على الصور الجلدية وأظهر دقة تصنيف متفوقة. المساهمة الأخيرة هي تطوير DA-UNet-Plus، وهو نموذج هجيني جديد للتقطيع. تم تطبيق DA-UNet-Plus في البداية لاكتشاف السوائل داخل الشبكية في صور OCT مع أداء ملحوظ، ثم تم تعديله لاحقاً لتحديد موقع الآفات في الصور الجلدية، محققاً تحسينات في تحديد حدود الآفات.
بوجه عام، تُظهر هذه المساهمات الإمكانات القوية للتعلم العميق في أتمتة المهام التشخيصية المعقدة وتعزيز الرعاية الصحية الشخصية، مما يمهد الطريق للابتكارات المستقبلية في تحليل الصور الطبية.
