Multi-label classification of retinal diseases based on fundus images using Resnet and Transformer.

Abstract

Retinal disorders are a major cause of irreversible vision loss, which can be mitigated through accurate and early diagnosis. Conventionally, fundus images are used as the gold diagnosis standard in detecting retinal diseases. In recent years, more and more researchers have employed deep learning methods for diagnosing ophthalmic diseases using fundus photography datasets. Among the studies, most of them focus on diagnosing a single disease in fundus images, making it still challenging for the diagnosis of multiple diseases. In this paper, we propose a framework that combines ResNet and Transformer for multi-label classification of retinal disease. This model employs ResNet to extract image features, utilizes Transformer to capture global information, and enhances the relationships between categories through learnable label embedding. On the publicly available Ocular Disease Intelligent Recognition (ODIR-5 k) dataset, the proposed method achieves a mean average precision of 92.86%, an area under the curve (AUC) of 97.27%, and a recall of 90.62%, which outperforms other state-of-the-art approaches for the multi-label classification. The proposed method represents a significant advancement in the field of retinal disease diagnosis, offering a more accurate, efficient, and comprehensive model for the detection of multiple retinal conditions.© 2024. International Federation for Medical and Biological Engineering.

Show Full Text

Multi-label classification of retinal diseases based on fundus images using Resnet and Transformer.

Researchers

Journal

Modalities

Models

Abstract

Glycosyltransferases: Mining, engineering and applications in biosynthesis of glycosylated plant natural products.

Assessing the added value of apparent diffusion coefficient, cerebral blood volume, and radiomic magnetic resonance features for differentiation of pseudoprogression versus true tumor progression in patients with glioblastoma.

Arterial Blood Pressure Estimation Method from Electrocardiogram Signals using U-Net.

A deep residual learning network for predicting lung adenocarcinoma manifesting as ground-glass nodule on CT images.

Image-based discrimination of the early stages of mesenchymal stem cell differentiation.

MTMC-AUR2CNet: Multi-textural multi-class attention recurrent residual convolutional neural network for COVID-19 classification using chest X-ray images.

Leave a Reply Cancel reply

Researchers

Journal

Modalities

Models

Abstract

Similar Posts

Leave a Reply Cancel reply