EXPLAINABLE HYBRID VISION TRANSFORMER AND GENERATIVE AI FRAMEWORK FOR REAL-TIME INTELLIGENT MEDICAL IMAGE DISEASE DIAGNOSIS
编号:127 访问权限:仅限参会人 更新:2026-07-22 16:11:10 浏览:11次 In-person

报告开始:2026年07月31日 14:25(Asia/Kolkata)

报告时间:15min

所在会场:[S4] Computer Vision and Pattern Recognition [S4-5] Computer Vision and Pattern Recognition

演示文件

提示:该报告下的文件权限为仅限参会人,您尚未登录,暂时无法查看。

摘要
Abstract - The use of medical image analysis is now an essential aspect of today’s healthcare systems due to early disease detection, accurate diagnosis, and sound clinical decision making. Nevertheless, current deep learning methods exhibit low interpretability, low generalization across different types of medical imaging datasets, and inability to perform in real time. Moreover, the lack of explainability in several convolutional and transformer models makes clinicians skeptical about their usage in critical healthcare settings. Therefore, this research offers a new Explainable Hybrid Vision Transformer with Generative Intelligence Network (EHVT-GIN) for real-time medical image analysis and diagnosis of diseases in an interpretable manner. The presented architecture combines Vision Transformer based global features extraction with local features learning using lightweight convolutional networks to maintain contextual and spatial information. The Generative AI module provides realistic augmentation of medical data samples to solve the class imbalance issue and increase feature diversity during training. The explainability component uses attention visualization and gradient attribution to provide clinically relevant diagnostic information, and the adaptive confidence calibration allows for more reliable predictions for multi-class diseases classification. Experiments on a number of public medical image datasets show that the proposed EHVT-GIN architecture surpasses other EfficientNet, Vision Transformer, and Hybrid CNN-Transformer architectures in terms of classification accuracy, which is achieved at 98.72% accuracy, 98.46% precision, 98.31% recall, 98.38% F1 score, and an average inference time of 17.1 ms per image. These findings prove that the proposed approach provides significant improvement in diagnostics, computation, robustness, and explainability, thus making it an appropriate choice for trustworthy real-time medical image analysis and disease diagnostics.
 
关键词
Vision Transformer, Generative AI, Explainable Artificial Intelligence, Medical Image Analysis, Disease Diagnosis
报告人
SANKARAI V
AP K.Ramakrishnan College of Engineering;TRICHY - TAMILNADU

稿件作者
PRADEEPKUMAR S K B Ramachandra College of Engineering
Mohd Akram Khan Madhav University, Aburoad
Meeha D Dr. Sakunthala Engineering College,
DIWAKAR AGARWAL GLA University, Mathura
Anandakumar Haldorai Sri Eshwar College of Engineering
SANKARAI V K.Ramakrishnan College of Engineering;TRICHY - TAMILNADU
发表评论
验证码 看不清楚,更换一张
全部评论
重要日期
  • 会议日期

    07月30日

    2026

    08月01日

    2026

  • 07月26日 2026

    初稿截稿日期

  • 07月30日 2026

    注册截止日期

主办单位
The United Societies of Science
承办单位
Kongunadu College of Engineering and Technology
协办单位
IEEE Section
IEEE Madras Section
历届会议
移动端
在手机上打开
小程序
打开微信小程序
客服
扫码或点此咨询