Q1 2026

Hierarchical context modeling and prototype-mediated cross-modal alignment for histopathology report generation

Chengxin Ye · Jingqin Lv · Guangli Li · Renzhong Wu · Shiying Zeng · Nan Jiang · Boyang Liu · Jianguo Wu · Donghong Ji · Hongbin Zhang
10.1007/s44443-026-00672-z 395 المشاهدات 0 الاقتباسات
0
الاقتباسات
395
المشاهدات
الملخص

Abstract

Histopathology in whole slide images (WSIs) serves as the gold standard for cancer diagnosis, with clinical reports playing a critical role in decision-making. However, the time-consuming nature of conventional pathological examination has driven increasing and urgent demand for automated report generation. Deep learning methods offer a certain potential to revolutionize this requirement by Histopathology Report Generation (HRG). Nevertheless, existing HRG approaches suffer from low-quality generation results due to ineffective exploration of multi-scale visual context in gigapixel WSIs and the inherent semantic gap between heterogeneous vision-language modalities. To address these challenges, we propose HC-Gen, a novel framework which synergistically combines hierarchical context modeling with prototype-mediate cross-modal alignment for HRG. Inspired by pathologists’ anatomically-grounded diagnostic logic, we design a hierarchical context fusion module to integrate multi-scale visual-semantic context and implicit hierarchy prior in WSIs. Furthermore, we propose a cross-modal prototypical memory module to establish learnable semantic prototypes as intermediate bridges to achieve unified and efficient vision-language alignment. Model performance was assessed through natural language generation metrics and human evaluation, extensive experiments on two benchmark datasets demonstrate that HC-Gen outperforms state-of-the-art methods. Extra visualization provides crucial support for the interpretability of the decision process. Our code is available at:
https://github.com/Modaoshuangming/HC-Gen
.

الاستشهاد بهذا المقال (APA)
Chengxin, Y., Jingqin, L., Guangli, L., Renzhong, W., Shiying, Z., Nan, J., Boyang, L., Jianguo, W., Donghong, J., Hongbin, Z. (2026). Hierarchical context modeling and prototype-mediated cross-modal alignment for histopathology report generation. Journal of King Saud University - Computer and Information Sciences. https://doi.org/10.1007/s44443-026-00672-z
أبحاث ذات صلة
A lightweight model for indoor object detection in unstructured scenes based on joint attention and …
Zhizhong Xing; Leping Li; Ying Yang; Wei Zhou; Guolan Ma; Shaochun Chen; Lechun · 2026
13
استشهاد
423
DDM-YOLO: A lightweight oriented detection model for mature daylily fruits in complex environments
Minqiu Kuang; Xuejie Zou; Fangping Xie; Xiaojian Li; Shang Chen; Dawei Liu; Yuxu · 2026
8
استشهاد
429
Information guided Levy flight for robot search in unknown environments
Weitao Zhao; Zati Hakim Azizul; Xin Lyu; Weijie Kuang · 2026
4
استشهاد
411
3
استشهاد
502
Bridging the gap: A comprehensive survey on AI-driven digital twin networks for future wireless syst…
Yousef Sanjalawe; Salam Fraihat; Salam Al-E’mari; Sharif Naser Makhadmeh · 2026
3
استشهاد
416
الوصول
عرض النص الكامل عبر DOI
نُشر في
الرقم الدولي ISSN 1319-1578
الربعية Q1
درجة المؤشر القياس العربي 100
التخصص Computer Science & AI
الناشر Elsevier / King Saud University
الدولة 🇸🇦 Saudi Arabia
عرض ملف المجلة →
المؤلفون
تفاصيل النشر
السنة 2026
اللغة English
أُضيف في 06 Jul 2026