[2511.10065] RadHiera: Semantic Hierarchical Reinforcement Learning

[2511.10065] RadHiera: Semantic Hierarchical Reinforcement Learning for Medical Report Generation

arXiv - AI March 24, 2026 4 min read

About this article

Abstract page for arXiv paper 2511.10065: RadHiera: Semantic Hierarchical Reinforcement Learning for Medical Report Generation

Computer Science > Artificial Intelligence arXiv:2511.10065 (cs) [Submitted on 13 Nov 2025 (v1), last revised 23 Mar 2026 (this version, v2)] Title:RadHiera: Semantic Hierarchical Reinforcement Learning for Medical Report Generation Authors:Bodong Du, Honglong Yang, Xiaomeng Li View a PDF of the paper titled RadHiera: Semantic Hierarchical Reinforcement Learning for Medical Report Generation, by Bodong Du and 2 other authors View PDF HTML (experimental) Abstract:Vision-language models have shown promising results in radiology report generation. However, most existing methods generate reports as flat text and do not explicitly model the semantic dependency between the Findings and Impression sections, which can lead to inconsistencies between clinical observations and diagnostic conclusions. In this paper, we propose RadHiera, a semantic hierarchical reinforcement learning framework for radiology report generation. RadHiera follows the semantic organization of radiology reports by first optimizing overall report quality, then improving the diagnostic accuracy of the Impression section, and finally enforcing consistency between Findings and Impression so that diagnostic conclusions are supported by clinical evidence. Specifically, we begin with a base reward that combines linguistic quality and medical factuality to provide supervision on the whole report. On this basis, we introduce a severity-aware reward for the Impression section that places greater emphasis on errors in...

Originally published on March 24, 2026. Curated by AI News.

Llms

[2603.17839] How do LLMs Compute Verbal Confidence

Abstract page for arXiv paper 2603.17839: How do LLMs Compute Verbal Confidence

arXiv - AI · 4 min · 16 minutes ago

Llms

[2603.15970] 100x Cost & Latency Reduction: Performance Analysis of AI Query Approximation using Lightweight Proxy Models

Abstract page for arXiv paper 2603.15970: 100x Cost & Latency Reduction: Performance Analysis of AI Query Approximation using Lightweight...

arXiv - AI · 4 min · 16 minutes ago

Llms

[2603.10062] Multi-Agent Memory from a Computer Architecture Perspective: Visions and Challenges Ahead

Abstract page for arXiv paper 2603.10062: Multi-Agent Memory from a Computer Architecture Perspective: Visions and Challenges Ahead

arXiv - AI · 3 min · 16 minutes ago

Llms

[2603.09085] Not All News Is Equal: Topic- and Event-Conditional Sentiment from Finetuned LLMs for Aluminum Price Forecasting

Abstract page for arXiv paper 2603.09085: Not All News Is Equal: Topic- and Event-Conditional Sentiment from Finetuned LLMs for Aluminum ...

arXiv - AI · 4 min · 16 minutes ago

[2511.10065] RadHiera: Semantic Hierarchical Reinforcement Learning for Medical Report Generation

About this article

Related Articles

[2603.17839] How do LLMs Compute Verbal Confidence

[2603.15970] 100x Cost & Latency Reduction: Performance Analysis of AI Query Approximation using Lightweight Proxy Models

[2603.10062] Multi-Agent Memory from a Computer Architecture Perspective: Visions and Challenges Ahead

[2603.09085] Not All News Is Equal: Topic- and Event-Conditional Sentiment from Finetuned LLMs for Aluminum Price Forecasting

No comments

Stay updated with AI News