Home | Publications | WRM26

Weighting What Matters: Boosting Sample Efficiency in Medical Report Generation via Token Reweighting

MCML Authors

Alexander Weers

→ Group Martin Menten
Artificial Intelligence in Healthcare and Medicine

Daniel Rückert

Prof. Dr.

Director

Artificial Intelligence in Healthcare and Medicine

Martin Menten

Dr.

JRG Leader AI for Vision

Artificial Intelligence in Healthcare and Medicine

Abstract

Training vision-language models (VLMs) for medical report generation is often hindered by the scarcity of high-quality annotated data. This work evaluates the use of a weighted loss function to improve data efficiency. Compared to standard cross-entropy loss, which treats all token prediction errors equally, the reweighted loss shifts the focus to semantically salient tokens with outsized clinical importance. In experiments on ophthalmological report generation, we show that this simple method improves efficiency across multiple data scales, achieving similar report quality with up to ten times less training data.

inproceedings WRM26

MIDL 2026

Medical Imaging with Deep Learning. Taipei, Taiwan, Jul 08-10, 2026. To be published. Preprint available.

Authors

A. Weers • D. Rückert • M. J. Menten

Links

URL

Research Area

C1 | Medicine

BibTeXKey: WRM26

#p-menten #p-rueckert