23.07.2025

Teaser image to How Reliable Are Machine Learning Methods? With Anne-Laure Boulesteix and Milena Wünsch

How Reliable Are Machine Learning Methods? With Anne-Laure Boulesteix and Milena Wünsch

Research Film

Often a new machine learning method claims to outperform the last. Whether it’s in bioinformatics, finance, or image recognition, the message is the same: this algorithm is faster, more accurate, more powerful. But can we trust those claims?

«It’s not just about the algorithms. It’s about how we compare them—and what we choose to report or ignore.»


Milena Wünsch

MCML Junior Member

Beneath the surface of many benchmarking studies lies a quiet problem: subtle biases that skew comparisons and inflate performance. These issues often go unnoticed — but they can have real consequences, especially when such models are used to inform research or high-stakes decisions.

«It doesn’t matter whether the bias is deliberate or not. It still shapes how methods are judged and used.»


Anne-Laure Boulesteix

MCML PI

Anne-Laure Boulesteix, Professor of Biometry at LMU and MCML PI, and Milena Wünsch, PhD student at LMU and MCML, study how seemingly harmless methodological choices can lead to misleading results.

One common issue: when a method fails on a dataset, researchers may simply drop it from the analysis. While convenient, this can introduce bias and overstate performance.

Bias can also arise from less obvious sources — like spending more time tuning one method, being more familiar with a tool, or unconsciously interpreting results in its favor.

With so many studies promoting the “next best” algorithm, it’s hard to know which results to trust. Researchers may end up using a method that only looked good due to biased comparisons. Still, the researchers are hopeful. In recent years, the methodological machine learning community has made real progress — pushing for better standards, more transparency, and more careful benchmarking.

Watch in Full Quality on youTube

The film was produced and edited by Nicole Huminski and Nikolai Huber.

 

23.07.2025


Subscribe to RSS News feed

Related

Link to AI for Personalized Psychiatry - with researcher Clara Vetter

01.09.2025

AI for Personalized Psychiatry - With Researcher Clara Vetter

AI research by Clara Vetter uses brain, genetic and smartphone data to personalize psychiatry and improve diagnosis and treatment.

Link to Satellite Insights for a Sustainable Future - with researcher Ivica Obadic

25.08.2025

Satellite Insights for a Sustainable Future - With Researcher Ivica Obadic

AI from satellite imagery helps design livable cities, improve well-being & food systems with transparent models by Ivica Obadić.

Link to Mingyang Wang receives Award at ACL 2025

18.08.2025

Mingyang Wang Receives Award at ACL 2025

MCML Junior Member Mingyang Wang wins SAC Highlights Award at ACL 2025 for research on cross-lingual consistency in language models.

Link to Digital Twins for Surgery - with researcher Azade Farshad

18.08.2025

Digital Twins for Surgery - With Researcher Azade Farshad

Azade Farshad develops patient digital twins at TUM & MCML to improve personalized treatment, surgical planning, and training.

Link to From Physics Dreams to Algorithm Discovery - with Niki Kilbertus

13.08.2025

From Physics Dreams to Algorithm Discovery - With Niki Kilbertus

Niki Kilbertus develops AI algorithms to uncover cause and effect, making science smarter and decisions in fields like medicine more reliable.