Home  | News

10.06.2025

Tiny logo
Teaser image to MCML at CVPR 2025

MCML at CVPR 2025

35 Accepted Papers (29 Main, and 6 Workshops)

IEEE/CVF Conference on Computer Vision and Pattern Recognition, Nashville, TN, USA, Jun 11-15, 2025

We are happy to announce that MCML researchers have contributed a total of 35 papers to CVPR 2025: 29 Main, and 6 Workshop papers. Congrats to our researchers!

Main Track (29 papers)

S. A. Baumann • F. Krause • M. Neumayr • N. Stracke • M. Sevi • V. T. HuB. Ommer
Continuous, Subject-Specific Attribute Control in T2I Models by Identifying Semantic Directions.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI GitHub

Q. Bouniot • I. Redko • A. Mallasto • C. Laclau • O. Struckmeier • K. Arndt • M. Heinonen • V. Kyrki • S. Kaski
From Alexnet to Transformers: Measuring the Non-linearity of Deep Neural Networks with Affine Optimal Transport.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

H. Chen • H. Li • Y. ZhangG. ZhangJ. Bi • P. Torr • J. Gu • D. Krompass • V. Tresp
FedBiP: Heterogeneous One-Shot Federated Learning with Personalized Latent Diffusion Models.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

C. CurreliD. MuhleA. Saroha • Z. Ye • R. MarinD. Cremers
Nonisotropic Gaussian Diffusion for Realistic 3D Human Motion Prediction.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI GitHub

Z. Chen • Y. Wang • L. Nan • X. Zhu
Parametric Point Cloud Completion for Polygonal Surface Reconstruction.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI GitHub

S. Dziadzio • V. Udandarao • K. Roth • A. Prabhu • Z. Akata • S. Albanie • M. Bethge
How to Merge Your Multimodal Models Over Time?
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

T. DagèsS. WeberY.-W. E. Lin • R. Talmon • D. Cremers • M. Lindenbaum • A. M. Bruckstein • R. Kimmel
Finsler Multi-Dimensional Scaling: Manifold Learning for Asymmetric Dimensionality Reduction and Embedding.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

T. Hannan • M. M. Islam • J. Gu • T. Seidl • G. Bertasius
ReVisionLLM: Recursive Vision-Language Model for Temporal Grounding in Hour-Long Videos.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI GitHub

O. Hahn • C. ReichN. AraslanovD. Cremers • C. Rupprecht • S. Roth
Scene-Centric Unsupervised Panoptic Segmentation.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI GitHub

S. KimR. XiaoM.-I. GeorgescuS. AlanizZ. Akata
COSMOS: Cross-Modality Self-Distillation for Vision Language Pre-training.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI GitHub

T. Liu • Z. Lai • J. Wang • G. ZhangS. Chen • P. Torr • V. Demberg • V. Tresp • J. Gu
Multimodal Pragmatic Jailbreak on Text-to-image Models.
CVPR 2025 - 2nd Workshop on Responsible Generative AI at the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. Best Paper Award. URL GitHub

W. Li • H. Xu • J. HuangH. Jung • P. K. Yu • N. NavabB. Busam
GCE-Pose: Global Context Enhancement for Category-level Object Pose Estimation.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI GitHub

D. Mildenberger • P. Hager • D. RückertM. J. Menten
A Tale of Two Classes: Adapting Supervised Contrastive Learning to Binary Imbalanced Datasets.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

E. ÖzsoyC. Pellegrini • T. Czempiel • F. TristramK. YuanD. Bani-HarouniU. EckB. BusamM. KeicherN. Navab
MM-OR: A Large Multimodal Operating Room Dataset for Semantic Understanding of High-Intensity Surgical Environments.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI GitHub

R. Qorbani • G. Villani • T. Panagiotakopoulos • M. B. Colomer • L. Härenstam-Nielsen • M. Segu • P. L. Dovesi • J. Karlgren • D. Cremers • F. Tombari • M. Poggi
Semantic Library Adaptation: LoRA Retrieval and Fusion for Open-Vocabulary Semantic Segmentation.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

K. RothZ. Akata • D. Damen • I. Balažević • O. J. Hénaff
Context-Aware Multimodal Pretraining.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

P. Roetzer • V. EhmD. Cremers • Z. Lähner • F. Bernard
Higher-Order Ratio Cycles for Fast and Globally Optimal Shape Matching.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

D. SchnausN. AraslanovD. Cremers
It's a (Blind) Match! Towards Vision-Language Correspondence without Parallel Data.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI GitHub

N. Stracke • S. A. Baumann • K. Bauer • F. Fundel • B. Ommer
CleanDIFT: Diffusion Features without Noise.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

L. Sang • Z. Canfes • D. Cao • R. Marin • F. Bernard • D. Cremers
4Deform: Neural Surface Deformation for Robust Shape Interpolation.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

J. Schusterbauer • M. Gui • F. Fundel • B. Ommer
Diff2Flow: Training Flow Matching Models via Diffusion Model Alignment.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

D. SinitsynL. Härenstam-NielsenD. Cremers
PRaDA: Projective Radial Distortion Averaging.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

F. WimbauerW. ChenD. Muhle • C. Rupprecht • D. Cremers
AnyCam: Learning to Recover Camera Poses and Intrinsics from Casual Videos.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

Y. Xie • V. Ehm • P. Roetzer • N. Amrani • M. Gao • F. Bernard • D. Cremers
EchoMatch: Partial-to-Partial Shape Matching via Correspondence Reflection.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

R. XiaoS. KimM.-I. GeorgescuZ. AkataS. Alaniz
FLAIR: VLM with Fine-grained Language-informed Image Representations.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI GitHub

Y. YeganehA. Farshad • I. Charisiadis • M. Hasny • M. Hartenberger • B. OmmerN. Navab • E. Adeli
Latent Drifting in Diffusion Models for Counterfactual Medical Image Synthesis.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. Highlight Paper. DOI

Y. Yuan • Y. XiaD. Cremers • M. Sester
SparseAlign: a Fully Sparse Framework for Cooperative Object Detection.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

D. Zhu • Y. Di • S. Gavranovic • S. Ilic
SeaLion: Semantic Part-Aware Latent Point Diffusion Models for 3D Generation.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

G. Zhang • M. L. A. Fok • J. Ma • Y. XiaD. Cremers • P. Torr • V. Tresp • J. Gu
Localizing Events in Videos with Multimodal Queries.
CVPR 2025 - IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

Workshops (6 papers)

L. BastianM. RashedN. Navab • T. Birdal
Continuous-Time SO(3) Forecasting with Savitzky--Golay Neural Controlled Differential Equations.
4DVision @CVPR 2025 - Workshop on 4D Vision: Modeling the Dynamic World at the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. arXiv

Y. Luo • R. Hoffmann • Y. Xia • O. Wysocki • B. Schwab • T. H. Kolbe • D. Cremers
RADLER: Radar Object Detection Leveraging Semantic 3D City Models and Self-Supervised Radar-Image Learning.
PBVS @CVPR 2025 - 21st IEEE Workshop on Perception Beyond the Visible Spectrum at the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI GitHub

E. ÖzsoyF. HolmC. Pellegrini • T. Czempiel • M. Saleh • N. NavabB. Busam
Location-Free Scene Graph Generation.
MULA @CVPR 2025 - 8th Multimodal Learning and Applications Workshop at the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

W. Tang • W. Li • X. Liang • O. Wysocki • F. Biljecki • C. Holst • B. Jutzi
Texture2LoD3: Enabling LoD3 Building Reconstruction With Panoramic Images.
USM3D @CVPR 2025 - 2nd Workshop on Urban Scene Modeling at the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI GitHub

L. Waldmann • A. Shah • Y. Wang • N. LehmannA. J. StewartZ. XiongX. ZhuS. Bauer • J. Chuang
Panopticon: Advancing Any-Sensor Foundation Models for Earth Observation.
EARTHVISION @CVPR 2025 - Workshop EarthVision: Large Scale Computer Vision for Remote Sensing Imagery at the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. DOI

D. Zverev • T. Wiedemer • A. Prabhu • M. Bethge • W. Brendel • A. S. Koepke
VGGSounder: Audio-Visual Evaluations for Foundation Models.
Sight and Sound @CVPR 2025 - Workshop Sight and Sound at the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville, TN, USA, Jun 11-15, 2025. PDF

#research #top-tier-work #akata #bauer-s #busam #cremers #jegelka #koepke #marin-riccardo #menten #navab #ommer #rueckert #seidl #thuerey #tresp #wang-xi #zhu

Related

Link to Cordelia Schmid Featured in Süddeutche Zeitung

11.05.2026

Cordelia Schmid Featured in Süddeutche Zeitung

Cordelia Schmid, a member of the MCML Advisory Board, was recently featured in Süddeutsche Zeitung for her work in computer vision and robotics.

Read more
Link to Right answer, wrong reasoning - Is AI Thinking or Cheating?

08.05.2026

Right Answer, Wrong Reasoning - Is AI Thinking or Cheating?

Can AI cheat without us noticing? Our PI Barbara Plank and her team introduce a new detection method at ICLR 2026.

Read more
Link to MCML Delegation Visit to the UK

07.05.2026

MCML Delegation Visit to the UK

MCML delegation visited top U.S. universities to advance AI X-Change and foster collaboration in generative and medical AI.

Read more
Link to Matthias Niessner Featured in Handelsblatt Disrupt podcast

07.05.2026

Matthias Niessner Featured in Handelsblatt Disrupt Podcast

Matthias Niessner joined the Handelsblatt Disrupt podcast to discuss the growing hype around “World Models”.

Read more
Link to MemAgents Workshop at ICLR 2026

06.05.2026

MemAgents Workshop at ICLR 2026

The workshop MemAgents workshop focused on the role of memory in long-horizon AI agents. Co-Organized by Hinrich Schütze, Yunpu Ma and Ercong Nie.

Read more
Back to Top