-
Prediction Degeneracy in Medical Vision-Language Models: Implications for Robustness and Interpretability2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCVW) conference-paper Computer Science Multimodal Machine Learning Applications
Martin Goetze, Dennis Eschweiler, Brendan Huang, Gustav Anton Müller‐Franzes, Carolina Ramirez, Madeline Hess, +3 more
0citations -
Case-Grounded Evidence Verification: A Framework for Constructing Evidence-Sensitive Supervision2026 arXiv (Cornell University) preprint Computer Science Multimodal Machine Learning Applications Open access
Soroosh Tayebi Arasteh, Mehdi Joodaki, Mahshad Lotfinia, Sven Nebelung, Daniel Truhn
0citations -
Systematische Untersuchung der Leistungsfähigkeit visueller Sprachmodelle in der radiologischen Bildinterpretation2026 RöFo - Fortschritte auf dem Gebiet der Röntgenstrahlen und der bildgebenden Verfahren conference-paper Computer Science Multimodal Machine Learning Applications
M von der Stück, Roman Vuskov, Simon D. Westfechtel, Robert Malte Siepmann, C Kuhl, S Nebelung
0citations -
Vision-language models for chest radiography do not always need the image2026 arXiv (Cornell University) preprint Computer Science Multimodal Machine Learning Applications Open access
Mahshad Lotfinia, Sebastian Ziegelmayer, Lisa Adams, Tri-Thien Nguyen, Daniel Truhn, Andreas Maier, +1 more
0citations -
Sparse Local Latents for Explainable Zero-Shot Reasoning in Medical Vision-Language Models2026 Lecture notes in computer science conference-paper Computer Science Multimodal Machine Learning Applications
Martin Goetze, Patrick Wienholt, Dennis Eschweiler, Marvin Gazibaric, Christiane Kuhl, Sven Nebelung, +1 more
0citations -
HotelMatch-LLM: Joint Multi-Task Training of Small and Large Language Models for Efficient Multimodal Hotel Retrieval2025 Annual Meeting of the Association for Computational Linguistics (ACL) conference-paper Computer Science Multimodal Machine Learning Applications Open access
Arian Askari, Emmanouil Stergiadis, Ilya Gusev, Moran Beladev
0citations -
Optimizing VLP-aligned Multimodal Intent Representation with Correct Visual Instantiation for Zero-Shot Composed Image Retrieval2026 arXiv (Cornell University) preprint Computer Science Multimodal Machine Learning Applications Open access
Xuri Ge, Chunhao Wang, Junchen Fu, Haokun Wen, Zhiwei Xu, Ying Zhou, +4 more
0citations -
LEGO Co-builder: Exploring Fine-Grained Vision-Language Modeling for Multimodal Assembly Assistants2026 INTERNATIONAL CONFERENCE ON MULTIMODAL INTERACTION conference-paper Computer Science Multimodal Machine Learning Applications Open access
Haochen Huang, Yue Su, Xin Sun, Moonisa Ahsan, Mohammad Aliannejadi, Irene Viola, +8 more
0citations