← Search

Sophie Ostmeier

8 accepted papers

2026

Diffusion MRI Transformer with a Diffusion Space Rotary Positional Embedding (D-RoPE)

CVPR 2026

Diffusion Magnetic Resonance Imaging (dMRI) plays a critical role in studying microstructural changes in the brain. It is, therefore, widely used in clinical practice; yet progress in learning general-purpose representations from dMRI has been limited. A key challenge is that existing deep learning

Cited by 0SourcecodeScholar
2026

Symbal: Detecting Systematic Misalignments in Model-Generated Captions

ICML 2026poster

Multimodal large language models (MLLMs) often introduce errors when generating image captions, resulting in misaligned image-text pairs. Our work focuses on a class of captioning errors that we refer to as systematic misalignments, where a recurring error in MLLM-generated captions is closely assoc…

Cited by 0SourceScholar
2025

Automated Structured Radiology Report Generation

ACL 2025long

Automated radiology report generation from chest X-ray (CXR) images has the potential to improve clinical efficiency and reduce radiologists’ workload. However, most datasets, including the publicly available MIMIC-CXR and CheXpert Plus, consist entirely of free-form reports, which are inherently va…

Cited by 0SourcePDFScholar
2025

CheXalign: Preference fine-tuning in chest X-ray interpretation models without human feedback

ACL 2025long

Radiologists play a crucial role in translating medical images into actionable reports. However, the field faces staffing shortages and increasing workloads. While automated approaches using vision-language models (VLMs) show promise as assistants, they require exceptionally high accuracy. Most curr…

2025

LieRE: Lie Rotational Positional Encodings

ICML 2025poster

Transformer architectures depend on explicit position encodings to capture token positional information. Rotary Position Encoding (RoPE) has emerged as a popular choice in language models due to its efficient encoding of relative position information through key-query rotations. However, RoPE faces…

2025

Structuring Radiology Reports: Challenging LLMs with Lightweight Models

EMNLP 2025

Radiology reports are critical for clinical decision-making but often lack a standardized format, limiting both human interpretability and machine learning (ML) applications. While large language models (LLMs) have shown strong capabilities in reformatting clinical text, their high computational req

Cited by 0SourcePDFScholar
2025

TRoVe: Discovering Error-Inducing Static Feature Biases in Temporal Vision-Language Models

NeurIPS 2025poster

Vision-language models (VLMs) have made great strides in addressing temporal understanding tasks, which involve characterizing visual changes across a sequence of images. However, recent works have suggested that when making predictions, VLMs may rely on static feature biases, such as background or…

Cited by 0SourcecodeScholar
2024

GREEN: Generative Radiology Report Evaluation and Error Notation

EMNLP 2024finding

Evaluating radiology reports is a challenging problem as factual correctness is extremely important due to its medical nature. Existing automatic evaluation metrics either suffer from failing to consider factual correctness (e.g., BLEU and ROUGE) or are limited in their interpretability (e.g., F1Che…

Cited by 19SourcePDFScholar