← Search

Andreas Maier

4 accepted papers

2026

SPEECHCT-CLIP: DISTILLING TEXT-IMAGE KNOWLEDGE TO SPEECH FOR VOICE-NATIVE MULTIMODAL CT ANALYSIS

ICASSP 2026oral

Spoken communication plays a central role in clinical workflows. In radiology, for example, most reports are created through dictation. Yet, nearly all medical AI systems rely exclusively on written text. In this work, we address this gap by exploring the feasibility of learning visual-language repr…

Cited by 0SourcePDFScholar
2023

PLIKS: A Pseudo-Linear Inverse Kinematic Solver for 3D Human Body Estimation

CVPR 2023highlight

We introduce PLIKS (Pseudo-Linear Inverse Kinematic Solver) for reconstruction of a 3D mesh of the human body from a single 2D image. Current techniques directly regress the shape, pose, and translation of a parametric model from an input image through a non-linear mapping with minimal flexibility t…

Cited by 38SourcePDFScholar
2022

Deepfilternet: A Low Complexity Speech Enhancement Framework for Full-Band Audio Based On Deep Filtering

ICASSP 2022accepted

Complex-valued processing has brought deep learning-based speech enhancement and signal extraction to a new level. Typically, the process is based on a time-frequency (TF) mask which is applied to a noisy spectrogram, while complex masks (CM) are usually preferred over real-valued masks due to their…

Cited by 0SourceScholar
2019

Estimating the Fundamental Matrix Without Point Correspondences With Application to Transmission Imaging

ICCV 2019poster

We present a general method to estimate the fundamental matrix from a pair of images under perspective projection without the need for image point correspondences. Our method is particularly well-suited for transmission imaging, where state-of-the-art feature detection and matching approaches genera…

Cited by 8PDFScholar