← Search

Zeeshan Hayder

13 accepted papers

2026

ARMFlow: AutoRegressive MeanFlow for Online 3D Human Reaction Generation

CVPR 2026

3D human reaction generation faces three main challenges: (1) high motion fidelity, (2) real-time inference, and (3) autoregressive adaptability for online scenarios. Existing methods fail to meet all three simultaneously. We propose ARMFlow, a MeanFlow-based autoregressive framework that models tem

Cited by 0SourcecodeScholar
2026

BiFM: Bidirectional Flow Matching for Few-Step Image Editing and Generation

CVPR 2026

Recent diffusion and flow matching models have demonstrated strong capabilities in image generation and editing by progressively removing noise through iterative sampling. While this enables flexible inversion for semantic-preserving edits, few-step sampling regimes suffer from poor forward process

Cited by 1SourceScholar
2026

DTO-KD: Dynamic Trade-off Optimization for Effective Knowledge Distillation

ICLR 2026oral

Knowledge Distillation (KD) is a widely adopted framework for compressing large models into compact student models by transferring knowledge from a high-capacity teacher. Despite its success, KD presents two persistent challenges: (1) the trade-off between optimizing for the primary task loss and mi…

Cited by 0SourceScholar
2026

Disentangled Hierarchical VAE for 3D Human-Human Interaction Generation

ICLR 2026poster

Generating realistic 3D Human-Human Interaction (HHI) requires coherent modeling of the physical plausibility of the agents and their interaction semantics. Existing methods compress all motion information into a single latent representation, limiting their ability to capture fine-grained actions an…

Cited by 0SourcecodeScholar
2025

Auto-Regressive Diffusion for Generating 3D Human-Object Interactions

AAAI 2025technical

Text-driven Human-Object Interaction (Text-to-HOI) generation is an emerging field with applications in animation, video games, virtual reality, and robotics. A key challenge in HOI generation is maintaining interaction consistency in long sequences. Existing Text-to-Motion-based approaches, such as…

2024

Backpropagation-free Network for 3D Test-time Adaptation

CVPR 2024poster

Real-world systems often encounter new data over time which leads to experiencing target domain shifts. Existing Test-Time Adaptation (TTA) methods tend to apply computationally heavy and memory-intensive backpropagation-based approaches to handle this. Here we propose a novel method that uses a bac…

2024

Canonical Shape Projection is All You Need for 3D Few-shot Class Incremental Learning

ECCV 2024poster

"In recent years, robust pre-trained foundation models have been successfully used in many downstream tasks. Here, we would like to use such powerful models to address the problem of few-shot class incremental learning (FSCIL) tasks on 3D point cloud objects. Our approach is to reprogram the well-kn…

2023

Hyperbolic Audio-visual Zero-shot Learning

ICCV 2023poster

Audio-visual zero-shot learning aims to classify samples consisting of a pair of corresponding audio and video sequences from classes that are not present during training. An analysis of the audio-visual data reveals a large degree of hyperbolicity, indicating the potential benefit of using a hyperb…

Cited by 22PDFScholar