← Search

Amogh Joshi

4 accepted papers

2026

LSD-3D: Large-Scale 3D Driving Scene Generation with Geometry Grounding

AAAI 2026technical

Large-scale scene data is essential for training and testing in robot learning. Neural reconstruction methods have promised the capability of reconstructing large physically-grounded outdoor scenes from captured sensor data. However, these methods have baked-in static environments and only allow for

Cited by 0SourcePDFScholar
2024

FEDORA: A Flying Event Dataset fOr Reactive behAvior

IROS 2024poster

The ability of resource-constrained biological systems such as fruitflies to perform complex and high-speed maneuvers in cluttered environments has been one of the prime sources of inspiration for developing vision-based autonomous systems. To emulate this capability, the perception pipeline of such…

Cited by 1SourceScholar
2024

Understanding the Limits of Vision Language Models Through the Lens of the Binding Problem

NeurIPS 2024poster

Recent work has documented striking heterogeneity in the performance of state-of-the-art vision language models (VLMs), including both multimodal language models and text-to-image models. These models are able to describe and generate a diverse array of complex, naturalistic images, yet they exhibit…

Cited by 7SourcePDFScholar