← Search

Jinzhao Li

5 accepted papers

2026

EgoProx: Evaluating MLLMs on Egocentric 3D Proximity Reasoning Across a Cognitive Hierarchy

CVPR 2026

Humans constantly reason about 3D proximity, the relations between their body and surrounding objects, to guide perception and action in daily life. Whether multimodal large language models (MLLMs) can perform such embodied 3D reasoning remains unclear. To this end, we introduce EgoProx, a benchmark

Cited by 0SourceScholar
2026

Human-LLM Collaborative Feature Engineering for Tabular Data

ICLR 2026poster

Large language models (LLMs) are increasingly used to automate feature engineering in tabular learning. Given task-specific information, LLMs can propose diverse feature transformation operations to enhance downstream model performance. However, current approaches typically assign the LLM as a black…

Cited by 0SourceScholar
2025

Scene Splatter: Momentum 3D Scene Generation from Single Image with Video Diffusion Model

CVPR 2025poster

In this paper, we propose Scene Splatter, a momentum-based paradigm for video diffusion to generate generic scenes from single image. Existing methods, which employ video generation models to synthesize novel views, suffer from limited video length and scene inconsistency, leading to artifacts and d…

Cited by 1SourcePDFScholar
2024

Solving Satisfiability Modulo Counting for Symbolic and Statistical AI Integration with Provable Guarantees

AAAI 2024technical

Satisfiability Modulo Counting (SMC) encompasses problems that require both symbolic decision-making and statistical reasoning. Its general formulation captures many real-world problems at the intersection of symbolic and statistical AI. SMC searches for policy interventions to control probabilistic…