← Search

Jiayi Wu

20 accepted papers

2026

Benchmarking Overton Pluralism in LLMs

ICLR 2026poster

We introduce a novel framework for measuring Overton pluralism in LLMs—the extent to which diverse viewpoints are represented in model outputs. We (i) formalize Overton pluralism as a set-coverage metric (OVERTONSCORE), (ii) conduct a large-scale US-representative human study (N=1209; 60 questions;…

Cited by 0SourcecodeScholar
2026

ICLR: Inter-Chrominance and Luminance Interaction for Natural Color Restoration in Low-Light Image Enhancement

AAAI 2026technical

Low-Light Image Enhancement (LLIE) task aims at improving contrast while restoring details and textures for images captured in low-light conditions. HVI color space has made significant progress in this task by enabling precise decoupling of chrominance and luminance. However, for the interaction of

Cited by 0SourcePDFScholar
2026

NavMoE: Hybrid Model and Learning-Based Traversability Estimation for Local Navigation Via Mixture of Experts

ICRA 2026poster

This paper explores traversability estimation for robot navigation. A key bottleneck in traversability estimation lies in efficiently achieving reliable and robust predictions while accurately encoding both geometric and semantic information across diverse environments. We introduce Navigation via M…

2026

PolarDepth: Polarization-Guided Monocular Depth for Visual Odometry

RA-L 2026

Glass surfaces remain challenging for indoor robot perception. Depth sensors and RGB-only monocular depth estimation often fail because of reflections, refractions, and low-texture regions. To this end, we present PolarDepth, a polarization-enhanced monocular depth framework for glass-dominant envir

Cited by 0SourceScholar
2025

AquaFuse: Waterbody Fusion for Physics-Guided View Synthesis of Underwater Scenes

RA-L 2025

In this letter, we introduce the idea of AquaFuse, a physics-based method for synthesizing <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">waterbody properties</i> in underwater imagery. We formulate a closed-form solution for waterbody fusion that f

Cited by 6SourceScholar
2025

Controlling The Spread of Epidemics on Networks with Differential Privacy

NeurIPS 2025poster

Designing effective strategies for controlling epidemic spread by vaccination is an important question in epidemiology, especially in the early stages when vaccines are limited. This is a challenging question when the contact network is very heterogeneous, and strategies based on controlling network…

Cited by 0SourceScholar
2025

Learning Normal Flow Directly From Events

ICCV 2025poster

Event-based motion field estimation is an important task. However, current optical flow methods face challenges: learning-based approaches, often frame-based and relying on CNNs, lack cross-domain transferability, while model-based methods, though more robust, are less accurate. To address the limit…

2025

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching

ICCV 2025poster

Leveraging the vision foundation models has emerged as a mainstream paradigm that improves the performance of image feature matching. However, previous works have ignored the misalignment when introducing the foundation models into feature matching. The misalignment arises from the discrepancy betwe…

Cited by 0SourcePDFScholar
2025

PA-RAG: RAG Alignment via Multi-Perspective Preference Optimization

NAACL 2025long

The emergence of Retrieval-augmented generation (RAG) has alleviated the issues of outdated and hallucinatory content in the generation of large language models (LLMs), yet it still reveals numerous limitations. When a general-purpose LLM serves as the RAG generator, it often suffers from inadequate…

2025

ShotVL: Human-Centric Highlight Frame Retrieval via Language Queries

AAAI 2025technical

Existing research on human-centric video understanding typically focuses on analyzing specific moments or entire videos. However, many applications require higher precision at the frame level. In this work, we propose a novel task, BestShot, which aims to locate highlight frames within human-centric…

2025

Simulating Society Requires Simulating Thought

NeurIPS 2025poster

Simulating society with large language models (LLMs), we argue, requires more than generating plausible behavior; it demands cognitively grounded reasoning that is structured, revisable, and traceable. LLM-based agents are increasingly used to emulate individual and group behavior, primarily through…

Cited by 0SourceScholar
2025

ViewActive: Active viewpoint optimization from a single image

IROS 2025

When observing objects, humans benefit from their spatial visualization and mental rotation ability to envision potential optimal viewpoints based on the current observation. This capability is crucial for enabling robots to achieve efficient and robust scene perception during operation, as optimal

Cited by 4SourcecodeScholar
2024

AdaSwitch: Adaptive Switching between Small and Large Agents for Effective Cloud-Local Collaborative Learning

EMNLP 2024main

Recent advancements in large language models (LLMs) have been remarkable. Users face a choice between using cloud-based LLMs for generation quality and deploying local-based LLMs for lower computational cost. The former option is typically costly and inefficient, while the latter usually fails to de…

Cited by 2SourcePDFScholar
2024

Cross-model Control: Improving Multiple Large Language Models in One-time Training

NeurIPS 2024poster

The number of large language models (LLMs) with varying parameter scales and vocabularies is increasing. While they deliver powerful performance, they also face a set of common optimization needs to meet specific requirements or standards, such as instruction following or avoiding the output of sens…

2024

Event3DGS: Event-Based 3D Gaussian Splatting for High-Speed Robot Egomotion

CoRL 2024poster

By combining differentiable rendering with explicit point-based scene representations, 3D Gaussian Splatting (3DGS) has demonstrated breakthrough 3D reconstruction capabilities. However, to date 3DGS has had limited impact on robotics, where high-speed egomotion is pervasive: Egomotion introduc…

Cited by 11SourceScholar
2024

MARVIS: Motion & Geometry Aware Real and Virtual Image Segmentation

IROS 2024poster

Tasks such as autonomous navigation, 3D reconstruction, and object recognition near the water surfaces are crucial in marine robotics applications. However, challenges arise due to dynamic disturbances, e.g., light reflections and refraction from the random air-water interface, irregular liquid flow…

Cited by 3SourcecodeScholar
2024

Structure-aware Fine-tuning for Code Pre-trained Models

COLING 2024main

Over the past few years, we have witnessed remarkable advancements in Code Pre-trained Models (CodePTMs). These models achieved excellent representation capabilities by designing structure-based pre-training tasks for code. However, how to enhance the absorption of structural knowledge when fine-tun…

Cited by 1SourcePDFScholar
2023

UDepth: Fast Monocular Depth Estimation for Visually-guided Underwater Robots

ICRA 2023poster

In this paper, we present a fast monocular depth estimation method for enabling 3D perception capabilities of low-cost underwater robots. We formulate a novel end-to-end deep visual learning pipeline named UDepth, which incorporates domain knowledge of image formation characteristics of natural unde…

Cited by 51SourcecodeScholar