← Search

Eugene Ie

8 accepted papers

2026

Beyond Markovian: Reflective Exploration via Bayes-Adaptive RL for LLM Reasoning

ICLR 2026poster

Large Language Models (LLMs) trained via Reinforcement Learning (RL) have exhibited strong reasoning capabilities and emergent reflective behaviors, such as rethinking and error correction, as a form of in-context exploration. However, the Markovian policy obtained from conventional RL training does…

Cited by 0SourcecodeScholar
2025

Chatbot Arena Estimate: towards a generalized performance benchmark for LLM capabilities

NAACL 2025industry

In industrial LLM development, evaluating large language models (LLMs) is critical for tasks like benchmarking internal models and detecting regressions during fine-tuning, but existing benchmark aggregation methods, such as Elo-based systems, can be resource-intensive, public facing, and time-consu…

2025

Improving Rectified Flow with Boundary Conditions

ICCV 2025poster

Rectified Flow offers a simple and effective approach to high-quality generative modeling by learning a velocity field. However,we identify a limitation in directly modeling the velocity with an unconstrained neural network: the learned velocity often fails to satisfy certain boundary conditions, le…

Cited by 0SourcePDFScholar
2024

Improving Multi-Agent Debate with Sparse Communication Topology

EMNLP 2024finding

Multi-agent debate has proven effective in improving large language models quality for reasoning and factuality tasks. While various role-playing strategies in multi-agent debates have been explored, in terms of the communication among agents, existing approaches adopt a brute force algorithm – each…

Cited by 19SourcePDFScholar
2023

Pedestrian Crossing Action Recognition and Trajectory Prediction with 3D Human Keypoints

ICRA 2023poster

Accurate understanding and prediction of human behaviors are critical prerequisites for autonomous vehicles, especially in highly dynamic and interactive scenarios such as intersections in dense urban areas. In this work, we aim at identifying crossing pedestrians and predicting their future traject…

Cited by 19SourceScholar
2020

Environment-agnostic Multitask Learning for Natural Language Grounded Navigation

ECCV 2020poster

Recent research efforts enable study for natural language grounded navigation in photo-realistic environments, e.g., following natural language instructions or dialog. However, existing methods tend to overfit training data in seen environments and fail to generalize well in previously unseen enviro…

2019

Transferable Representation Learning in Vision-and-Language Navigation

ICCV 2019poster

Vision-and-Language Navigation (VLN) tasks such as Room-to-Room (R2R) require machine agents to interpret natural language instructions and learn to act in visually realistic environments to achieve navigation goals. The overall task requires competence in several perception problems: successful age…

Cited by 101PDFScholar