← Search

Jiuzhou Han

7 accepted papers

2026

Uncertainty-Based Methods for Automated Process Reward Data Construction and Output Aggregation in Mathematical Reasoning

AAAI 2026technical

Large language models have demonstrated remarkable capabilities in complex mathematical reasoning tasks, but they inevitably generate errors throughout multi-step solutions. Process-level Reward Models (PRMs) have shown great promise by providing supervision and evaluation at each intermediate step,

Cited by 0SourcePDFScholar
2025

Agent S: An Open Agentic Framework that Uses Computers Like a Human

ICLR 2025poster

We present Agent S, an open agentic framework that enables autonomous interaction with computers through Graphical User Interface (GUI), aimed at transforming human-computer interaction by automating complex, multi-step tasks. Agent S addresses three key challenges in automating computer tasks: acqu…

2024

PiVe: Prompting with Iterative Verification Improving Graph-based Generative Capability of LLMs

ACL 2024findings

Large language models (LLMs) have shown great abilities of solving various natural language tasks in different domains. Due to the training objective of LLMs and their pre-training data, LLMs are not very well equipped for tasks involving structured data generation. We propose a framework, Prompting…

2023

POSQA: Probe the World Models of LLMs with Size Comparisons

EMNLP 2023long findings

Embodied language comprehension emphasizes that language understanding is not solely a matter of mental processing in the brain but also involves interactions with the physical and social environment. With the explosive growth of Large Language Models (LLMs) and their already ubiquitous presence in…

Cited by 0SourcecodeScholar