← Search

Yanzhao Zheng

4 accepted papers

2026

SERL: Self-Examining Reinforcement Learning on Open-Domain

AAAI 2026technical

Reinforcement Learning (RL) has been shown to improve the capabilities of large language models (LLMs). However, applying RL to open-domain tasks faces two key challenges: (1) the inherent subjectivity of these tasks prevents the verifiable rewards as required by Reinforcement Learning with Verifiab

Cited by 0SourcePDFScholar
2025

Reason from Future: Reverse Thought Chain Enhances LLM Reasoning

ACL 2025finding

It has been demonstrated that carefully designed reasoning paradigms, like Chain-of-Thought(CoT) and Tree-of-Thought(ToT), can enhance the reasoning capabilities of small language models by detailed thinking and extensive thought searching, unbounded branching factors in the searching space create p…

Cited by 0SourcePDFScholar
2023

COOP: Decoupling and Coupling of Whole-Body Grasping Pose Generation

ICCV 2023poster

Generating life-like whole-body human grasping has garnered significant attention in the field of computer graphics. Existing works have demonstrated the effectiveness of keyframe-guided motion generation framework, witch focus on modeling the grasping motions of humans in temporal sequence when the…

Cited by 7PDFcodeScholar
2022

HIE-SQL: History Information Enhanced Network for Context-Dependent Text-to-SQL Semantic Parsing

ACL 2022findings

Recently, context-dependent text-to-SQL semantic parsing which translates natural language into SQL in an interaction process has attracted a lot of attentions. Previous works leverage context dependence information either from interaction history utterances or previous predicted queries but fail in…

Cited by 35SourcePDFScholar