← Search

Min He

2 accepted papers

2026

Decoupling Reasoning and Confidence: Resurrecting Calibration in Reinforcement Learning from Verifiable Rewards

ICML 2026poster

Reinforcement Learning from Verifiable Rewards (RLVR) significantly enhances large language models (LLMs) reasoning but severely suffers from calibration degeneration, where models become excessively over-confident in incorrect answers. Previous studies devote to directly incorporating calibration o…

Cited by 0SourceScholar
2023

Exploiting CCTV Cameras for Hand Hygiene Recognition in ICU

ICASSP 2023accepted

The monitoring of hand hygiene activities can effectively reduce infection and contamination in the Intensive Care Unit (ICU). In this paper, we created a clinical dataset using CCTV cameras installed in ICU to explore the feasibility of recognizing the hand-washing steps of clinicians. A video proc…

Cited by 0SourceScholar