← Search

Zilin Kang

3 accepted papers

2026

Entropy Regularizing Activation: Boosting Continuous Control, Large Language Models, and Image Classification with Activation as Entropy Constraints

ICLR 2026poster

We propose ERA, a new paradigm for entropy-constrained policy via output activation. It guarantees minimum sampling entropy by transforming the outputs of the last layer. Our approach demonstrates broad effectiveness across different domains: 1) for large language models~(LLMs), boosting the AIME 20…

Cited by 0SourcecodeScholar
2026

Euclean: Automated Geometry Problem Formalization with Unified Verification in Lean

ICML 2026poster

Recent formal reasoning systems achieve IMO-level performance, but create a fragmented landscape: algebra and number theory use Lean, while geometry relies on domain-specific languages with limited formal guarantees. This fragmentation increases the trusted computing base and hinders unified model d…

Cited by 0SourceScholar
2025

A Forget-and-Grow Strategy for Deep Reinforcement Learning Scaling in Continuous Control

ICML 2025poster

Deep reinforcement learning for continuous control has recently achieved impressive progress. However, existing methods often suffer from primacy bias—a tendency to overfit early experiences stored in the replay buffer—which limits an RL agent’s sample efficiency and generalizability. A common exist…

Cited by 0SourcePDFScholar