← Search

Zhizhuo Yang

2 accepted papers

2025

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection

ACL 2025finding

We present a novel pipeline, ReflectEvo, to demonstrate that small language models (SLMs) can enhance meta introspection through reflection learning. This process iteratively generates self-reflection for self-training, fostering a continuous and self-evolving process. Leveraging this pipeline, we c…

Cited by 0SourcePDFScholar
2025

SR-AIF: Solving Sparse-Reward Robotic Tasks From Pixels with Active Inference and World Models

ICRA 2025

Although research has produced promising results demonstrating the utility of active inference (AIF) in Markov decision processes (MDPs), there is relatively less work that builds AIF models in the context of environments and problems that take the form of partially observable Markov decision proces

Cited by 10SourceScholar