← Search

Fan Lin

4 accepted papers

2025

Diagnosing Failures in Large Language Models’ Answers: Integrating Error Attribution into Evaluation Framework

ACL 2025finding

With the widespread application of Large Language Models (LLMs) in various tasks, the mainstream LLM platforms generate massive user-model interactions daily. In order to efficiently analyze the performance of models and diagnose failures in their answers, it is essential to develop an automated fra…

2024

IDGen: Item Discrimination Induced Prompt Generation for LLM Evaluation

NeurIPS 2024poster

As Large Language Models (LLMs) become more capable of handling increasingly complex tasks, the evaluation set must keep pace with these advancements to ensure it remains sufficiently discriminative. Item Discrimination (ID) theory, which is widely used in educational assessment, measures the abilit…

2024

Trend-Heuristic Reinforcement Learning Framework for News-Oriented Stock Portfolio Management

ICASSP 2024accepted

Recent studies have shown that reinforcement learning (RL) methods have brought significant performance gains for stock portfolio management (PM) because they effectively utilize historical price information and directly generate portfolio weights. We found, however, that there is still great room f…

Cited by 0SourceScholar