← Search

Ding Chen

3 accepted papers

2025

GuessArena: Guess Who I Am? A Self-Adaptive Framework for Evaluating LLMs in Domain-Specific Knowledge and Reasoning

ACL 2025long

The evaluation of large language models (LLMs) has traditionally relied on static benchmarks, a paradigm that poses two major limitations: (1) predefined test sets lack adaptability to diverse application domains, and (2) standardized evaluation protocols often fail to capture fine-grained assessmen…

Cited by 0SourcePDFScholar
2025

xFinder: Large Language Models as Automated Evaluators for Reliable Evaluation

ICLR 2025poster

The continuous advancement of large language models (LLMs) has brought increasing attention to the critical issue of developing fair and reliable methods for evaluating their performance. Particularly, the emergence of cheating phenomena, such as test set leakage and prompt format overfitting, poses…

Cited by 0SourcePDFScholar
2023

Parallel Spiking Neurons with High Efficiency and Ability to Learn Long-term Dependencies

NeurIPS 2023poster

Vanilla spiking neurons in Spiking Neural Networks (SNNs) use charge-fire-reset neuronal dynamics, which can only be simulated serially and can hardly learn long-time dependencies. We find that when removing reset, the neuronal dynamics can be reformulated in a non-iterative form and parallelized. B…