← Search

Wanying Wang

3 accepted papers

2026

TESTAGENT: AUTOMATIC BENCHMARKING AND EXPLORATORY INTERACTION FOR EVALUATING LLMS IN VERTICAL DOMAINS

ICASSP 2026oral

As Large Language Models (LLMs) are increasingly deployed in highly specialized vertical domains, the evaluation of their domain-specific performance becomes critical. However, existing evaluations for vertical domains typically rely on the labor-intensive construction of static single-turn datasets…

Cited by 0SourcePDFScholar
2024

Exploring Gradient Explosion in Generative Adversarial Imitation Learning: A Probabilistic Perspective

AAAI 2024technical

Generative Adversarial Imitation Learning (GAIL) stands as a cornerstone approach in imitation learning. This paper investigates the gradient explosion in two types of GAIL: GAIL with deterministic policy (DE-GAIL) and GAIL with stochastic policy (ST-GAIL). We begin with the observation that the tra…

Cited by 6SourcePDFScholar