← Search

Hongli Sun

2 accepted papers

2025

MedOdyssey: A Medical Domain Benchmark for Long Context Evaluation Up to 200K Tokens

NAACL 2025findings

Numerous advanced Large Language Models (LLMs) now support context lengths up to 128K, and some extend to 200K. Some benchmarks in the generic domain have also followed up on evaluating long-context capabilities. In the medical domain, tasks are distinctive due to the unique contexts and need for do…

2025

Text-to-ES Bench: A Comprehensive Benchmark for Converting Natural Language to Elasticsearch Query

ACL 2025long

Elasticsearch (ES) is a distributed RESTful search engine optimized for large-scale and long-text search scenarios. Recent research on text-to-Query has explored using large language models (LLMs) to convert user query intent to executable code, making it an increasingly popular research topic. To o…

Cited by 0SourcePDFScholar