← Search

Md. Maruf Hossain

1 accepted papers

2025

Evaluating Large Language Models with Enterprise Benchmarks

NAACL 2025industry

The advancement of large language models (LLMs) has led to a greater challenge of having a rigorous and systematic evaluation of complex tasks performed, especially in enterprise applications. Therefore, LLMs need to be benchmarked with enterprise datasets for a variety of NLP tasks. This work explo…

Cited by 0SourcePDFScholar