← Search

Graham Horwood

4 accepted papers

2025

Active Evaluation Acquisition for Efficient LLM Benchmarking

ICML 2025poster

As large language models (LLMs) become increasingly versatile, numerous large scale benchmarks have been developed to thoroughly assess their capabilities. These benchmarks typically consist of diverse datasets and prompts to evaluate different aspects of LLM performance. However, comprehensive eval…

Cited by 1SourcePDFScholar
2025

MetaSynth: Meta-Prompting-Driven Agentic Scaffolds for Diverse Synthetic Data Generation

ACL 2025finding

Recent smaller language models such Phi-3.5 and Phi-4 rely on synthetic data generated using larger Language models. Questions remain about leveraging synthetic data for other use cases, such as adapting LLMs to specific domains. A key limitation of synthetic data is low diversity, which negatively…

Cited by 0SourcePDFScholar
2023

Contrastive Training Improves Zero-Shot Classification of Semi-structured Documents

ACL 2023findings

We investigate semi-structured document classification in a zero-shot setting. Classification of semi-structured documents is more challenging than that of standard unstructured documents, as positional, layout, and style information play a vital role in interpreting such documents. The standard cla…

2022

Contrastive Representation Learning for Cross-Document Coreference Resolution of Events and Entities

NAACL 2022long

Identifying related entities and events within and across documents is fundamental to natural language understanding. We present an approach to entity and event coreference resolution utilizing contrastive representation learning. Earlier state-of-the-art methods have formulated this problem as a bi…