← Search

Michael Hind

3 accepted papers

2025

BenchmarkCards: Standardized Documentation for Large Language Model Benchmarks

NeurIPS 2025poster

Large language models (LLMs) are powerful tools capable of handling diverse tasks. Comparing and selecting appropriate LLMs for specific tasks requires systematic evaluation methods, as models exhibit varying capabilities across different domains. However, finding suitable benchmarks is difficult gi…

Cited by 0SourcecodeScholar
2025

Granite Guardian: Comprehensive LLM Safeguarding

NAACL 2025industry

The deployment of language models in real-world applications exposes users to various risks, including hallucinations and harmful or unethical content. These challenges highlight the urgent need for robust safeguards to ensure safe and responsible AI. To address this, we introduce Granite Guardian,…

2019

Constructing and Compressing Frames in Blockchain-based Verifiable Multi-party Computation

ICASSP 2019accepted

In previous work, we proposed a scalable multi-party verification scheme for expensive iterative computations on a Blockchain substrate by appropriate storage and endorsement of frames of iterates. In this work, we extend the framework to verify sets of complete computations with different unordered…

Cited by 0SourceScholar