← Search

Yaniv Galron

2 accepted papers

2025

Beyond Token Probes: Hallucination Detection via Activation Tensors with ACT-ViT

NeurIPS 2025poster

Detecting hallucinations in Large Language Model-generated text is crucial for their safe deployment. While probing classifiers show promise, they operate on isolated layer–token pairs and are LLM-specific, limiting their effectiveness and hindering cross-LLM applications. In this paper, we introduc…

Cited by 0SourceScholar