← Search

András Balogh

3 accepted papers

2026

Verification of the Implicit World Model in a Generative Model via Adversarial Sequences

ICLR 2026poster

Generative sequence models are typically trained on sample sequences from natural or formal languages. It is a crucial question whether—or to what extent—sample-based training is able to capture the true structure of these languages, often referred to as the "world model". Theoretical re…

Cited by 0SourcecodeScholar
2025

How Not to Stitch Representations to Measure Similarity: Task Loss Matching Versus Direct Matching

AAAI 2025technical

Measuring the similarity of the internal representations of deep neural networks is an important and challenging problem. Model stitching has been proposed as a possible approach, where two half-networks are connected by mapping the output of the first half-network to the input of the second one. Th…