← Search

Adrian Cosma

5 accepted papers

2026

On Model and Data Scaling for Skeleton-based Self-Supervised Gait Recognition

AAAI 2026technical

Gait recognition from video streams is a challenging problem in computer vision biometrics due to the subtle differences between gaits and numerous confounding factors. Recent advancements in self-supervised pretraining have led to the development of robust gait recognition models that are invariant

Cited by 0SourcePDFScholar
2025

The Strawberry Problem: Emergence of Character-level Understanding in Tokenized Language Models

EMNLP 2025

Despite their remarkable progress across diverse domains, Large Language Models (LLMs) consistently fail at simple character-level tasks, such as counting letters in words, due to a fundamental limitation: tokenization. In this work, we frame this limitation as a problem of low mutual information an

2024

How Hard is this Test Set? NLI Characterization by Exploiting Training Dynamics

EMNLP 2024main

Natural Language Inference (NLI) evaluation is crucial for assessing language understanding models; however, popular datasets suffer from systematic spurious correlations that artificially inflate actual model performance. To address this, we propose a method for the automated creation of a challeng…

2024

RoCode: A Dataset for Measuring Code Intelligence from Problem Definitions in Romanian

COLING 2024main

Recently, large language models (LLMs) have become increasingly powerful and have become capable of solving a plethora of tasks through proper instructions in natural language. However, the vast majority of testing suites assume that the instructions are written in English, the de facto prompting la…

2020

Black-Box Ripper: Copying black-box models using generative evolutionary algorithms

NeurIPS 2020oral

We study the task of replicating the functionality of black-box neural models, for which we only know the output class probabilities provided for a set of input images. We assume back-propagation through the black-box model is not possible and its training images are not available, e.g. the model co…