← Search

Athanasios Katsamanis

6 accepted papers

2025

Krikri: Advancing Open Large Language Models for Greek

EMNLP 2025

We introduce Llama-Krikri-8B, a cutting-edge Large Language Model tailored for the Greek language, built on Meta’s Llama 3.1-8B. Llama-Krikri-8B has been extensively trained on high-quality Greek data to ensure superior adaptation to linguistic nuances. With 8 billion parameters, it offers advanced

Cited by 0SourcePDFScholar
2023

Designing and Evaluating Speech Emotion Recognition Systems: A Reality Check Case Study with IEMOCAP

ICASSP 2023accepted

There is an imminent need for guidelines and standard test sets to allow direct and fair comparisons of speech emotion recognition (SER). While resources, such as the Interactive Emotional Dyadic Motion Capture (IEMOCAP) database, have emerged as widely-adopted reference corpora for researchers to d…

Cited by 0SourceScholar
2023

Exploring Language-Agnostic Speech Representations Using Domain Knowledge for Detecting Alzheimer's Dementia

ICASSP 2023accepted

We explore ways to use speech data to screen for indications of Alzheimer’s dementia (AD). In particular, we describe our approach to the ICASSP 2023 Signal Processing Grand Challenge, which involves extrapolating from models learned from English speech samples, to Greek speech samples, to determine…

Cited by 0SourceScholar
2018

Multi-View Audio-Articulatory Features for Phonetic Recognition on RTMRI-TIMIT Database

ICASSP 2018accepted

In this paper, we investigate the use of articulatory information, and more specifically real time Magnetic Resonance Imaging (rtMRI) data of the vocal tract, to improve speech recognition performance. For the purpose of our experiments, we use data from the rtMRI-TIMIT database. Firstly, Scale Inva…

Cited by 0SourceScholar
2016

Multimodal human action recognition in assistive human-robot interaction

ICASSP 2016accepted

Within the context of assistive robotics we develop an intelligent interface that provides multimodal sensory processing capabilities for human action recognition. Human action is considered in multimodal terms, containing inputs such as audio from microphone arrays, and visual inputs from high defi…

Cited by 0SourceScholar
2016

Towards a behaviorally-validated computational audiovisual saliency model

ICASSP 2016accepted

Computational saliency models aim at predicting, in a bottom-up fashion, where human attention is drawn in the presented (visual, auditory or audiovisual) scene and have been proven useful in applications like robotic navigation, image compression and movie summarization. Despite the fact that well-…

Cited by 0SourceScholar