← Search

Vasilios Mavroudis

6 accepted papers

2026

Beyond Training-time Poisoning: Component-level and Post-training Backdoors in Deep Reinforcement Learning

AAAI 2026technical

Deep Reinforcement Learning (DRL) systems are increasingly used in safety-critical applications, yet their security remains severely underexplored. This work investigates backdoor attacks, which implant hidden triggers that cause malicious actions only when specific inputs appear in the observation

Cited by 0SourcePDFScholar
2026

DRMD: Deep Reinforcement Learning for Malware Detection Under Concept Drift

AAAI 2026technical

Malware detection in real-world settings must deal with evolving threats, limited labeling budgets, and uncertain predictions. Traditional classifiers, without additional mechanisms, struggle to maintain performance under concept drift in malware domains, as their supervised learning formulation can

Cited by 9SourcePDFScholar
2024

Online Convex Optimisation: The Optimal Switching Regret for all Segmentations Simultaneously

NeurIPS 2024spotlight

We consider the classic problem of online convex optimisation. Whereas the notion of static regret is relevant for stationary problems, the notion of switching regret is more appropriate for non-stationary problems. A switching regret is defined relative to any segmentation of the trial sequence, an…

Cited by 1SourcePDFScholar