← Search

Sylvain Kubler

2 accepted papers

2026

CSPO: Constraint-Sensitive Policy Optimization for Safe Reinforcement Learning

ICML 2026spotlight

Safe reinforcement learning (Safe RL) aims to maximize expected return while satisfying safety constraints, typically modeled as constrained Markov decision processes. While primal-dual methods scale well to deep RL, they often suffer from delayed constraint correction, leading to oscillatory behavi…

Cited by 0SourceScholar
2026

Counterfactual eXplainable AI (XAI) Method for Deep Learning-Based Multivariate Time Series Classification

AAAI 2026technical

Recent advances in deep learning have improved multivariate time series (MTS) classification and regression by capturing complex patterns, but their lack of transparency hinders decision-making. Explainable AI (XAI) methods offer partial insights, yet often fall short of conveying the full decision

Cited by 0SourcePDFScholar