← Search

Tolga Ok

2 accepted papers

2025

Rank-One Modified Value Iteration

ICML 2025poster

In this paper, we provide a novel algorithm for solving planning and learning problems of Markov decision processes. The proposed algorithm follows a policy iteration-type update by using a rank-one approximation of the transition probability matrix in the policy evaluation step. This rank-one app…

Cited by 0SourcePDFScholar
2024

Scalable Kernel Inverse Optimization

NeurIPS 2024poster

Inverse Optimization (IO) is a framework for learning the unknown objective function of an expert decision-maker from a past dataset. In this paper, we extend the hypothesis class of IO objective functions to a reproducing kernel Hilbert space (RKHS), thereby enhancing feature representation to an i…