← Search

Wooseong Cho

3 accepted papers

2024

Randomized Exploration for Reinforcement Learning with Multinomial Logistic Function Approximation

NeurIPS 2024poster

We study reinforcement learning with _multinomial logistic_ (MNL) function approximation where the underlying transition probability kernel of the _Markov decision processes_ (MDPs) is parametrized by an unknown transition core with features of state and action. For the finite horizon episodic setti…

Cited by 0SourcePDFScholar
2023

Semi-Parametric Contextual Pricing Algorithm using Cox Proportional Hazards Model

ICML 2023poster

Contextual dynamic pricing is a problem of setting prices based on current contextual information and previous sales history to maximize revenue. A popular approach is to postulate a distribution of customer valuation as a function of contextual information and the baseline valuation. A semi-paramet…

Cited by 3SourcePDFScholar