ICML 2020poster12 citations

Bisection-Based Pricing for Repeated Contextual Auctions against Strategic Buyer

Anton Zhiyanov, Alexey Drutsa

Abstract

We are interested in learning algorithms that optimize revenue in repeated contextual posted-price auctions where a single seller faces a single strategic buyer. In our setting, the buyer maximizes his expected cumulative discounted surplus, and his valuation of a good is assumed to be a fixed function of a $d$-dimensional context (feature) vector. We introduce a novel deterministic learning algorithm that is based on ideas of the Bisection method and has strategic regret upper bound of $O(\log^2 T)$. Unlike previous works, our algorithm does not require any assumption on the distribution of context information, and the regret guarantee holds for any realization of feature vectors (adversarial upper bound). To construct our algorithm we non-trivially adopted techniques of integral geometry to act against buyer strategicness and improved the penalization trick to work in contextual auctions.

BibTeX
@InProceedings{pmlr-v119-zhiyanov20a,
  title = 	 {Bisection-Based Pricing for Repeated Contextual Auctions against Strategic Buyer},
  author =       {Zhiyanov, Anton and Drutsa, Alexey},
  booktitle = 	 {Proceedings of the 37th International Conference on Machine Learning},
  pages = 	 {11469--11480},
  year = 	 {2020},
  editor = 	 {III, Hal Daumé and Singh, Aarti},
  volume = 	 {119},
  series = 	 {Proceedings of Machine Learning Research},
  month = 	 {13--18 Jul},
  publisher =    {PMLR},
  pdf = 	 {http://proceedings.mlr.press/v119/zhiyanov20a/zhiyanov20a.pdf},
  url = 	 {https://proceedings.mlr.press/v119/zhiyanov20a.html},
  abstract = 	 {We are interested in learning algorithms that optimize revenue in repeated contextual posted-price auctions where a single seller faces a single strategic buyer. In our setting, the buyer maximizes his expected cumulative discounted surplus, and his valuation of a good is assumed to be a fixed function of a $d$-dimensional context (feature) vector. We introduce a novel deterministic learning algorithm that is based on ideas of the Bisection method and has strategic regret upper bound of $O(\log^2 T)$. Unlike previous works, our algorithm does not require any assumption on the distribution of context information, and the regret guarantee holds for any realization of feature vectors (adversarial upper bound). To construct our algorithm we non-trivially adopted techniques of integral geometry to act against buyer strategicness and improved the penalization trick to work in contextual auctions.}
}
Bisection-Based Pricing for Repeated Contextual Auctions against Strategic Buyer · ICML 2020