← Search

Porter Jenkins

5 accepted papers

2026

ICPO: Provable and Practical In-Context Policy Optimization for Test-Time Scaling

ICLR 2026poster

We study test-time scaling, where a model improves its answer through multi-round self-reflection at inference. We introduce In-Context Policy Optimization (ICPO), in which an agent optimizes its response in context using self-assessed or externally observed rewards without modifying its parameters.…

Cited by 0SourceScholar
2025

Fully Heteroscedastic Count Regression with Deep Double Poisson Networks

ICML 2025poster

Neural networks capable of accurate, input-conditional uncertainty representation are essential for real-world AI systems. Deep ensembles of Gaussian networks have proven highly effective for continuous regression due to their ability to flexibly represent aleatoric uncertainty via unrestricted hete…

Cited by 1SourcePDFScholar
2024

Probabilistic Offline Policy Ranking with Approximate Bayesian Computation

AAAI 2024technical

In practice, it is essential to compare and rank candidate policies offline before real-world deployment for safety and reliability. Prior work seeks to solve this offline policy ranking (OPR) problem through value-based methods, such as Off-policy evaluation (OPE). However, they fail to analyze spe…

2021

Neural Utility Functions

AAAI 2021technical

Current neural network architectures have no mechanism for explicitly reasoning about item trade-offs. Such trade-offs are important for popular tasks such as recommendation. The main idea of this work is to give neural networks inductive biases that are inspired by economic theories. To this end, w…