2025
Compute-Optimal Scaling for Value-Based Deep RL
NeurIPS 2025poster
As models grow larger and training them becomes expensive, it becomes increasingly important to scale training recipes not just to larger models and more data, but to do so in a compute-optimal manner that extracts maximal performance per unit of compute. While such scaling has been well studied for…