← Search

Richard Ren

3 accepted papers

2025

Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs

NeurIPS 2025spotlight

As AIs rapidly advance and become more agentic, the risk they pose is governed not only by their capabilities but increasingly by their propensities, including goals and values. Tracking the emergence of goals and values has proven a longstanding problem, and despite much interest over the years it…

Cited by 0SourceScholar
2024

Safetywashing: Do AI Safety Benchmarks Actually Measure Safety Progress?

NeurIPS 2024poster

Performance on popular ML benchmarks is highly correlated with model scale, suggesting that most benchmarks tend to measure a similar underlying factor of general model capabilities. However, substantial research effort remains devoted to designing new benchmarks, many of which claim to measure nove…

Cited by 22SourcecodeScholar
2023

Deep Reinforcement Learning for Decentralized Multi-Robot Exploration With Macro Actions

RA-L 2023

Cooperative multi-robot teams need to be able to explore cluttered and unstructured environments while dealing with communication dropouts that prevent them from exchanging local information to maintain team coordination. Therefore, robots need to consider high-level teammate intentions during actio

Cited by 53SourceScholar