2025
RLPF: Reinforcement Learning from Prediction Feedback for User Summarization with LLMs
AAAI 2025technical
LLM-powered personalization agent systems employ Large Language Models (LLMs) to predict users’ behavior from their past activities. However, their effectiveness often hinges on the ability to effectively leverage extensive, long user historical data due to its inherent noise and length of such data…