An MRP Formulation for Supervised Learning: Generalized Temporal Difference Learning Models (Abstract Reprint)
Background: Traditional supervised learning (SL) assumes data points are independently and identically distributed (i.i.d.), which overlooks dependencies in real-world data. Reinforcement learning (RL), in contrast, models dependencies through state transitions. Objectives: This study aims to bridge