Satinder P. Singh; Tommi Jaakkola; Michael I. Jordan
Reinforcement Learning with Soft State Aggregation