This week's #PaperILike is "Plan-based Reward Shaping for Reinforcement Learning" (Grzes & Kudenko, 2008).
This week's #PaperILike is "Plan-based Reward Shaping for Reinforcement Learning" (Grzes & Kudenko, 2008).
A nice combo of planning and RL that takes seriously the policy invariance ideas from Ng, Harada, & Russell (1999) [another paper I
How it ranks: Backlist reads my Twitter/X timeline, scores every tweet for substance with an LLM rubric (not engagement), and publishes the daily top picks with a one-line takeaway. Curated by Surya Dantuluri.