Prototype reward-modulated STDP learner #155
Reference in New Issue
Block a user
Delete Branch "%!s()"
Deleting a branch is permanent. Although the deleted branch may continue to exist for a short time before it actually gets removed, it CANNOT be undone in most cases. Continue?
Question
Add the learning algorithm: reward-modulated STDP. Using the continuous learning loop (no episodes, wait-then-evaluate, shaped inverse reward), implement spike-timing tracking, eligibility traces, and reward-modulated weight updates. The SNN outputs a target angle, waits for gun to arrive, measures error, computes reward = 1.0 / (1.0 + error), and updates weights via R-STDP. Debug overlay should show error dropping over time.
Blocked by: #152 (SNN prototype must exist first)
Parent map: #147
Resolution
Reward-modulated STDP implemented:
Compiles clean. Ready for training runs.