Policy regularization as a unifying theory of the striatal division of labor in learning
Learning novel behaviors requires balancing previously learned actions with the ability to flexibly adapt to changing reward contingencies. This trade-off is well documented in the division of labor between dorsolateral striatum (DLS), which promotes selection of cached, history-dependent actions, and dorsomedial striatum (DMS), which supports flexible learning as reward contingencies change.…
We haven't written up this one. bioRxiv has the full story — the link below goes straight to it.