Conduit: An Experience Data Plane for Distributed Reinforcement Learning
Distributed reinforcement learning (RL) scales training by parallelizing actors and learners around an Experience Buffer. As RL workloads grow, however, the buffer becomes more than a replay queue: it is the storage substrate of a large-capacity, latency-critical experience path that every iteration traverses to move, transform, sample, and batch experiences before learner updates can begin.…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.