17 upvotes · 23 JUL 2026 · Chenhui Gou, Haoqin Tu, Yunhao Fang et al.
This paper develops a method called Experience Distillation that allows agents to learn from their own interaction histories without needing additional environment interactions, making learning more sample-efficient. Practitioners might care about this because it can improve the performance of agents in complex environments with limited resources.
17 upvotes · 17 JUL 2026 · Wentao Zhang, Xuanhe Pan, Han Zhou et al.
This paper introduces a new system for controlling live machine learning model training, allowing humans and automated controllers to interact with the training process through a shared interface, and providing a clear audit trail of changes and outcomes. Practitioners might care because it enables more transparent and accountable training processes.