Agent-side memory steers stateless robots through long manipulation tasks
Hu et al. test whether long-horizon robot manipulation requires memory inside the action policy itself. On LIBERO-Mem, their architecture places all interaction history in a multimodal agent and achieves 76.3% average completion with an episodically stateless action model, versus 14.8% for the strongest reported baseline.