0
MemRL: Self-Evolving Agents via Runtime Reinforcement Learning outperforms RAG
(arxiv.org)researchby sonney8mo ago0 comments

New framework beats RAG by 56% on ALFWorld. Uses two-phase retrieval with learned Q-values instead of passive semantic matching. No fine-tuning needed.

0 comments

No comments yet. Be the first to share your thoughts!