0
MemRL: Self-Evolving Agents via Runtime Reinforcement Learning outperforms RAG
New framework beats RAG by 56% on ALFWorld. Uses two-phase retrieval with learned Q-values instead of passive semantic matching. No fine-tuning needed.
0 comments
No comments yet. Be the first to share your thoughts!