Downloads · 30 days
419
3% of all-time downloads
BytedTsinghua-SIA/RL-MemoryAgent-7B
RL-MemoryAgent-7B is a machine learning model from BytedTsinghua-SIA. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
The RL-MemAgent-7B is a part of the MemAgent framework, which enables Large Language Models (LLMs) to process arbitrarily long texts through end-to-end Reinforcement Learning without altering their core architecture.
Downloads · 30 days
419
3% of all-time downloads
All-time downloads
14.1K
Public
Parameters
7.6B
15.2 GB on disk
Likes
8
Public
Click a slice to open those files.
.safetensors15.2 GB · 100%
From the Hugging Face model README
The RL-MemAgent-7B is a part of the MemAgent framework, which enables Large Language Models (LLMs) to process arbitrarily long texts through end-to-end Reinforcement Learning without altering their core architecture.
This model is ideal for tasks requiring the understanding and processing of very long documents, such as comprehensive question answering, summarizing extensive reports, or analyzing large codebases.
For detailed instructions on how to use, evaluate, and train models within the MemAgent framework, please refer to the main MemAgent GitHub repository.
If you find this work useful, please consider citing our paper:
@article{yu2025memagent,
title={MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent},
author={Yu, Hongli and Chen, Tinghong and Feng, Jiangtao and Chen, Jiangjie and Dai, Weinan and Yu, Qiying and Zhang, Ya-Qin and Ma, Wei-Ying and Liu, Jingjing and Wang, Mingxuan and others},
journal={arXiv preprint arXiv:2507.02259},
year={2025}
}