Downloads · 30 days
29
12% of all-time downloads
yaoyueduzhen/RAG-R1-sq-7b
RAG-R1-sq-7b is a machine learning model from yaoyueduzhen. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
Model Name: RAG-R1-sq-7b Version: 1.0 Model Type: RAG Developers: Zhiwen Tan, Jiaming Huang, Qintong Wu, Hongxuan Zhang, Chenyi Zhuang, Jinjie Gu
Downloads · 30 days
29
12% of all-time downloads
All-time downloads
245
Public
Parameters
7.6B
30.5 GB on disk
Likes
2
Public
Click a slice to open those files.
.safetensors30.5 GB · 100%
From the Hugging Face model README
RAG-R1 is a deepsearch training framework designed to enable LLMs to adaptively leverage internal and external knowledge during the reasoning process. We further expand the generation and retrieval processes within the framework from single-query mode to multi-query parallelism, aimed at reducing inference time and enhancing the model's capabilities. Extensive experiments on seven question-answering benchmarks demonstrate that our method outperforms the strongest baseline by up to 13.2% and decreases inference time by 11.1%.
RAG-R1 is inspired by Deepseek-R1 with its implementation based on veRL and Search-r1. We deeply appreciate the contributions of these teams to open-source research and development.
Please cite our repo if our works are helpful for your research.
@article{RAG-R1,
title={RAG-R1 : Incentivize the Search and Reasoning Capabilities of LLMs through Multi-query Parallelism},
author={Zhiwen Tan and Jiaming Huang and Qintong Wu and Hongxuan Zhang and Chenyi Zhuang and Jinjie Gu},
journal={arXiv preprint arXiv:2507.02962},
year={2025}
}