Downloads · 30 days
0
ys-qu/found-rl_vlms
found-rl_vlms is a image-text-to-text model from ys-qu. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
These VLMs serve for the paper "Found-RL: Foundation Model-Enhanced Reinforcement Learning for Autonomous Driving".
Downloads · 30 days
0
Access
Public
Updated Feb 14, 2026
Repo size
66.2 GB
Likes
0
Public
Click a slice to open those files.
.safetensors60.3 GB · 91%
From the Hugging Face model README
These VLMs serve for the paper "Found-RL: Foundation Model-Enhanced Reinforcement Learning for Autonomous Driving".
In this work, we use fine-tuned VLMs to provide feedback for reinforcement learning agents in autonomous driving scenarios.
RGB + Text (LoRA SFT):
Rendered BEV + Text (Full SFT):
If you use these VLMs in your research, please cite our paper:
@misc{qu2026foundrl,
title={Found-RL: foundation model-enhanced reinforcement learning for autonomous driving},
author={Yansong Qu and Zihao Sheng and Zilin Huang and Jiancong Chen and Yuhao Luo and Tianyi Wang and Yiheng Feng and Samuel Labi and Sikai Chen},
year={2026},
eprint={2602.10458},
archivePrefix={arXiv},
primaryClass={cs.AI},
url={https://arxiv.org/abs/2602.10458},
}