Downloads · 30 days
6
8% of all-time downloads
HunterJiang97/PABU-Agent-8B
PABU-Agent-8B is a question answering model from HunterJiang97. Use it when the input is a question plus a passage. The card lists the license as mit.
<p align="center" 📃 <a href="https://arxiv.org/pdf/2602.09138" target="blank"Paper</a • 🌐 <a href="https://pabu-agent.github.io/" target="blank"Project Page</a • 🤗 <a href="https://huggingface.co/datasets/HunterJia…
Downloads · 30 days
6
8% of all-time downloads
All-time downloads
77
Public
Parameters
8B
16.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors16.1 GB · 100%
From the Hugging Face model README
PABU-Agent-8B is a Large Language Model (LLM) agent built on top of LLaMA‑3.1‑8B, fine-tuned for interactive decision making using step-level supervision from the PABU Dataset. The model is trained to operate in sequential action–observation environments while maintaining a compact belief state via Progress-Aware Belief Update (PABU).
Instead of conditioning on full interaction histories, the model learns to predict relative task progress at each step and selectively retain informative past interactions. This results in improved task completion and reduced interaction length across diverse long-horizon environments.
Users should evaluate the model in their target environment and avoid extrapolating performance gains beyond AgentGym-style tasks.
The model is intended to be used within an agent loop that alternates between observations and actions and maintaining a belief memory buffer as described in PABU.
@misc{jiang2026pabuprogressawarebeliefupdate,
title={PABU: Progress-Aware Belief Update for Efficient LLM Agents},
author={Haitao Jiang and Lin Ge and Hengrui Cai and Rui Song},
year={2026},
eprint={2602.09138},
archivePrefix={arXiv},
primaryClass={cs.AI},
url={https://arxiv.org/abs/2602.09138},
}