Downloads ยท 30 days
22
100% of all-time downloads
chibifire/EditScore-7B
EditScore-7B is a text generation model from chibifire. Use it when you need the model to write or continue text. It is set up for peft.
<p align="center" <img src="assets/logo.png" width="65%" </p
Downloads ยท 30 days
22
100% of all-time downloads
All-time downloads
22
Public
Repo size
324 MB
Likes
0
Public
Click a slice to open those files.
.safetensors323 MB ยท 100%
From the Hugging Face model README
EditScore is a series of state-of-the-art open-source reward models (7Bโ72B) designed to evaluate and enhance instruction-guided image editing.
While Reinforcement Learning (RL) holds immense potential for this domain, its progress has been severely hindered by the absence of a high-fidelity, efficient reward signal.
To overcome this barrier, we provide a systematic, two-part solution:
A Rigorous Evaluation Standard: We first introduce EditReward-Bench, a new public benchmark for the direct and reliable evaluation of reward models. It features 13 diverse subtasks and expert human annotations, establishing a gold standard for measuring reward signal quality.
A Powerful & Versatile Tool: Guided by our benchmark, we developed the EditScore model series. Through meticulous data curation and an effective self-ensembling strategy, EditScore sets a new state of the art for open-source reward models, even surpassing the accuracy of leading proprietary VLMs.
We demonstrate the practical utility of EditScore through two key applications:
This repository releases both the EditScore models and the EditReward-Bench dataset to facilitate future research in reward modeling, policy optimization, and AI-driven model improvement.
<p align="center"> <img src="assets/figure_edit_results.png" width="95%"> <br> <em>EditScore as a superior reward signal for image editing.</em> </p>We are actively working on improving EditScore and expanding its capabilities. Here's what's next:
# 1. Clone the repo
git clone [email protected]:VectorSpaceLab/EditScore.git
cd EditScore
# 2. (Optional) Create a clean Python environment
conda create -n editscore python=3.12
conda activate editscore
# 3. Install dependencies
# 3.1 Install PyTorch (choose correct CUDA version)
pip install torch==2.7.1 torchvision --extra-index-url https://download.pytorch.org/whl/cu126
# 3.2 Install other required packages
pip install -r requirements.txt
# EditScore runs even without vllm, though we recommend install it for best performance.
pip install vllm
# Install PyTorch from a domestic mirror
pip install torch==2.7.1 torchvision --index-url https://mirror.sjtu.edu.cn/pytorch-wheels/cu126
# Install other dependencies from Tsinghua mirror
pip install -r requirements.txt -i https://pypi.tuna.tsinghua.edu.cn/simple
# EditScore runs even without vllm, though we recommend install it for best performance.
pip install vllm -i https://pypi.tuna.tsinghua.edu.cn/simple
Using EditScore is straightforward. The model will be automatically downloaded from the Hugging Face Hub on its first run.
from PIL import Image
from editscore import EditScore
# Load the EditScore model. It will be downloaded automatically.
# Replace with the specific model version you want to use.
model_path = "Qwen/Qwen2.5-VL-7B-Instruct"
lora_path = "EditScore/EditScore-7B"
scorer = EditScore(
backbone="qwen25vl", # set to "qwen25vl_vllm" for faster inference
model_name_or_path=model_path,
enable_lora=True,
lora_path=lora_path,
score_range=25,
num_pass=1, # Increase for better performance via self-ensembling
)
input_image = Image.open("example_images/input.png")
output_image = Image.open("example_images/output.png")
instruction = "Adjust the background to a glass wall."
result = scorer.evaluate([input_image, output_image], instruction)
print(f"Edit Score: {result['final_score']}")
# Expected output: A dictionary containing the final score and other details.
We provide an evaluation script to benchmark reward models on EditReward-Bench. To evaluate your own custom reward model, simply create a scorer class with a similar interface and update the script.
# This script will evaluate the default EditScore model on the benchmark
bash evaluate.sh
# Or speed up inference with VLLM
bash evaluate_vllm.sh
If you find this repository or our work useful, please consider giving a star โญ and citation ๐ฆ, which would be greatly appreciated:
@article{luo2025editscore,
title={EditScore: Unlocking Online RL for Image Editing via High-Fidelity Reward Modeling},
author={Xin Luo and Jiahao Wang and Chenyuan Wu and Shitao Xiao and Xiyan Jiang and Defu Lian and Jiajun Zhang and Dong Liu and Zheng Liu},
journal={arXiv preprint arXiv:2509.23909},
year={2025}
}