Downloads · 30 days
13
30% of all-time downloads
Jayi2424/HumorGen_GRPO_7B
HumorGen_GRPO_7B is a text generation model from Jayi2424. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as apache-2.0.
Part of the HumorGen Collection · SaLT Lab, Carnegie Mellon University
Downloads · 30 days
13
30% of all-time downloads
All-time downloads
43
Public
Repo size
173 MB
Likes
0
Public
Click a slice to open those files.
.safetensors162 MB · 91%
From the Hugging Face model README
Part of the HumorGen Collection · SaLT Lab, Carnegie Mellon University
Offline Group Relative Policy Optimization (O-GRPO) on the SFT checkpoint. Trains on the full group of six persona candidates per headline rather than a single binary pair.
Paper(s): arXiv:2604.09629
| Property | Value |
|---|---|
| Stage | Offline Group Relative Policy Optimization (O-GRPO) |
| Initialized from | HumorGen_SFT_7B |
| Backbone | Qwen2.5-7B-Instruct (QLoRA 4-bit) |
| Group size | 6 (one per CSF persona) |
| Reward | HumorRank Bradley-Terry scores |
This is a PEFT LoRA adapter. Load the base model and apply the adapter:
from transformers import AutoModelForCausalLM, AutoTokenizer
from peft import PeftModel
import torch
tokenizer = AutoTokenizer.from_pretrained("Qwen/Qwen2.5-7B-Instruct")
model = AutoModelForCausalLM.from_pretrained("Qwen/Qwen2.5-7B-Instruct", torch_dtype=torch.bfloat16, device_map="auto")
model = PeftModel.from_pretrained(model, "Jayi2424/HumorGen_GRPO_7B")
headline = "AI writes entire novel; readers say it felt a little too human"
prompt = (
"<|im_start|>system\n"
"You are a comedy writer. Write one sharp, witty joke for the headline.\n<|im_end|>\n"
f"<|im_start|>user\n{headline}<|im_end|>\n"
"<|im_start|>assistant\n"
)
inputs = tokenizer(prompt, return_tensors="pt").to(model.device)
outputs = model.generate(**inputs, max_new_tokens=120, temperature=0.9, top_p=0.95)
print(tokenizer.decode(outputs[0], skip_special_tokens=True))
@misc{ajayi2026humorgen,
title = {HumorGen: Cognitive Synergy for Humor Generation in Large Language
Models via Persona-Based Distillation},
author = {Ajayi, Edward and others},
year = {2026},
eprint = {2604.09629},
archivePrefix = {arXiv},
primaryClass = {cs.CL},
url = {https://arxiv.org/abs/2604.09629}
}