Skip to content

ahmedehabb

Memory-R2

Based on Qwen/Qwen2.5-7B-Instruct

Tags

  • safetensors
  • qwen2
  • memory
  • long-horizon
  • reinforcement-learning
  • grpo
  • agent
  • conversational
  • arxiv:2605.21768
  • base_model:Qwen/Qwen2.5-7B-Instruct
  • base_model:finetune:Qwen/Qwen2.5-7B-Instruct
  • license:apache-2.0
Author
ahmedehabb
Task
text-generation
Architecture
qwen2
License
apache-2.0
Base model
Qwen/Qwen2.5-7B-Instruct
Downloads
247
Likes
0
Memory-R2 — AI Model — AIMarketly