Downloads · 30 days
0
AMD-PAVS-AI/deepseek_r1
deepseek_r1 is a text generation model from AMD-PAVS-AI. Use it when you need the model to write or continue text. It is set up for vllm. The card lists the license as mit.
Downloads · 30 days
0
Access
Public
Updated Aug 4, 2026
Repo size
150 KB
Likes
1
Public
Click a slice to open those files.
.png150 KB · 96%
From the Hugging Face model README

DeepSeek-R1-Distill-Qwen-7B is a distilled reasoning language model that generates chain-of-thought answers for math and logic problems. This repository packages evaluation/inference for text reasoning / math problem solving using vLLM, exported and validated for AMD ROCm so it runs efficiently on AMD GPUs and CPUs.
This is based on the implementation of DeepSeek-R1 found here. This repository contains configurations and scripts optimized for AMD® ROCm™ platforms. You can use the deepseek_r1 AMD scripts to reproduce results or export with custom configurations. More details on model performance can be found here.
Task: Text reasoning / math problem solving
Dataset: MATH-500 (500 competition math problems); sample prompts for infer-text
Output metrics: MATH-500 accuracy
vLLM note: MATH-500 evaluation uses symbolic answer verification through
math_verify, which parses\boxed{}expressions and checks equivalence to the reference solution.
This model export has been adapted and validated for AMD Instinct™ / Radeon™ GPUs running ROCm, as well as AMD CPUs. Key points:
7.2 and vLLM ROCm build 0.19.1 (built from source).| Runtime | Precision | Backend | Hardware | Notes |
|---|---|---|---|---|
| GPU | FP32 / FP16 / BF16 | vLLM | AMD RYZEN AI MAX+ 395 w/ Radeon 8060S | VLLM_ROCM_USE_SKINNY_GEMM=0 set to avoid bf16/fp16 GEMM segfaults on gfx1151 |
For setup instructions, evaluation scripts, and custom configuration options, see the deepseek_r1 on GitHub.
Model Type: Distilled reasoning language model (text generation)
Base Model: Qwen/Qwen2.5-7B (Qwen2.5-7B)
Model Stats:
7BHigher MATH-500 accuracy means the model produces mathematically equivalent answers to ground truth more often — 100% would be perfect, ~0% is chance-level. Strong distilled reasoning models typically score ~85–95% on this benchmark.
| Metric | Description |
|---|---|
| MATH-500 Accuracy | Primary metric — fraction of problems where the model's final boxed answer is symbolically equivalent to the reference solution. |
Full Dataset Evaluation (MATH-500) — filled from deepseek-ai/DeepSeek-R1-Distill-Qwen-7B; run make metrics to refresh:
| Device | Backend | Precision | Variant | Accuracy (%) |
|---|---|---|---|---|
| GPU | vLLM | FP32 | DeepSeek-R1-Distill-Qwen-7B | 90.00 |
| GPU | vLLM | FP16 | DeepSeek-R1-Distill-Qwen-7B | 90.00 |
| GPU | vLLM | BF16 | DeepSeek-R1-Distill-Qwen-7B | 90.00 |
Want to explore the full evaluation scripts, config options, and other AMD-optimized model examples?
📂 View the full project on GitHub
The GitHub repository includes: