Downloads · 30 days
0
runpodai09/old_poker
old_poker is a machine learning model from runpodai09. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for peft.
This private repository preserves the locally available checkpoints and experiment outputs from earlier QwenPoker work. The contents are historical evidence. They do not establish playing strength against independent…
Downloads · 30 days
0
Access
Public
Updated Sep 23, 2026
Repo size
1.5 GB
Likes
0
Public
Click a slice to open those files.
.safetensors793 MB · 45%
From the Hugging Face model README
This private repository preserves the locally available checkpoints and experiment outputs from earlier QwenPoker work. The contents are historical evidence. They do not establish playing strength against independent opponents.
| Directory | Observed base model | Contents |
|---|---|---|
checkpoints/phase1_sft/ | Qwen/Qwen2.5-3B-Instruct | Final LoRA adapter and final resumable checkpoint at step 10795. |
checkpoints/phase1_sft_1.5b_backup/ | Qwen/Qwen2.5-1.5B-Instruct | Final LoRA adapter and final resumable checkpoint at step 10795. |
checkpoints/phase2_rl/ | Local path /workspace/Qwen2.5-3B-Instruct in adapter metadata | Available evaluated steps 40960, 61440, 71680, and 122880, plus two adapter combinations. |
The RL adapter metadata records a local base path. Resolve the exact compatible base model and tokenizer before loading it outside the original environment. A LoRA adapter is not a standalone base model.
Earlier intermediate SFT and RL checkpoints also exist locally but are not included in this initial transfer. The table above is the exact checkpoint selection published here.
outputs/ contains available training logs, configuration snapshots, evaluation tables, and reports. Four historical files with untranslated text or presentation content were excluded from the English release.artifacts/platform_evidence/ contains Phase 0 research, engine validation, and implementation evidence produced during later platform work. These are not completed model training weights or independently verified real hand histories.The original raw OCR corpus and source videos are not included. Some historical records have incomplete experiment provenance. Do not treat a checkpoint directory name or a self-play metric as an independently validated model ranking.
The current platform code is maintained separately in runpodai09/poker_platform. New Qwen3.5-4B model weights are not part of this legacy repository because no completed local training checkpoint was identified for that model.