Downloads · 30 days
0
shaafsalman/shaafsalman
shaafsalman is a machine learning model from shaafsalman. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
ML/AI engineer working on medical vision-language models — chest X-ray report generation, clinical reasoning fine-tunes, and LLM production observability tooling.
Downloads · 30 days
0
Access
Public
Updated Jul 21, 2026
Repo size
—
Likes
0
Public
Click a slice to open those files.
.md7.5 KB · 83%
From the Hugging Face model README
ML/AI engineer working on medical vision-language models — chest X-ray report generation, clinical reasoning fine-tunes, and LLM production observability tooling.
Affiliated with NeuroCare.AI and SaveLife.AI. Writing at medium.com/@ishaafsalman.
Fine-tunes of Qwen3.5-9B for chest radiograph report generation, plus supporting object-detection and segmentation baselines.
| Model | Description |
|---|---|
| Qwen3.5-9B-CXR-VA | Vision-language CXR report generation, variant A |
| Qwen3.5-9B-CXR-DT | Vision-language CXR report generation, DT checkpoint |
| Qwen3.5-9B-CXR-LoRA-V2 | LoRA adapter, v2 |
| Qwen3.5-9B-CXR-GGUF | Quantized GGUF builds for local inference |
| RF-DETR-CXR-VinBigData | RF-DETR abnormality detection, VinBigData |
| RF-DETR-CXR-Detection-V3 (private) | RF-DETR abnormality detection, v3 training run |
| UNet-CXR-ChestXDet | UNet + EfficientNet-B4 CXR segmentation |
| Model | Description |
|---|---|
| Qwen2.5-32B-Med-Merge | 32B medical merge of Qwen2.5-32B-Instruct |
| Dataset | Description |
|---|---|
| medagentbench-tasks | Stanford MedAgentBench FHIR agentic task set |
| visionops-eval-data (private) | VLM CXR evaluation results across models |
| mimic-cxr-sft-70k (private) | MIMIC-CXR SFT set, 70K reports |
| mimic-cxr-100k-sft (private) | MIMIC-CXR SFT set, 100K variant |
| cxr-openi-nih-50k (private) | OpenI + NIH ChestX-ray14 combined, 50K |
| cxr-alpaca-v1 (private) | Alpaca-format CXR SFT dataset, 100K samples |
| qwen3.5-cxr-cot (private) | Chain-of-thought CXR reasoning traces |
| qwen3.5-cxr-cot-v3 (private) | CXR reasoning traces, v3 |
| cxr-evals (private) | Base vs. fine-tuned CXR eval comparison, 1K samples |
| cxr-golden-v1 (private) | Curated 100K-row "golden" CXR report reference set |
| cxr-vlm-case-study (private) | 70-case VLM comparison case study |
| cxr-case-study-50 (private) | 50-case multi-strategy prompting case study |
Several of these datasets were generated with Mimesis, a local vLLM-based synthetic training-data pipeline — see the writeup: Generating Synthetic Training Data at Scale with vLLM.
| Dataset | Description |
|---|---|
| reasoning-traces-37k | 37.5K math/science chain-of-thought traces |
| medical-reasoning-traces-11k | 11K medical reasoning traces |
| instruction-sft-11k | Alpaca-format instruction SFT set |
| prompt-seeds | Seed prompts for data generation pipelines |
| synthetic-code-gen-pipeline | Synthetic code generation pipeline outputs |
| mimesis-opus-4.6 | Mimesis output — Claude Opus 4.6 backend |
| mimesis-deepseek | Mimesis output — DeepSeek backend |
| mimesis-togetherapi | Mimesis output — Together AI backend |
| mimesis-octomed | Mimesis output — Octo backend, medical seeds |
| mimesis-medical-octo-5k | Mimesis output — Octo backend, 5K medical rows |
Production LLM observability datasets from internal services — audit logs, benchmark samples, and regression fixtures. Kept private due to internal service and clinical-adjacent content.
llm-audit-logs · llm-audit-logs-by-service · llm-bench-prompts · medcoding-prompts · medcoding-v3-test-cases · soap-reasoning-leaks
Selected posts from medium.com/@ishaafsalman:
Profile last reorganized 2026-07-21. No Spaces or Collections currently exist on this account.