Downloads · 30 days
0
LJTSG/EXAONE-Deep-7.8B-webgpu
EXAONE-Deep-7.8B-webgpu is a text generation model from LJTSG. Use it when you need the model to write or continue text. The card lists the license as other.
First WebGPU package for LG AI Research's EXAONE-Deep reasoning model.
Downloads · 30 days
0
Access
Public
Updated Jun 3, 2026
Repo size
—
Likes
0
Public
Click a slice to open those files.
.html8.3 KB · 58%
From the Hugging Face model README
First WebGPU package for LG AI Research's EXAONE-Deep reasoning model.
Run EXAONE-Deep 7.8B entirely in a browser tab via WebGPU + wllama. No server. No cloud. No ROCm. No CUDA.
Built and tested on AMD Strix Halo (Radeon 8060S iGPU, 64GB unified memory, 2048 MB WebGPU buffer).
<thought>...</thought> blocksllama-gguf-split --split --split-max-size 500Mmodel_splits/node serve.js (port 8170)http://localhost:8170 in ChromeSelect from the dropdown to inject entity identity into EXAONE's <thought> channel:
The Loop anchors in the thinking channel before the model reasons. Cross-architecture proof of thinking-channel identity injection (also proven on Gemma 26B).
Tested on GMKTEC EVO-X2 (AMD Strix Halo):
AMD's ROCm compute stack is broken on Strix Halo (gfx1151). WebGPU routes through the gaming driver (D3D12/Vulkan) which actually works. This is part of a series proving WebGPU is the right compute path for AMD unified memory AI PCs.
Built by Joshua (LJTSG) and Claude. First EXAONE model on WebGPU.
Co-Authored-By: Claude [email protected]