Skip to content

ctogaurav

GLM_OCR

ctogaurav/GLM_OCR

GLM_OCR is a image-to-text model from ctogaurav. Use it when you need a caption or text from an image. It is set up for peft. The card lists the license as mit.

Fine-tunes zai-org/GLM-OCR (0.9B vision-language model) with LoRA to transcribe handwritten university-level math answer sheets into complete, pdflatex-compilable LaTeX documents β€” ignoring printed headers, student id…

Downloads Β· 30 days

0

Access

Public

Updated Sep 15, 2026

Repo size

282 MB

Likes

2

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors282 MB Β· 93%

At a glance

Task
Image-to-Text
Library
peft
License
mit
Access
Public
Created
Aug 3, 2026
Updated
Sep 15, 2026
SHA
18620480

Base models

Task
Image-to-Text
Library
peft
License
mit
Languages
en
Created
Aug 3, 2026
Updated
Sep 15, 2026