Downloads · 30 days
9
4% of all-time downloads
bigdefence/Midm-2.0-Mini-Vision-Instruct
Midm-2.0-Mini-Vision-Instruct is a image-to-text model from bigdefence. Use it when you need a caption or text from an image. The card lists the license as apache-2.0.
- Midm-2.0-Mini-Vision-Instruct은 Midm-2.0-Mini-Vision-Instruct은 한국어 이미지 인식에 특화된 고성능, 경량 Vision-Language Model입니다. K-intelligence/Midm-2.0-Mini-Instruct 기반으로 구축되어 한국어 텍스트가 포함된 이미지 이해와 한국어 응답 생성에 최적화되었습니다. - End-to-End…
Downloads · 30 days
9
4% of all-time downloads
All-time downloads
212
Public
Parameters
2.6B
5.2 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors5.2 GB · 100%
From the Hugging Face model README

| 항목 | 세부사항 |
|---|---|
| 기반 모델 | K-intelligence/Midm-2.0-Mini-Instruct |
| 언어 | 한국어 (Korean) |
| 모델 크기 | ~2B 파라미터 |
| 작업 유형 | Image-to-Text 이미지 멀티모달 |
| 라이선스 | Apache 2.0 |
Midm-2.0-Mini-Vision-Instruct을 시작하려면 다음과 같이 레포지토리를 클론하고 환경을 설정하세요. 🛠️
레포지토리 클론:
git clone https://github.com/bigdefence/midm-vision
cd midm-vision
의존성 설치:
conda create -n midm-vision python=3.10 -y
conda activate midm-vision
pip install -e .
pip install flash-attn==2.5.2 --no-build-isolation
Huggingface CLI 사용:
pip install -U huggingface_hub
huggingface-cli download bigdefence/Midm-Vision --local-dir ./checkpoints
Snapshot Download 사용:
pip install -U huggingface_hub
from huggingface_hub import snapshot_download
snapshot_download(
repo_id="bigdefence/Midm-Vision",
local_dir="./checkpoints",
resume_download=True
)
Git 사용:
git lfs install
git clone https://huggingface.co/bigdefence/midm-vision
Midm-Vision으로 추론을 수행하려면 다음 단계를 따라 모델을 설정하고 로컬에서 실행하세요. 📡
모델 준비:
추론 실행:
python3 infer.py --model-path checkpoints --image-file test.jpg
이 모델은 Apache 2.0 라이선스 하에 배포됩니다. 상업적 사용이 가능하며, 자세한 내용은 LICENSE 파일을 참조하세요.