Downloads ยท 30 days
553
100% of all-time downloads
aixk/BAAR2-3M
BAAR2-3M is a text generation model from aixk. Use it when you need the model to write or continue text. The card lists the license as other.
<p align="center" <a href="https://baar.uk/model/" <img src="https://cdn.jsdelivr.net/npm/minicon1@latest/logo/baar.png" alt="BAAR Logo" width="180" </a </p
Downloads ยท 30 days
553
100% of all-time downloads
All-time downloads
553
Public
Parameters
4.4M
193 MB on disk
Likes
2
Public
Click a slice to open those files.
.pt98.9 MB ยท 87%
From the Hugging Face model README
BAAR2-3M is an ultra-compact multilingual conversational model designed to enable real-time chat under minimal hardware conditions. With just 3 million (3M) parameters, it brings lightweight, interactive dialogue capabilities across 11 languages to extreme resource-constrained environments, embedded devices, and on-device platforms where standard LLMs cannot operate.
BAAR2-3M์ ์ต์ํ์ ํ๋์จ์ด ์ฌ๊ฑด๊ณผ ๊ทน๋จ์ ์ธ ๋ฆฌ์์ค ์ ์ฝ ํ๊ฒฝ์์๋ 11๊ฐ ์ธ์ด๋ก ์ค์๊ฐ ๋ํ(Chat)๊ฐ ๊ฐ๋ฅํ๋๋ก ์ค๊ณ๋ 300๋ง(3M) ํ๋ผ๋ฏธํฐ ๊ท๋ชจ์ ์ด๊ฒฝ๋ ๋ค๊ตญ์ด ์ธ์ด ๋ชจ๋ธ์ ๋๋ค. ์ผ๋ฐ์ ์ธ ๋ํ ๋ชจ๋ธ ๊ตฌ๋์ด ๋ถ๊ฐ๋ฅํ ์ด์ ์ฌ์ ์ฃ์ง ๋ฐ ์๋ฒ ๋๋ ๊ธฐ๊ธฐ์์๋ ์ฆ๊ฐ์ ์ธ ๋ํํ ์ถ๋ก ๊ณผ ์ด๊ฒฝ๋ ์ค์๊ฐ RAG ์์คํ ๊ตฌ์ถ์ ์ง์ํฉ๋๋ค.
Hardware: Single NVIDIA Tesla T4 | FP16 Inference
| Task / Category | Prompt | Generated Output | Latency |
|---|---|---|---|
| ํ๊ตญ์ด ์ผ๋ฐ QA | ๋ํ๋ฏผ๊ตญ์ ์๋๋ ์ด๋์ธ๊ฐ์? | ๋ํ๋ฏผ๊ตญ์ ์๋๋ ์ฌ๋ฌ ๊ฐ์ง ์ฃผ์ ์์๋ฅผ ๋งํ๊ณ | ~298.8ms |
| ์์ด ์ผ๋ฐ QA | What is artificial intelligence? | It is a complex region of the subject in the | ~283.0ms |
| ์ผ์์ด ๋ฒ์ญ ์์ฒญ | ์๋ ์ด ์์ด๋ก ๋ญ์ผ? | ์๋
์ ์์ด๋ก Hello ์
๋๋ค. | ~215.0ms |
| ์ผ์์ด ๋ฒ์ญ ์์ฒญ 2 | ์ฌ๊ณผ๊ฐ ์์ด๋ก ๋ญ์ผ? | ์ฌ๊ณผ๋ ์์ด๋ก apple ์
๋๋ค. | ~212.2ms |
| ์->ํ ๋ฒ์ญ ์์ฒญ | What is 'Seoul' in Korean? | A monument in the first time of the park in the U | ~275.6ms |
| ์ผ๋ณธ์ด ์ผ๋ฐ QA | ๆฅๆฌใฎ้ฆ้ฝใฏใฉใใงใใ๏ผ | ๆฅๆฌใฎ้ฆ้ฝใฏใ ๆฅๆฌใฎ้ฆ้ฝใงใใ | ~299.2ms |
| ์ค๊ตญ์ด ์ผ๋ฐ QA | ไบบๅทฅๆบ่ฝ็ไธป่ฆ็น็นๆฏไปไน๏ผ | ไบบๅทฅๆบ่ฝ็ไธป่ฆ็น็นๆฏไธ็งๅ
ฌๅธ | ~285.6ms |
| ์คํ์ธ์ด ์ผ๋ฐ QA | ยฟCuรกl es la capital de Espaรฑa? | Espaรฑa de Espaรฑa de Espaรฑa de Espaรฑa de | ~286.0ms |
| ๋ฌ์์์ด ์ผ๋ฐ QA | ะะฐะบะฐั ััะพะปะธัะฐ ะ ะพััะธะธ? | ะ ะพััะธะธ ะ ะพััะธะธ โ ััะพ ะฟัะตะดะปะฐะณะฐััะธะน ะฟัะตะดะป | ~270.6ms |
| ํ๋์ด ์ผ๋ฐ QA | เคญเคพเคฐเคค เคเฅ เคฐเคพเคเคงเคพเคจเฅ เคเฅเคฏเคพ เคนเฅ? | เคญเคพเคฐเคค เคเฅ เคฐเคพเคเคงเคพเคจเฅ เคเฅเคฏเคพ เคนเฅ เคเคฟเคธเคเคพ เคเคชเคฏเฅเค เคเคฟเคฏเคพ เคเคพ เคธเคเคคเคพ | ~271.1ms |
Non-Commercial Use Only (๋น์์
์ ์ฐ๊ตฌยท๊ฐ์ธ ์ด์ฉ ํ์ ):
This model is strictly provided for non-commercial, research, and educational purposes only. Any commercial use without a separate commercial agreement is strictly prohibited.
๋ณธ ๋ชจ๋ธ์ ์์ ์ฐ๊ตฌ, ๊ต์ก ๋ฐ ๊ฐ์ธ ๋น์๋ฆฌ ๋ชฉ์ ์ ํํด์๋ง ๋ฌด๋ฃ๋ก ์ ๊ณต๋ฉ๋๋ค. ์ฌ์ ํ๊ฐ ์๋ ์ผ์ฒด์ ์๋ฆฌ์ ์ด์ฉ์ ์๊ฒฉํ ๊ธ์ง๋ฉ๋๋ค.
Commercial Licensing & Inquiries for Revenue-Generating Entities (๋งค์ถ ๋ฐ์ ๊ธฐ์
๋ฐ ์์ฉํ ๋์
๋ฌธ์):
If your organization generates revenue, or if you plan to incorporate this model into commercial products, services, or paid APIs, you are required to purchase a commercial license. Please contact us directly via email for enterprise licensing and integration support.
๊ธฐ์
/์ฌ์
์ฒด์ ๋งค์ถ์ด ๋ฐ์ํ๊ณ ์๊ฑฐ๋, ๋ณธ ๋ชจ๋ธ์ ์์ฉ ์๋น์คยท์ ๋ฃ ์ ํยท์์ต ์ฐฝ์ถ ๋ชฉ์ ์ ์์คํ
์ ๋์
ํ๊ณ ์ ํ์๋ ๊ฒฝ์ฐ ๋ฐ๋์ ๋ณ๋์ ์์ฉ ๋ผ์ด์ ์ค๋ฅผ ์ทจ๋ํ์
์ผ ํฉ๋๋ค. ๋์
์กฐ๊ฑด ๋ฐ ๊ธฐ์
์ฉ ๋ผ์ด์ ์ค ๋ฐ๊ธ์ ์๋ ์ด๋ฉ์ผ๋ก ๋ฌธ์ํด ์ฃผ์๊ธฐ ๋ฐ๋๋๋ค.
This project demonstrates practical competency in designing custom micro-scale architectures, end-to-end multi-language pre-training, and ultra-low-latency deployment under extreme compute constraints.
I am actively seeking AI Engineering / Research opportunities, recruitment offers, and technical partnerships.
๊ทน๋จ์ ์ธ ์ ์ฌ์ ๋ฆฌ์์ค ํ๊ฒฝ์์๋ ์ฆ๊ฐ ์๋ํ๋ ๊ณ ํจ์จ ๋ค๊ตญ์ด ๋ชจ๋ธ ์ํคํ ์ฒ๋ฅผ ์ค๊ณํ๊ณ , ๋ฐ์ดํฐ ์ ์ ๋ถํฐ ํ์ตยท์๋น๊น์ง ์ ๊ณผ์ ์ ์ง์ ๊ตฌํํ๋ AI ์์ง๋์ด์ ๋๋ค. ๊ธฐ์ ๋์ , ์ฑ์ฉ ์ ์ ๋ฐ ํ์ ์ ๊ด์ฌ ์๋ ๊ธฐ์ ๊ณผ ํ์ ๋ฌธ์๋ฅผ ํ์ํฉ๋๋ค.