Downloads · 30 days
25
9% of all-time downloads
NexaAI/Granite-4-Micro-NPU
Granite-4-Micro-NPU is a text generation model from NexaAI. Use it when you need the model to write or continue text.
Run Granite-4.0-Micro optimized for Qualcomm NPUs with nexaSDK.
Downloads · 30 days
25
9% of all-time downloads
All-time downloads
286
Public
Repo size
3.3 GB
Likes
3
Public
Click a slice to open those files.
.nexa3.3 GB · 100%
From the Hugging Face model README
Run Granite-4.0-Micro optimized for Qualcomm NPUs with nexaSDK.
Install NexaSDK and create a free account at sdk.nexa.ai
Activate your device with your access token:
nexa config set license '<access_token>'
Run the model on Qualcomm NPU in one line:
nexa infer NexaAI/Granite-4-Micro-NPU
Granite-4.0-Micro is a 3B parameter instruction-tuned model in the Granite 4.0 family, developed by IBM.
It’s optimized for long-context reasoning (128K tokens), efficient inference, and enterprise-ready capabilities such as tool calling and retrieval-augmented generation. The model balances compact size with strong performance across general NLP tasks, making it suitable for both experimentation and production workloads.
Input: Natural language text prompts, chat conversations, or tool-augmented requests.
Output: Natural language responses—answers, explanations, summaries, structured JSON for function calls, or code snippets.
This model is released under the Creative Commons Attribution–NonCommercial 4.0 (CC BY-NC 4.0) license.
Non-commercial use, modification, and redistribution are permitted with attribution.
For commercial licensing, please contact [email protected].