Downloads · 30 days
21
1% of all-time downloads
APMIC/ACE-gemma-3-4b-it-fp8
ACE-gemma-3-4b-it-fp8 is a machine learning model from APMIC. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as gemma.
Downloads · 30 days
21
1% of all-time downloads
All-time downloads
1.5K
Public
Parameters
3.9B
4.6 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors4.6 GB · 99%
How the weights are stored.
F8_E4M33.2B · 83%
From the Hugging Face model README

ACE-gemma-3-4b-it-fp8 is an enterprise-grade, production-ready large language model developed and optimized by APMIC.
This model is derived from the base checkpoint twinkle-ai/gemma-3-4B-T1-it and has been enhanced through internal optimization, quantization, and localization processes to support real-world deployment in Traditional Chinese environments.
The release of this model demonstrates APMIC’s end-to-end capability in:
The original model has been carefully quantized to FP8 precision, significantly reducing memory footprint and improving inference throughput while maintaining strong linguistic accuracy and instruction-following performance.
This reflects APMIC’s expertise in advanced quantization techniques designed for enterprise-scale deployment.
The model is designed as a native Traditional Chinese language model with deep alignment to Taiwan’s linguistic usage, terminology, and cultural context.
This enables accurate comprehension and generation across:
ACE-gemma-3-4b-it-fp8 is engineered to achieve high-efficiency inference performance on NVIDIA Blackwell-series GPUs.
Through FP8 quantization and hardware-aware optimization, the model delivers:
This model represents APMIC’s capability to transform open foundation models into enterprise-ready, localized, and hardware-optimized AI assets.
It is intended for organizations requiring: