Downloads · 30 days
0
simaai/LFM2-350M-a16w4
LFM2-350M-a16w4 is a text generation model from simaai. Use it when you need the model to write or continue text. It is set up for llima. The card lists the license as other.
This repository contains the LFM2-350M-a16w4 model, optimized and compiled for the SiMa.ai Modalix platform.
Downloads · 30 days
0
Access
Public
Updated Sep 4, 2026
Repo size
2.4 GB
Likes
0
Public
Click a slice to open those files.
.elf775 MB · 91%
From the Hugging Face model README
This repository contains the LFM2-350M-a16w4 model, optimized and compiled for the SiMa.ai Modalix platform.
The following measurements characterize text generation on Modalix.
Measured with MoLE using batch size 1, five samples per input length, and up to 128 generated tokens. Values are arithmetic means. TTFT includes language-model prefill and the first generated token; generation rate is measured after the first token.
| Input tokens | Mean TTFT (seconds) | Mean generation rate (tokens/second) |
|---|---|---|
| 128 | 0.02 | 197.53 |
| 256 | 0.03 | 196.33 |
| 512 | 0.06 | 191.87 |
| 1024 | 0.11 | 184.61 |
| 2048 | 0.24 | 158.50 |
| 3072 | 0.45 | 147.60 |
| 4096 | 0.70 | 137.77 |
| 5120 | 1.09 | 120.23 |
| 6144 | 1.55 | 120.36 |
| 7168 | 2.13 | 112.89 |
To run this model, you need:
Follow these steps to deploy the model to your Modalix device.
Note: This is a one-time setup. If the Neat Library is already installed on your Modalix device, you can skip this step and continue with model download.
Follow the SiMa.ai Neat getting started guide to install or update the Neat Library on your Modalix device.
The llima CLI is available on Modalix after the Neat runtime is installed. It manages precompiled GenAI models under /media/nvme/llima/models by default. Set LLIMA_MODELS_PATH to use a different model directory.
Download the compiled model assets from this repository directly to your device.
# Download the model to a local directory
llima pull LFM2-350M-a16w4
Alternatively, you can download the compiled model to a Host and copy it to the Modalix device:
hf download simaai/LFM2-350M-a16w4 --local-dir LFM2-350M-a16w4
scp -r LFM2-350M-a16w4 sima@<modalix-ip>:/media/nvme/llima/models/
Replace <modalix-ip> with the IP address of your Modalix device.
Expected Directory Structure:
/media/nvme/llima/
└── models/
└── LFM2-350M-a16w4/ # The compiled model
Run the model directly on Modalix:
llima run LFM2-350M-a16w4
For all runtime options, run:
llima run -h
The GenAI demo application is separate from LLiMa installation. Use the GenAI Multimodal Assistant page to install and run the demo app. Once installed, the demo app can use precompiled models such as this one.
To serve this model with OpenAI- or Ollama-compatible APIs and send requests to it, use the GenAI server workflow in Serve GenAI Models.
For direct LLM calls without setting up a server, see Run an LLM.
sima-cli not found: Ensure that sima-cli is installed on your Modalix device.llima not found: Install or update the Neat Library. See Getting Started./media/nvme/llima/models/ and not nested (e.g., /media/nvme/llima/models/LFM2-350M-a16w4/LFM2-350M-a16w4)./media/nvme directory.