Downloads · 30 days
81
25% of all-time downloads
LSXPrime/ProseFlow-v1-360M-Instruct-GGUF
ProseFlow-v1-360M-Instruct-GGUF is a text generation model from LSXPrime. Use it when you need the model to write or continue text. It is set up for gguf. The card lists the license as apache-2.0.
ProseFlow-v1-360M-Instruct is a lightweight, experimental instruction-tuned model created for the ProseFlow desktop application. This model is a fine-tune of HuggingFace's SmolLM-360M-Instruct and was created to explo…
Downloads · 30 days
81
25% of all-time downloads
All-time downloads
318
Public
Repo size
747 MB
Likes
0
Public
Click a slice to open those files.
.gguf747 MB · 100%
From the Hugging Face model README
ProseFlow-v1-360M-Instruct is a lightweight, experimental instruction-tuned model created for the ProseFlow desktop application. This model is a fine-tune of HuggingFace's SmolLM-360M-Instruct and was created to explore the capabilities of smaller language models on a diverse set of text-processing tasks.
The model was fine-tuned on the ProseFlow-Actions-v1 dataset.
Note: This model is provided for research and experimental purposes and low-resource devices. For the best user
experience in the ProseFlow application, the larger and more capable
ProseFlow-v1-1.5B-Instruct model is strongly recommended.
ProseFlow is a universal AI text processor that allows users to create and execute custom AI "Actions" on text in any application. This model was an experiment to see if a ~360M parameter model could reliably perform the wide range of tasks defined in the training dataset.
Evaluations show that while this model is extremely fast and has very low resource requirements, its capabilities are limited.
This repository provides multiple versions of the model, allowing users to choose the best balance of performance and resource usage for their specific hardware. All quantized versions are provided in the GGUF format for broad compatibility.
| File Name (Quantization) | VRAM Usage (Approx.) | Performance | Recommended Use Case |
|---|---|---|---|
Q8_0 | ~1 GB | Best Overall. Nearly identical to FP16. | The recommended default for most users. |
Q4_K_M | ~900 MB | Low Quality. Noticeable degradation in nuance. | For maximum speed on low-power devices. |
Note on Quantization: To maintain the highest possible quality, the token embeddings and the final output layer were kept at F16 precision. Additionally, an importance matrix was used for calibration during the quantization process. This is why the quantized files are larger than what might typically be expected, as this method significantly improves their performance and coherence.
To Paragraph), the model hallucinates content completely unrelated to the
input.This model is intended for experimental use and for users on extremely resource-constrained systems who are willing to accept a significant trade-off in performance and reliability. It may be suitable for a very limited subset of simple, repetitive text-formatting tasks.
It is designed to be used within the ProseFlow desktop application, but it is not the recommended model for general use.
ProseFlow-v1-360M-Instruct from the "Available for
Download" list. We recommend starting with Q8_0.This model is licensed under the Apache License, Version 2.0.