Downloads · 30 days
0
coreai-community/FastContext-1.0-4B-CoreAI
FastContext-1.0-4B-CoreAI is a text generation model from coreai-community. Use it when you need the model to write or continue text. The card lists the license as mit.
Core AI is Apple's on-device ML runtime in iOS 27 / macOS 27 and the successor to Core ML: PyTorch models are exported with Apple's coreai-torch (LLMs: coreai.llm.export) into .aimodel bundles that run on the GPU or t…
Downloads · 30 days
0
Access
Public
Updated Sep 5, 2026
Repo size
2.3 GB
Likes
0
Public
Click a slice to open those files.
.bin2.3 GB · 99%
From the Hugging Face model README
Core AI is Apple's on-device ML runtime in iOS 27 / macOS 27 and the successor to Core ML: PyTorch models are exported with Apple's coreai-torch (LLMs: coreai.llm.export) into .aimodel bundles that run on the GPU or the Neural Engine, e.g. Qwen3-8B 4-bit decodes at 94 tok/s on an M4 Max GPU, MLX 90 under the same protocol (apple-silicon-llm-bench, macOS 27 beta, 2026-06).
Mirror of
mlboydaisuke/FastContext-1.0-4B-CoreAI— the canonical repo (CoreAI Model Zoo). Updates land there first.
microsoft/FastContext-1.0-4B-SFT converted to Apple Core AI for on-device inference, served by the CoreAIChat app.
FastContext is a long-context repository-exploration agent (Qwen3-4B-Instruct backbone, supervised-fine-tuned on exploration traces: broad first-turn search, multi-turn evidence gathering, and precise file:line citation generation).
gpu/ — AOT-compiled GPU (MPSGraph) bundle, 4-bit linear-INT4, h18p
(iPhone 17 / iPhone 18 class). ~2.1 GB. Drop-in for CoreAIChat's
Documents/models/fastcontext_4b_gpu.Device class: this is an AOT h18p bundle (iPhone 17 / 18 class). On-device specialization of a 4B graph is not viable, so the bundle is precompiled for the h18p GPU — the same approach Core AI uses for the Gemma-4B bundle.
Install CoreAIChat, open the model picker, and select FastContext 4B.
MIT, inherited from the base model
microsoft/FastContext-1.0-4B-SFT.
See LICENSE.