# llamafile

> Run an LLM from a single file

Packs a model and runtime into one executable that runs on most systems.

- Source: https://github.com/mozilla-ai/llamafile
- Homepage: https://docs.mozilla.ai/llamafile
- License: NOASSERTION
- Language: C++
- Stars: 26159
- Forks: 1634
- Contributors: 85
- Last commit: 2026-09-30
- Latest release: 0.10.6 (2026-09-15)
- Purpose: AI & LLM tooling, Local inference
- Runs on: CLI
- For: Personal

## Worth score: 71/100

How much DigGitHub recommends it, from activity, adoption, docs, license and security signals.

- Popularity: 22/25
- Momentum: 10/20
- Maintenance: 25/25
- Community: 11/15
- Readiness: 3/15

## Alternatives

- [Ollama](https://diggithub.com/ollama/ollama.md): Run open-weight LLMs locally with one command
- [llama.cpp](https://diggithub.com/ggml-org/llama.cpp.md): LLM inference in C/C++ on CPUs and GPUs
- [whisper.cpp](https://diggithub.com/ggml-org/whisper.cpp.md): Whisper speech recognition in C/C++
- [Colibri](https://diggithub.com/JustVugg/colibri.md): Run large MoE models on everyday hardware
- [llmfit](https://diggithub.com/AlexsJones/llmfit.md): Find which LLMs run on your hardware
- [exo](https://diggithub.com/exo-explore/exo.md): Run frontier AI models across your own devices
- [ds4](https://diggithub.com/antirez/ds4.md): Local DeepSeek 4 inference engine
- [LocalAI](https://diggithub.com/mudler/LocalAI.md): Drop-in OpenAI API replacement that runs locally

---

Source page: https://diggithub.com/mozilla-ai/llamafile
Updated: 2026-10-03
