Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
C++mozilla-ai/llamafile

llamafile

Distribute and run LLMs with a single file.

87.4/100
25.9KForks: 1.6K
View on GitHubHomepage →
Loading report...

Similar Projects

RCLI

84

Talk to your Mac, query your docs, no cloud required. On-device voice AI + RAG

C++1.5K

lucebox

73

LLM speculative inference server for heterogeneous hardware & consumer GPUs

C++2.8K

BigMoeOnEdge

77

Run MoE models bigger than your RAM. Frontier-size MoE on a 12 GB phone, CPU only, lossless, on stock llama.cpp

C++548

whisper.cpp

87

Port of OpenAI's Whisper model in C/C++

C++53.5K
Back to List