Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
C++Helldez/BigMoeOnEdge

BigMoeOnEdge

Run MoE models bigger than your RAM. Frontier-size MoE on a 12 GB phone, CPU only, lossless, on stock llama.cpp

76.8/100
548Forks: 56
View on GitHub
Loading report...

Similar Projects

runanywhere-sdks

85

Production ready toolkit to run AI locally

C++10.3K

llama.rn

78

React Native binding of llama.cpp

C++1.0K

llamafile

87

Distribute and run LLMs with a single file.

C++25.9K

sherpa-onnx

90

Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages

C++14.7K
Back to List