Popular repositories Loading
-
llama.cpp
llama.cpp PublicForked from ggml-org/llama.cpp
llama.cpp fork for running Qwen3.8-Flash-Next V3 (95.5 GiB) on a 64 GB Apple Silicon Mac by streaming MoE experts from SSD. Setup guide: docs/qwen38-flash-next-v3.md
-
splash-plus
splash-plus PublicEmpirical benchmarks and quality evaluations of Qwen3.8-27B with Multi-Token Prediction (MTP) across Splash, MTPLX, MLX, and llama.cpp on Apple Silicon
Python 8
-
splash
splash PublicForked from incoai/splash
A local inference engine for Apple silicon, built around the model.
Python 6
-
starling
starling PublicForked from starling/starling
Starling Message Queue - please contribute if you want commit access
Ruby
-
QuantSoftwareToolkit
QuantSoftwareToolkit PublicForked from QuantSoftware/QuantSoftwareToolkit
QuantSoftwareToolkit
Python
-
If the problem persists, check the GitHub status page or contact support.


