gavamedia/deltafin
Python
Run Kimi K3, a 2.8T-parameter Mixture-of-Experts LLM, on a single Apple Silicon Mac. Streams MXFP4 experts on demand over HTTP into a local disk cache — fused NEON kernels, Metal/MPS compute, exact reproducible decoding, and an OpenAI-compatible API server for local chat and coding agents.
★ +5 today 26 total stars
Star History
kimi kimi-k3 local-ai local-llm
Quality 100
🔥 41