Skip to main content
buildradar
Sign in

ARahim3/mlx-dspark

@ARahim3

Up to 4× faster LLM decoding on Apple Silicon, lossless. Native MLX port of DeepSeek's DSpark & z-lab's DFlash speculative decoding — Gemma-4, Qwen3.8, Muse-Glimmer, Nemotron, LFM2.5, Ornith-1.0, ternary Bonsai-27B.

Stars
647
Forks
53
Language
Python
License
MIT
Last push
1 day ago
Pythonmacosmetalclaude-codecodexllm-inferencelocal-llmpiapple-siliconmlxinference-enginespeculative-decodingdflash2dsparkqwen3-8

No related intel yet

This repo has not appeared in any of the sources the radar tracks. The collector runs on a schedule — check back once it covers this repo.