on-device-ai
Tracked open-source repos tagged on-device-ai, sorted by stars.
Related topics
Topics that frequently appear alongside on-device-ai on the same repo.
Recent risers
Repos created in the last 90 days, tagged on-device-ai.
- #1
Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
★ 6,674 - #2
Swiftlet is a Swift and Metal runtime that runs large Qwen Mixture-of-Experts models locally on Apple devices by streaming expert weights from storage, enabling 35B and 80B models to run with low RAM, including on iPhone.
★ 628 - #3
Run MoE models bigger than your RAM. Frontier-size MoE on a 12 GB phone, CPU only, lossless, on stock llama.cpp
★ 548
- #1
Production ready toolkit to run AI locally
★ 10,281+1Star change over the last 7 days - #2
14MB foundation model for tiny devices; phones, wearables, smart home, and robots.
★ 10,106+445Star change over the last 7 days - #3
Run frontier LLMs and VLMs locally on Qualcomm devices across NPU, GPU, and CPU with a few lines of code
★ 8,347+9Star change over the last 7 days - #4
Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
★ 6,674+187Star change over the last 7 days - #5
LiteRT-LM is Google's production-ready, high-performance, open-source inference framework for deploying Large Language Models on edge devices.
★ 6,358+39Star change over the last 7 days - #6
Quantization, kernels, runtime and inference engine for mobiles, wearables, smart home and robots.
★ 5,975+20Star change over the last 7 days - #7
AI You Control: Choose your models. Own your data. Eliminate vendor lock-in.
★ 4,761-2Star change over the last 7 days - #8
An AI-powered file management tool that ensures privacy by organizing local texts, images. Using Llama3.2 3B and Llava v1.6 models with the Nexa SDK, it intuitively scans, restructures, and organizes files for quick, seamless access and easy retrieval.
★ 3,334+0Star change over the last 7 days - #9
Run Claude Code 100% on-device with local AI on Apple Silicon. MLX-native Anthropic-API server. 6 fighters incl. Muse-Glimmer 30B (now multimodal — reads images, abliterated), Gemma 4 31B, Qwen 3.5 122B (65 tok/s), DeepSeek V4 Flash (1M ctx). Private, offline, airgap-ready. Built for NDA / legal / healthcare workflows.
★ 3,255+12Star change over the last 7 days - #10
Mano-P: Open-source GUI-VLA agent for edge devices. #1 on OSWorld (specialized, 58.2%). Runs locally on Apple M4 Mac mini/MacBook — no data leaves your device.Mano-P 是一个开源 GUI-VLA 项目,支持在 Mac mini/MacBook 上或通过算力棒本地运行推理,实现纯视觉驱动的跨平台 GUI 自动化操作。数据完全本地处理,支持复杂多步骤任务规划与执行。
★ 2,615+15Star change over the last 7 days - #11★ 2,048+7Star change over the last 7 days
- #12
Core ML model zoo for iOS/macOS — PyTorch models converted to ready-to-use .mlpackage, each with a conversion script and SwiftUI sample app. Sibling repos cover Apple's Core AI framework (iOS/macOS 27) and on-device LLMs.
★ 1,860+4Star change over the last 7 days - #13
Declarative way to run AI models in React Native on device, powered by ExecuTorch.
★ 1,707+5Star change over the last 7 days - #14★ 1,542+0Star change over the last 7 days
- #15
[ICLR 2019] ProxylessNAS: Direct Neural Architecture Search on Target Task and Hardware
★ 1,447+0Star change over the last 7 days - #16
Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unsloth-compatible API.
★ 1,398+8Star change over the last 7 days - #17★ 1,392+3Star change over the last 7 days
- #18
PhoneClaw turns phones into local AI agent runtimes with on-device models, native mobile Skills, LiveLand, and optional Mac Gateway inference.
★ 1,231+6Star change over the last 7 days - #19
A practical lab for building, testing, and evaluating apps with Apple's Foundation Models framework.
★ 1,178+1Star change over the last 7 days - #20
Official PyTorch implementation of "EdgeSAM: Prompt-In-the-Loop Distillation for On-Device Deployment of SAM"
★ 1,173-1Star change over the last 7 days - #21
NativeMind: Your fully private, open-source, on-device AI assistant
★ 1,130+1Star change over the last 7 days - #22
NobodyWho is an inference engine that lets you run LLMs locally and efficiently on any device.
★ 1,094+13Star change over the last 7 days - #23
Muesli: agent-native local meeting transcription + dictation for macOS (Granola + WisprFlow alternative)
★ 1,091+43Star change over the last 7 days - #24
PokeClaw (PocketClaw) — first on-device AI that controls your Android phone. Gemma 4, no cloud, no API key. Poke is short for Pocket.
★ 1,040+14Star change over the last 7 days - #25
[CVPR 2025] Official PyTorch implementation of "EdgeTAM: On-Device Track Anything Model"
★ 975+2Star change over the last 7 days - #26
TinyChatEngine: On-Device LLM Inference Library
★ 961+0Star change over the last 7 days - #27
Sudoless Apple Silicon system monitor (native SwiftUI GUI) with ANE / Media Engine / memory-bandwidth tracking
★ 928+28Star change over the last 7 days - #28
Local-first, open-source AI assistant for your data. Unify tasks, notes, docs, photos, and bookmarks. Private, self-hosted, and extensible via APIs.
★ 918+6Star change over the last 7 days - #29
Optimized Whisper models for streaming and on-device use
★ 897+0Star change over the last 7 days - #30
Shared Single-file memory layer for all your agents, sub mili-second RAG over text, photo and video on Apple Silicon.. No Server. No API. One File. Pure Swift
★ 789-2Star change over the last 7 days