edge-ai
Tracked open-source repos tagged edge-ai, sorted by stars.
Related topics
Topics that frequently appear alongside edge-ai on the same repo.
Recent risers
Repos created in the last 90 days, tagged edge-ai.
- #1
An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, and more, with highly optimized performance. No Python dependency.
★ 2,214 - #2
Run MoE models bigger than your RAM. Frontier-size MoE on a 12 GB phone, CPU only, lossless, on stock llama.cpp
★ 543 - #3
Build AI agents that run 100% on-device. Sub-100ms latency on Qualcomm NPU. Zero cloud dependency.
★ 486
- #1
High-Performance server for NATS.io, the cloud and edge native messaging system.
★ 20,653+23Star change over the last 7 days - #2
PyTorch implementation of YOLOv3, YOLOv3-SPP, and YOLOv3-tiny for real-time object detection with training, validation, inference, and multi-format export.
★ 10,602+3Star change over the last 7 days - #3
LiteRT-LM is Google's production-ready, high-performance, open-source inference framework for deploying Large Language Models on edge devices.
★ 6,358+39Star change over the last 7 days - #4
Quantization, kernels, runtime and inference engine for mobiles, wearables, smart home and robots.
★ 5,975+20Star change over the last 7 days - #5★ 5,246+3Star change over the last 7 days
- #6
FEDML - The unified and scalable ML library for large-scale distributed training, model serving, and federated learning. FEDML Launch, a cross-cloud scheduler, further enables running any AI jobs on any GPU cloud or on-premise cluster. Built on this library, TensorOpera AI (https://TensorOpera.ai) is your generative AI platform at scale.
★ 4,062+1Star change over the last 7 days - #7
The Swiss Army Knife of Offline AI. Chat, see, speak, and generate images on your phone or Mac — GGUF LLMs, vision, Whisper speech-to-text, Stable Diffusion, tool calling, and local-network servers. Runs on your CPU, GPU, or NPU. No account, no API key, zero data leaves your device.
★ 3,038+17Star change over the last 7 days - #8
An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, and more, with highly optimized performance. No Python dependency.
★ 2,214+115Star change over the last 7 days - #9★ 1,712+0Star change over the last 7 days
- #10
Microsoft AI for Good Lab — Biodiversity research hub. Open-source AI models, edge devices, and tools for biodiversity monitoring and conservation. Your source for MegaDetector, SPARROW, PytorchWildlife, Bioacoustics, and more.
★ 1,068+1Star change over the last 7 days - #11
DefraDB is a Peer-to-Peer Edge-First Database. It's the core data storage system for the Source Ecosystem.
★ 895+4Star change over the last 7 days - #12
Android 17 local LLM prototype with Jetpack Compose and ONNX Runtime for offline AI inference experiments.
★ 853+0Star change over the last 7 days - #13★ 843+1Star change over the last 7 days
- #14
Real-time voice assistant — WebRTC streaming, faster-whisper ASR, local LLM, Vui Nano (300M) TTS. OpenAI Realtime API compatible. Voice cloning, barge-in, ~9× realtime on a 4090. Apache 2.0.
★ 756+9Star change over the last 7 days - #15
The Hailo Model Zoo includes pre-trained models and a full building and evaluation environment
★ 703+2Star change over the last 7 days - #16
speech to text benchmark framework
★ 697+0Star change over the last 7 days - #17
EfficientSAM3 compresses SAM3 into lightweight, edge-friendly models via progressive knowledge distillation for fast promptable concept segmentation and tracking.
★ 666+8Star change over the last 7 days - #18
Production-grade C++ edge AI engine for video analytics and on-device VLM across Sophon, Rockchip RKNN, and x86, with visual orchestration, real-time OSD, events, and reproducible benchmarks.
★ 656+45Star change over the last 7 days - #19
AI-in-a-Box leverages the expertise of Microsoft across the globe to develop and provide AI and ML solutions to the technical community. Our intent is to present a curated collection of solution accelerators that can help engineers establish their AI/ML environments and solutions rapidly and with minimal friction.
★ 596+0Star change over the last 7 days - #20
Run MoE models bigger than your RAM. Frontier-size MoE on a 12 GB phone, CPU only, lossless, on stock llama.cpp
★ 543—Star change over the last 7 days - #21
On-Device Training Under 256KB Memory [NeurIPS'22]
★ 524+0Star change over the last 7 days - #22
A curated list of awesome edge computing, including Frameworks, Simulators, Tools, etc.
★ 520-1Star change over the last 7 days