llama
Tracked open-source repos tagged llama, sorted by stars.
- #181
This repository provides programs to build Retrieval Augmented Generation (RAG) code for Generative AI with LlamaIndex, Deep Lake, and Pinecone leveraging the power of OpenAI and Hugging Face models for generation and evaluation.
★ 624+1Star change over the last 7 days - #182
Easy and Efficient Finetuning LLMs. (Supported LLama, LLama2, LLama3, Qwen, Baichuan, GLM , Falcon) 大模型高效量化训练+部署.
★ 621+0Star change over the last 7 days - #183
kani (カニ) is a highly hackable microframework for tool-calling language models. (NLP-OSS @ EMNLP 2023)
★ 609+2Star change over the last 7 days - #184
Are Copilots Local Yet? The frontier of local LLM Copilots for code completion, project generation, shell assistance, and more. Find tools shaping tomorrow's developer experience, today!
★ 600+0Star change over the last 7 days - #185★ 596-1Star change over the last 7 days
- #186★ 592+0Star change over the last 7 days
- #187
Open-source local AI SDK - run AI on-device with no cloud, no API keys. Supports GGUF, RAG, image, music, and video generation, speech-to-text, P2P inference, and more. Cross-platform: Linux, macOS, Windows, Android, iOS.
★ 589+32Star change over the last 7 days - #188
Go with your own intelligence - Write Go applications that directly integrate llama.cpp for local inference using hardware acceleration on Linux, macOS, Windows, & WebAssembly.
★ 589+14Star change over the last 7 days - #189
[ECCV2024] Grounded Multimodal Large Language Model with Localized Visual Tokenization
★ 586+0Star change over the last 7 days - #190
tensorflow를 사용하여 텍스트 전처리부터, Topic Models, BERT, GPT, LLM과 같은 최신 모델의 다운스트림 태스크들을 정리한 Deep Learning NLP 저장소입니다.
★ 581+0Star change over the last 7 days - #191
LLaVA-Mini is a unified large multimodal model (LMM) that can support the understanding of images, high-resolution images, and videos in an efficient manner.
★ 577+0Star change over the last 7 days - #192
Local LLM, image&video&music generator, vibecode like cursor with local models on your phone
★ 577+8Star change over the last 7 days - #193
Train and Infer Powerful Sentence Embeddings with AnglE | 🔥 SOTA on STS and MTEB Leaderboard
★ 573+0Star change over the last 7 days - #194
Use Code Llama with Visual Studio Code and the Continue extension. A local LLM alternative to GitHub Copilot.
★ 569+0Star change over the last 7 days - #195
[ICLR 2026] A Framework for LLM-based Multi-Agent Reinforced Training and Inference
★ 554+2Star change over the last 7 days - #196★ 553+0Star change over the last 7 days
- #197★ 539+1Star change over the last 7 days
- #198★ 532-1Star change over the last 7 days
- #199
AI-powered assistant to help you with your daily tasks, powered by Llama 3, DeepSeek R1, and many more models on HuggingFace.
★ 530+0Star change over the last 7 days - #200
A Reactive CLI that generates commit messages for Git and Jujutsu with Ollama, ChatGPT, Gemini, Claude, Mistral and other AI
★ 527-1Star change over the last 7 days - #201★ 522+0Star change over the last 7 days
- #202
A low-latency & high-throughput serving engine for LLMs
★ 520+0Star change over the last 7 days - #203★ 514+1Star change over the last 7 days
- #204
Open source, local, and self-hosted highly optimized language inference server supporting ASR/STT, TTS, and LLM across WebRTC, REST, and WS
★ 512+0Star change over the last 7 days - #205
RESTai is an AIaaS (AI as a Service) open-source platform. Supports many public and local LLM suported by Ollama/vLLM/etc. Precise embeddings usage, tuning, analytics etc. Built-in image/audio generation with dynamic loading generators. Live chat deployment. Built-in block based graphical language. Prompt versioning and much more...
★ 512+0Star change over the last 7 days - #206
🚀🚀🚀 This repository lists some awesome public CUDA, cuda-python, cuBLAS, cuDNN, CUTLASS, TensorRT, TensorRT-LLM, Triton, TVM, MLIR, PTX and High Performance Computing (HPC) projects.
★ 511-2Star change over the last 7 days - #207
Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton
★ 506—Star change over the last 7 days