Skip to main content
buildradar
Sign in
Topic · llama

llama

Tracked open-source repos tagged llama, sorted by stars.

207 repos
  • RAG-Driven-Generative-AI@Denis2054

    This repository provides programs to build Retrieval Augmented Generation (RAG) code for Generative AI with LlamaIndex, Deep Lake, and Pinecone leveraging the power of OpenAI and Hugging Face models for generation and evaluation.

    624+1Star change over the last 7 days
  • LLamaTuner@jianzhnie

    Easy and Efficient Finetuning LLMs. (Supported LLama, LLama2, LLama3, Qwen, Baichuan, GLM , Falcon) 大模型高效量化训练+部署.

    621+0Star change over the last 7 days
  • kani@zhudotexe

    kani (カニ) is a highly hackable microframework for tool-calling language models. (NLP-OSS @ EMNLP 2023)

    609+2Star change over the last 7 days
  • Are Copilots Local Yet? The frontier of local LLM Copilots for code completion, project generation, shell assistance, and more. Find tools shaping tomorrow's developer experience, today!

    600+0Star change over the last 7 days
  • yalm@andrewkchan

    Yet Another Language Model: LLM inference in C++/CUDA, no libraries except for I/O

    596-1Star change over the last 7 days
  • swift@ai-ng

    Fast voice assistant powered by Groq, Cartesia, and Vercel.

    592+0Star change over the last 7 days
  • qvac@tetherto

    Open-source local AI SDK - run AI on-device with no cloud, no API keys. Supports GGUF, RAG, image, music, and video generation, speech-to-text, P2P inference, and more. Cross-platform: Linux, macOS, Windows, Android, iOS.

    589+32Star change over the last 7 days
  • yzma@hybridgroup

    Go with your own intelligence - Write Go applications that directly integrate llama.cpp for local inference using hardware acceleration on Linux, macOS, Windows, & WebAssembly.

    589+14Star change over the last 7 days
  • Groma@FoundationVision

    [ECCV2024] Grounded Multimodal Large Language Model with Localized Visual Tokenization

    586+0Star change over the last 7 days
  • tensorflow-nlp-tutorial@ukairia777

    tensorflow를 사용하여 텍스트 전처리부터, Topic Models, BERT, GPT, LLM과 같은 최신 모델의 다운스트림 태스크들을 정리한 Deep Learning NLP 저장소입니다.

    581+0Star change over the last 7 days
  • LLaVA-Mini@ictnlp

    LLaVA-Mini is a unified large multimodal model (LMM) that can support the understanding of images, high-resolution images, and videos in an efficient manner.

    577+0Star change over the last 7 days
  • LLM-Hub@timmyy123

    Local LLM, image&video&music generator, vibecode like cursor with local models on your phone

    577+8Star change over the last 7 days
  • AnglE@SeanLee97

    Train and Infer Powerful Sentence Embeddings with AnglE | 🔥 SOTA on STS and MTEB Leaderboard

    573+0Star change over the last 7 days
  • Use Code Llama with Visual Studio Code and the Continue extension. A local LLM alternative to GitHub Copilot.

    569+0Star change over the last 7 days
  • MARTI@TsinghuaC3I

    [ICLR 2026] A Framework for LLM-based Multi-Agent Reinforced Training and Inference

    554+2Star change over the last 7 days
  • snowChat@kaarthik108

    Chat snowflake - Text to SQL

    553+0Star change over the last 7 days
  • aikit@kaito-project

    🏗️ Fine-tune, build, and deploy open-source LLMs easily!

    539+1Star change over the last 7 days
  • LESS@princeton-nlp

    [ICML 2024] LESS: Selecting Influential Data for Targeted Instruction Tuning

    532-1Star change over the last 7 days
  • llama-assistant@nrl-ai

    AI-powered assistant to help you with your daily tasks, powered by Llama 3, DeepSeek R1, and many more models on HuggingFace.

    530+0Star change over the last 7 days
  • aicommit2@tak-bro

    A Reactive CLI that generates commit messages for Git and Jujutsu with Ollama, ChatGPT, Gemini, Claude, Mistral and other AI

    527-1Star change over the last 7 days
  • ReMind@DonTizi

    Your Local Artificial Memory on your Device.

    522+0Star change over the last 7 days
  • sarathi-serve@microsoft

    A low-latency & high-throughput serving engine for LLMs

    520+0Star change over the last 7 days
  • LLaMA-Pro@TencentARC

    [ACL 2024] Progressive LLaMA with Block Expansion.

    514+1Star change over the last 7 days
  • Open source, local, and self-hosted highly optimized language inference server supporting ASR/STT, TTS, and LLM across WebRTC, REST, and WS

    512+0Star change over the last 7 days
  • restai@apocas

    RESTai is an AIaaS (AI as a Service) open-source platform. Supports many public and local LLM suported by Ollama/vLLM/etc. Precise embeddings usage, tuning, analytics etc. Built-in image/audio generation with dynamic loading generators. Live chat deployment. Built-in block based graphical language. Prompt versioning and much more...

    512+0Star change over the last 7 days
  • 🚀🚀🚀 This repository lists some awesome public CUDA, cuda-python, cuBLAS, cuDNN, CUTLASS, TensorRT, TensorRT-LLM, Triton, TVM, MLIR, PTX and High Performance Computing (HPC) projects.

    511-2Star change over the last 7 days
  • ome@ome-projects

    Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton

    506Star change over the last 7 days
← Back to topics