Skip to main content
buildradar
Sign in
Topic · gpu

gpu

Tracked open-source repos tagged gpu, sorted by stars.

279 repos
  • warp@NVIDIA

    A Python framework for GPU-accelerated simulation, robotics, and machine learning.

    7,070+17Star change over the last 7 days
  • Halide@halide

    a language for fast, portable data-parallel computation

    6,595+2Star change over the last 7 days
  • flashinfer@flashinfer-ai

    FlashInfer: Kernel Library for LLM Serving

    6,321+38Star change over the last 7 days
  • x11docker@mviereck

    Run GUI applications and desktops in docker and podman containers. Focus on security.

    6,306+3Star change over the last 7 days
  • harfbuzz@harfbuzz

    HarfBuzz text shaping engine

    6,051+11Star change over the last 7 days
  • DALI@NVIDIA

    A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference applications.

    5,740+2Star change over the last 7 days
  • lemonade@lemonade-sdk

    Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk

    5,585+55Star change over the last 7 days
  • LACT@ilya-zlobintsev

    Linux GPU Configuration And Monitoring Tool

    5,538+32Star change over the last 7 days
  • tf-quant-finance@google

    High-performance TensorFlow library for quantitative finance.

    5,493+7Star change over the last 7 days
  • koharu@koharu-rs

    AI-powered manga translator, written in Rust.

    5,465+32Star change over the last 7 days
  • gpuweb@gpuweb

    Where the GPU for the Web work happens!

    5,464+3Star change over the last 7 days
  • rust-cuda@Rust-GPU

    Ecosystem of libraries and tools for writing and executing fast GPU code fully in Rust.

    5,334+3Star change over the last 7 days
  • cuml@NVIDIA

    NVIDIA cuML: GPU-Accelerated Machine Learning

    5,269-1Star change over the last 7 days
  • FluidX3D@ProjectPhysX

    The fastest and most memory efficient lattice Boltzmann CFD software, running on all GPUs and CPUs via OpenCL. Free for non-commercial use.

    5,257+3Star change over the last 7 days
  • nccl@NVIDIA

    Optimized primitives for collective multi-GPU communication

    5,041+9Star change over the last 7 days
  • Ultralight@ultralight-ux

    Lightweight, high-performance HTML renderer for game and app developers.

    5,012+3Star change over the last 7 days
  • executorch@pytorch

    On-device AI across mobile, embedded and edge for PyTorch

    4,980+15Star change over the last 7 days
  • Time series forecasting with PyTorch

    4,979+1Star change over the last 7 days
  • arrayfire@arrayfire

    ArrayFire: a general purpose GPU library.

    4,903+0Star change over the last 7 days
  • MegEngine@MegEngine

    MegEngine 是一个快速、可拓展、易于使用且支持自动求导的深度学习框架

    4,811+1Star change over the last 7 days
  • asitop@tlkh

    Perf monitoring CLI tool for Apple Silicon

    4,631+5Star change over the last 7 days
  • tiny-cuda-nn@NVlabs

    Lightning fast C++/CUDA neural network framework

    4,530+1Star change over the last 7 days
  • gpustat@wookayin

    📊 A simple command-line utility for querying and monitoring GPU status

    4,392+0Star change over the last 7 days
  • llm-d@llm-d

    Achieve state of the art inference performance with modern accelerators on Kubernetes

    4,390+69Star change over the last 7 days
  • deepflow@deepflowio

    eBPF Observability - Distributed Tracing and Profiling

    4,252+6Star change over the last 7 days
  • tvm-cn@hyperai

    TVM Documentation in Chinese Simplified / TVM 中文文档

    3,927+7Star change over the last 7 days
  • Brightroom@FluidGroup

    📷 A composable image editor using Core Image and Metal.

    3,668+1Star change over the last 7 days
  • StringZilla@ashvardanian

    Up to 100x faster strings for C, C++, CUDA, Python, Rust, Swift, JS, & Go, leveraging NEON, AVX2, AVX-512, SVE, GPGPU, & SWAR to accelerate search, hashing, sorting, edit distances, sketches, and memory ops 🦖

    3,549+6Star change over the last 7 days
  • ml-workspace@ml-tooling

    🛠 All-in-one web-based IDE specialized for machine learning and data science.

    3,543-1Star change over the last 7 days
  • A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance with lower memory utilization in both training and inference.

    3,518+9Star change over the last 7 days
← Back to topics