Skip to main content
buildradar
Sign in
Language · Cuda

Cuda

Tracked open-source repos with Cuda as the primary language, sorted by stars.

61 repos
  • raft@NVIDIA

    RAFT contains fundamental widely-used algorithms and primitives for machine learning and information retrieval. The algorithms are CUDA-accelerated and form building blocks for more easily writing high performance applications.

    1,040+2Star change over the last 7 days
  • cuopt@NVIDIA

    GPU accelerated decision optimization

    1,036+4Star change over the last 7 days
  • NKSR@nv-tlabs

    [CVPR 2023 Highlight] Neural Kernel Surface Reconstruction

    990+2Star change over the last 7 days
  • CudaSift@Celebrandil

    A CUDA implementation of SIFT for NVidia GPUs (1.2 ms on a GTX 1060)

    953+0Star change over the last 7 days
  • nvbench@NVIDIA

    CUDA Kernel Benchmarking Library

    924+0Star change over the last 7 days
  • Examples demonstrating available options to program multiple GPUs in a single node or a cluster

    913+2Star change over the last 7 days
  • tron@NovaCodeu

    USDT钱包靓号生成器,Trx靓号,波场GPU靓号源码,TRON钱包靓号生成,波场地址靓号生成器

    868+0Star change over the last 7 days
  • lietorch@princeton-vl
    855+1Star change over the last 7 days
  • cuvs@NVIDIA

    cuVS - a library for vector search and clustering on the GPU

    848+4Star change over the last 7 days
  • GPUMD@brucefan1983

    Graphics Processing Units Molecular Dynamics

    833+3Star change over the last 7 days
  • NeuS2@19reborn

    [ICCV 2023] Official code for NeuS2

    736+0Star change over the last 7 days
  • ai-infra-hpc@jinbooooom

    hpc 教程,包含集合通信(mpi、nccl)、cuda 编程、向量化 SIMD、RDMA 通信等

    703+13Star change over the last 7 days
  • AMGX@NVIDIA

    Distributed multigrid linear solver library on GPU

    691+1Star change over the last 7 days
  • radfoam@theialab

    Original implementation of "Radiant Foam: Real-Time Differentiable Ray Tracing"

    665+0Star change over the last 7 days
  • unet.cu@clu0

    UNet diffusion model in pure CUDA

    663+0Star change over the last 7 days
  • cuCollections@NVIDIA
    663+2Star change over the last 7 days
  • recommenders-addons@tensorflow

    Additional utils and helpers to extend TensorFlow when build recommendation systems, contributed and maintained by SIG Recommenders.

    635+0Star change over the last 7 days
  • GPU@a-hamdi

    100 days of building GPU kernels!

    629-1Star change over the last 7 days
  • fast.cu@pranjalssh

    Fastest kernels written from scratch

    621+5Star change over the last 7 days
  • DiffPhysDrone@HenryHuYu

    Published on Nature Machine Intelligence! The first real robot(quadrotor) based on differentiable physics training.

    615+6Star change over the last 7 days
  • gpuRIR@DavidDiazGuerra

    Python library for Room Impulse Response (RIR) simulation with GPU acceleration

    611+0Star change over the last 7 days
  • cudahandbook@ArchaeaSoftware

    Source code that accompanies The CUDA Handbook.

    598+0Star change over the last 7 days
  • cuda_hgemm@Bruce-Lee-LY

    Several optimization methods of half-precision general matrix multiplication (HGEMM) using tensor core with WMMA API and MMA PTX instruction.

    568+0Star change over the last 7 days
  • CUDA Learning guide

    568-1Star change over the last 7 days
  • DarkPose@ilovepose

    Distribution-Aware Coordinate Representation for Human Pose Estimation

    568+0Star change over the last 7 days
  • qwen600@yassa9

    Static suckless single batch CUDA-only qwen3-0.6B mini inference engine

    560+0Star change over the last 7 days
  • flash attention tutorial written in python, triton, cuda, cutlass

    534+0Star change over the last 7 days
  • megalodon@XuezheMax

    Reference implementation of Megalodon 7B model

    527+0Star change over the last 7 days
  • mnist-cuda@Infatoshi
    517+2Star change over the last 7 days
  • BasicCUDA@CalvinXKY

    A tutorial for CUDA&PyTorch

    515+3Star change over the last 7 days
← Back to languages