Skip to main content
buildradar
Sign in
Topic · model-compression

model-compression

Tracked open-source repos tagged model-compression, sorted by stars.

Repos
16
Total stars
30,933
Avg. stars
1,933
Share
0.00%

Topics that frequently appear alongside model-compression on the same repo.

Recent risers

Repos created in the last 90 days, tagged model-compression.

No new repos tagged with this topic in the last 90 days.

  • Efficient-AI-Backbones@huawei-noah

    Efficient AI Backbones including GhostNet, TNT and MLP, developed by Huawei Noah's Ark Lab.

    4,419+1Star change over the last 7 days
  • Awesome Knowledge Distillation

    3,903-1Star change over the last 7 days
  • Torch-Pruning@VainF

    [CVPR 2023] DepGraph: Towards Any Structural Pruning; LLMs, Vision Foundation Models, etc.

    3,351+3Star change over the last 7 days
  • Pretrained language model and its related optimization techniques developed by Huawei Noah's Ark Lab.

    3,166+0Star change over the last 7 days
  • A curated list of neural network pruning resources.

    2,497+0Star change over the last 7 days
  • A list of papers, docs, codes about model quantization. This repo is aimed to provide the info for model quantization research, we are continuously improving the project. Welcome to PR the works (papers, repositories) that are missed by the repo.

    2,433+0Star change over the last 7 days
  • micronet@666DZY666

    micronet, a model compression and deploy lib. compression: 1、quantization: quantization-aware-training(QAT), High-Bit(>2b)(DoReFa/Quantization and Training of Neural Networks for Efficient Integer-Arithmetic-Only Inference)、Low-Bit(≤2b)/Ternary and Binary(TWN/BNN/XNOR-Net); post-training-quantization(PTQ), 8-bit(tensorrt); 2、 pruning: normal、regular and group convolutional channel pruning; 3、 group convolution structure; 4、batch-normalization fuse for quantization. deploy: tensorrt, fp32/fp16/int8(ptq-calibration)、op-adapt(upsample)、dynamic_shape

    2,265-1Star change over the last 7 days
  • model-optimization@tensorflow

    A toolkit to optimize ML models for deployment for Keras and TensorFlow, including quantization and pruning.

    1,578+0Star change over the last 7 days
  • Efficient-Computing@huawei-noah

    Efficient computing methods developed by Huawei Noah's Ark Lab

    1,307+0Star change over the last 7 days
  • channel-pruning@ethanhe42

    Channel Pruning for Accelerating Very Deep Neural Networks (ICCV'17)

    1,088+0Star change over the last 7 days
  • DeepCache@horseee

    [CVPR 2024] DeepCache: Accelerating Diffusion Models for Free

    970-1Star change over the last 7 days
  • Collection of recent methods on (deep) neural network compression and acceleration.

    957+0Star change over the last 7 days
  • TinyNeuralNetwork@alibaba

    TinyNeuralNetwork is an efficient and easy-to-use deep learning model compression framework.

    879-1Star change over the last 7 days
  • List of papers related to neural network quantization in recent AI conferences and journals.

    851+6Star change over the last 7 days
  • SqueezeLLM@SqueezeAILab

    [ICML 2024] SqueezeLLM: Dense-and-Sparse Quantization

    724+1Star change over the last 7 days
  • Awesome machine learning model compression research papers, quantization, tools, and learning material.

    546+0Star change over the last 7 days
← Back to topics