data-augmentation
Tracked open-source repos tagged data-augmentation, sorted by stars.
Related topics
Topics that frequently appear alongside data-augmentation on the same repo.
Recent risers
Repos created in the last 90 days, tagged data-augmentation.
No new repos tagged with this topic in the last 90 days.
- #1★ 6,005+1Star change over the last 7 days
- #2
A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference applications.
★ 5,740+2Star change over the last 7 days - #3
🔥🔥High-Performance Face Recognition Library on PaddlePaddle & PyTorch🔥🔥
★ 3,591+1Star change over the last 7 days - #4
TextAttack 🐙 is a Python framework for adversarial attacks, data augmentation, and model training in NLP https://textattack.readthedocs.io/en/master/
★ 3,473+0Star change over the last 7 days - #5
A high-performance Python-based I/O system for large (and small) deep learning problems, with strong support for PyTorch.
★ 3,175+2Star change over the last 7 days - #6★ 2,440+0Star change over the last 7 days
- #7
A Python library for audio data augmentation. Useful for making audio ML models work well in the real world, not just in the lab.
★ 2,314+0Star change over the last 7 days - #8
🎨 NeMo Data Designer: Generate high-quality synthetic data from scratch or from seed data.
★ 2,197+7Star change over the last 7 days - #9
fastdup is a powerful, free tool designed to rapidly generate valuable insights from image and video datasets. It helps enhance the quality of both images and labels, while significantly reducing data operation costs, all with unmatched scalability.
★ 1,906-1Star change over the last 7 days - #10★ 1,879+1Star change over the last 7 days
- #11
List of useful data augmentation resources. You will find here some not common techniques, libraries, links to GitHub repos, papers, and others.
★ 1,639+0Star change over the last 7 days - #12
Code for TKDE paper "Self-supervised learning on graphs: Contrastive, generative, or predictive"
★ 1,436+0Star change over the last 7 days - #13
This repository collects papers for "A Survey on Knowledge Distillation of Large Language Models". We break down KD into Knowledge Elicitation and Distillation Algorithms, and explore the Skill & Vertical Distillation of LLMs.
★ 1,306+1Star change over the last 7 days - #14
Fast audio data augmentation in PyTorch. Inspired by audiomentations. Useful for deep learning.
★ 1,168+1Star change over the last 7 days - #15
Awesome papers about generative Information Extraction (IE) using Large Language Models (LLMs)
★ 1,062+2Star change over the last 7 days - #16
Natural Language Toolkit for Indic Languages aims to provide out of the box support for various NLP tasks that an application developer might need
★ 838+1Star change over the last 7 days - #17
A library for generating and evaluating synthetic tabular data for privacy, fairness and data augmentation.
★ 681+3Star change over the last 7 days - #18
CAIRI Supervised, Semi- and Self-Supervised Visual Representation Learning Toolbox and Benchmark
★ 658+0Star change over the last 7 days - #19
Light-weight Single Person Pose Estimator
★ 647+0Star change over the last 7 days - #20
Augmentation pipeline for rendering synthetic paper printing, faxing, scanning and copy machine processes
★ 576+2Star change over the last 7 days - #21
Next-generation Albumentations: dual-licensed for open-source and commercial use
★ 548+2Star change over the last 7 days