Skip to main content
buildradar
Sign in
Topic · sft

sft

Tracked open-source repos tagged sft, sorted by stars.

Repos
15
Total stars
63,142
Avg. stars
4,209
Share
0.00%

Topics that frequently appear alongside sft on the same repo.

Recent risers

Repos created in the last 90 days, tagged sft.

No new repos tagged with this topic in the last 90 days.

  • ms-swift@modelscope

    Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).

    15,400+76Star change over the last 7 days
  • bisheng@dataelement

    BISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and comprehensive features include: GenAI workflow, RAG, Agent, Unified model management, Evaluation, SFT, Dataset Management, Enterprise-level System Management, Observability and more.

    11,917+14Star change over the last 7 days
  • oumi@oumi-ai

    Easily fine-tune, evaluate and deploy Qwen, Gemma, or any open weight LLM!

    9,380+3Star change over the last 7 days
  • AgentGuide@adongwanai

    https://adongwanai.github.io/AgentGuide | AI Agent开发指南 | LangGraph实战 | 高级RAG | 转行大模型 | 大模型面试 | 算法工程师 | 面试题库 | 强化学习|数据合成

    8,947+308Star change over the last 7 days
  • hands-on-modern-rl@walkinglabs

    🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.

    4,145+94Star change over the last 7 days
  • Soup@MakazhanAlpamys

    Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.

    3,564+883Star change over the last 7 days
  • maxtext@AI-Hypercomputer

    A simple, performant, and scalable Jax LLM!

    2,407+6Star change over the last 7 days
  • chatglm_finetuning@ssbuild

    chatglm 6b finetuning and alpaca finetuning

    1,526+0Star change over the last 7 days
  • diy-llm@datawhalechina

    🎓 系统性大语言模型构建课程|🛠️ 覆盖预训练数据工程、Tokenizer、Transformer、MoE、GPU 编程 (CUDA/Triton)、分布式训练、Scaling Laws、推理优化及对齐 (SFT/RLHF/GRPO)|🚀 6 个渐进式作业 + 代码驱动,建立 LLM 全栈认知体系

    1,260+22Star change over the last 7 days
  • GraphGen@InternScience

    GraphGen: Enhancing Supervised Fine-Tuning for LLMs with Knowledge-Driven Synthetic Data Generation

    1,214+5Star change over the last 7 days
  • surogate@invergent-ai

    Training/Fine-tuning at the speed of light

    813+3Star change over the last 7 days
  • DeepSeek-671B-SFT-Guide@ScienceOne-AI

    An open-source solution for full parameter fine-tuning of DeepSeek-V3/R1 671B, including complete code and scripts from training to inference, as well as some practical experiences and conclusions. (DeepSeek-V3/R1 满血版 671B 全参数微调的开源解决方案,包含从训练到推理的完整代码和脚本,以及实践中积累一些经验和结论。)

    812+0Star change over the last 7 days
  • trainable-agents@choosewhatulike

    Code and datasets for "Character-LLM: A Trainable Agent for Role-Playing"

    645+1Star change over the last 7 days
  • tensorflow-nlp-tutorial@ukairia777

    tensorflow를 사용하여 텍스트 전처리부터, Topic Models, BERT, GPT, LLM과 같은 최신 모델의 다운스트림 태스크들을 정리한 Deep Learning NLP 저장소입니다.

    581+0Star change over the last 7 days
  • LoongForge@baidu-baige

    A high-performance framework for training LLMs, VLMs, diffusion, and embodied models on NVIDIA GPUs and Kunlun XPUs.

    538Star change over the last 7 days
← Back to topics