FireRedTeam
FireRedTeam's tracked open-source repos, sorted by stars.
- #1
FireRed-OpenStoryline is an AI video editing agent that transforms manual editing into intention-driven directing through natural language interaction, LLM-powered planning, and precise tool orchestration. It facilitates transparent, human-in-the-loop creation with reusable Style Skills for consistent, professional storytelling.
★ 3,356+43Star change over the last 7 days - #2
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering outstanding singing lyrics recognition capability.
★ 1,975+3Star change over the last 7 days - #3
Long-form streaming TTS system for multi-speaker dialogue generation
★ 1,430+1Star change over the last 7 days - #4
FireRed-Image-Edit is a powerful image editing foundation model achieving open-source state-of-the-art performance with precise instruction following, high-fidelity generation, superior identity consistency, and seamless multi-element fusion.
★ 1,350+6Star change over the last 7 days - #5
An Open-Sourced LLM-empowered Foundation TTS System
★ 921+1Star change over the last 7 days - #6
StoryMaker: Towards consistent characters in text-to-image generation
★ 723+0Star change over the last 7 days - #7
A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/accents), English, code-switching, and both speech and singing ASR. FireRedVAD supports speech/singing/music in 100+ langs. FireRedLID supports 100+ langs and 20+ zh dialects. FireRedPunc supports zh and en.
★ 669+10Star change over the last 7 days - #8
A Fully Self-Hosted Solution for Full-Duplex Voice Interaction
★ 585+3Star change over the last 7 days - #9
A SOTA Industrial-Grade Voice Activity Detection & Audio Event Detection, supporting 100+ languages, outperforming Silero-VAD, TEN-VAD, FunASR-VAD and WebRTC-VAD
★ 519+8Star change over the last 7 days