跳到主要內容
buildradar
登入

jishengpeng/WavTokenizer

@jishengpeng

[ICLR 2025] 針對音訊語言建模、每秒 40/75 個 token 的 SOTA 離散聲學編解碼器模型

星數
1,319
Fork 數
116
語言
Python
授權
MIT
最後推送
2 年前
Pythontext-to-speechgpt4osemanticcodecdacacousticspeech-representationaudio-representationencodecmusic-representation-learningsoundstreamspeech-language-model

還沒有相關情報

radar 追蹤的來源裡還沒有出現過這個 repo。收集器照排程執行——等它涵蓋到這個 repo 再回來看看。