You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
License and capability clarification (2026-07-14): FunASR is a toolkit, not a single checkpoint. The FunASR and SenseVoice repository source code is MIT; model weights follow each model card. SenseVoiceSmall supports Chinese, Cantonese, English, Japanese, and Korean, and its weights use the linked FunASR Model Open Source License Agreement. Fun-ASR-Nano-2512 is Apache-2.0. Language coverage, punctuation, and performance depend on the selected model and runtime configuration.
Feature Request
MeloTTS is an excellent multi-lingual TTS project! FunASR could help with training data preparation — automatic speech-to-text annotation for audio datasets.
Precedent: GPT-SoVITS (58K stars) already uses FunASR for training data labeling.
Why FunASR for TTS data prep?
SenseVoice: 50+ languages including Chinese, English, Japanese, Korean — matching MeloTTS's language coverage
Paraformer: Accurate character-level timestamps for audio-text alignment
Note
License and capability clarification (2026-07-14): FunASR is a toolkit, not a single checkpoint. The FunASR and SenseVoice repository source code is MIT; model weights follow each model card. SenseVoiceSmall supports Chinese, Cantonese, English, Japanese, and Korean, and its weights use the linked FunASR Model Open Source License Agreement. Fun-ASR-Nano-2512 is Apache-2.0. Language coverage, punctuation, and performance depend on the selected model and runtime configuration.
Feature Request
MeloTTS is an excellent multi-lingual TTS project! FunASR could help with training data preparation — automatic speech-to-text annotation for audio datasets.
Precedent: GPT-SoVITS (58K stars) already uses FunASR for training data labeling.
Why FunASR for TTS data prep?
Example: