fix: 속도 조절을 length_scale에서 사후 타임스트레치로 분리(발음 뭉개짐 개선)
- 자연 속도(1.0)로 합성 후 Rubberband(R3)로 배속 → 발음/피치 보존 - 폴백: ffmpeg atempo → librosa. rubberband-cli/pyrubberband 추가 - 원인: MeloTTS speed는 length_scale 축소라 고배속에서 자음이 뭉개짐
This commit is contained in:
@@ -7,6 +7,7 @@ python-multipart==0.0.32
|
||||
soundfile==0.14.0
|
||||
librosa==0.9.1
|
||||
numpy==1.26.4
|
||||
pyrubberband==0.4.0
|
||||
|
||||
# MeloTTS 런타임 의존성 (정확한 버전은 constraints.txt 로 고정)
|
||||
# gradio/tensorboard/unidic(full) 등 미사용 무거운 의존성은 제외한다.
|
||||
|
||||
Reference in New Issue
Block a user