HierSpeech++: Bridging the Gap between Semantic and Acoustic Representation of Speech by Hierarchical Variational Inference for Zero-shot Speech Synthesis
고려대 · arXiv · 2023 · 검증 완료 검증 완료복수의 신뢰 가능한 출처로 검증되었습니다.
계층적 변분추론으로 의미·음향 표현의 간극을 메워 제로샷 TTS와 음성변환을 강건하게 수행하는 음성합성 모델.
- arXiv
- 2311.12454
- 원문
- https://arxiv.org/abs/2311.12454
출처 및 검증
- [paper] https://arxiv.org/abs/2311.12454
마지막 검증일: 2026-07-22