Comprehensive-E2E-TTS

开源 开源免费 音频音乐
GitHub 星标 ★ 147
维护状态 低维护
是否开源
定价模式 开源免费

项目数据

分类音频音乐
开发团队keonlee9420
所属国家
官网地址
定价模式开源免费
价格区间
是否开源
开源协议
主要语言Python
技术栈/模型deep-learning,end-to-end,fastspeech2,hifi-gan,jets,multi-speaker,neural-tts,non-ar,non-autoregressive,pytorch,single-speaker,sota,speech-synthesis,text-to-speech,text-to-wav,tts,ultimate-tts,unsupervised
GitHub 星标★ 147
30天Star增速
HF 下载量
上线时间2022-03-30 00:00:00
最近更新2026-03-14 00:00:00
维护状态低维护
中文支持
访问方式
移动端支持
综合评分
收录时间2026-08-09
浏览次数0

核心亮点

    不足之处

      适用场景

        替代项目

        项目介绍

        Comprehensive-E2E-TTS 是一个音频音乐领域的开源项目,官方简介:A Non-Autoregressive End-to-End Text-to-Speech (text-to-wav), supporting a family of SOTA unsupervised duration modelings. This project grows with the research community, aiming to achieve the ultimate E2E-TTS。项目使用 Python 开发,在 GitHub 上获得 147 星标。

        上一篇:Tacotron2-PyTorch

        下一篇:tts-arabic-pytorch

        同类项目推荐

        StyleTTS2 开源

        音色情绪几乎以假乱真,听不出是机器在读

        StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversari···

        ★ 6328 2026-08-09
        tuneflow-py 开源

        编曲软件里直接跑 AI,灵感落地不用切工具

        + Build your music algorithms and AI models with the next-gen DAW

        ★ 890 2026-08-09
        NeuroRVQ 开源

        NeuroRVQ: Multi-Scale Biosignal Tokenization for Generative Foundation Models

        ★ 51 2026-08-09
        voc2vec 开源

        This repository contains the code for the paper "voc2vec: A Foundation Model for Non···

        ★ 58 2026-08-09