speech_dataset

开源 开源免费 音频音乐
GitHub 星标 ★ 466
维护状态 低维护
是否开源
定价模式 开源免费

项目数据

分类音频音乐
开发团队double22a
所属国家
官网地址
定价模式开源免费
价格区间
是否开源
开源协议Apache-2.0
主要语言
技术栈/模型asr,audio,automatic-speech-recognition,dataset,deep-learning,deep-neural-networks,speech,speech-diarization,speech-enhancement,speech-recognition,speech-segmentation,speech-separation,speech-synthesis,speech-to-text,speech-translation,text-to-speech,tts,voice-conversion,wav
GitHub 星标★ 466
30天Star增速
HF 下载量
上线时间2021-04-07 00:00:00
最近更新2026-07-31 00:00:00
维护状态低维护
中文支持
访问方式
移动端支持
综合评分
收录时间2026-08-09
浏览次数0

核心亮点

    不足之处

      适用场景

        替代项目

        项目介绍

        speech_dataset 是一个音频音乐领域的开源项目,官方简介:The dataset of Speech Recognition。项目使用 未知 开发,在 GitHub 上获得 466 星标。

        上一篇:StyleTTS

        下一篇:AivisSpeech

        同类项目推荐

        StyleTTS2 开源

        音色情绪几乎以假乱真,听不出是机器在读

        StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversari···

        ★ 6328 2026-08-09
        tuneflow-py 开源

        编曲软件里直接跑 AI,灵感落地不用切工具

        + Build your music algorithms and AI models with the next-gen DAW

        ★ 890 2026-08-09
        alda-clj 开源

        A Clojure library for live-coding music with Alda

        ★ 70 2026-08-09