FastSpeech2 Alternatives

Similar projects and alternatives to FastSpeech2

TTS

231 28,959 9.5 Python FastSpeech2 VS TTS

🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
tortoise-tts

144 11,686 8.2 Jupyter Notebook FastSpeech2 VS tortoise-tts

A multi-voice TTS system trained with an emphasis on quality
InfluxDB

www.influxdata.com
sponsored

Power Real-Time Data Analytics at Scale. Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.
Real-Time-Voice-Cloning

96 50,652 0.0 Python FastSpeech2 VS Real-Time-Voice-Cloning

Clone a voice in 5 seconds to generate arbitrary speech in real-time
tacotron2

28 4,882 0.0 Jupyter Notebook FastSpeech2 VS tacotron2

Tacotron 2 - PyTorch implementation with faster-than-realtime inference
NeMo

29 9,951 9.8 Python FastSpeech2 VS NeMo

NeMo: a framework for generative AI
speechbrain

26 7,836 9.8 Python FastSpeech2 VS speechbrain

A PyTorch-based Speech Toolkit
Parallel-Tacotron2

1 184 0.0 Python FastSpeech2 VS Parallel-Tacotron2

PyTorch Implementation of Google's Parallel Tacotron 2: A Non-Autoregressive Neural TTS Model with Differentiable Duration Modeling
WorkOS

workos.com
sponsored

The modern identity platform for B2B SaaS. The APIs are flexible and easy-to-use, supporting authentication, user identity, and complex enterprise features like SSO and SCIM provisioning.
vits

6 6,206 0.0 Python FastSpeech2 VS vits

VITS: Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech
tacotron

3 2,919 0.0 Python FastSpeech2 VS tacotron

A TensorFlow implementation of Google's Tacotron speech synthesis with pre-trained model (unofficial)
flowtron

6 882 0.0 Jupyter Notebook FastSpeech2 VS flowtron

Flowtron is an auto-regressive flow-based generative network for text to speech synthesis with control over speech variation and style transfer
voice100

1 25 4.9 Python FastSpeech2 VS voice100

Voice100 includes neural TTS/ASR models. Inference of Voice100 is low cost as its models are tiny and only depend on CNN without autoregression.
hifi-gan

5 1,744 0.0 Python FastSpeech2 VS hifi-gan

HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis
radtts

1 270 0.0 Roff FastSpeech2 VS radtts

Provides training, inference and voice conversion recipes for RADTTS and RADTTS++: Flow-based TTS models with Robust Alignment Learning, Diverse Synthesis, and Generative Modeling and Fine-Grained Control over of Low Dimensional (F0 and Energy) Speech Attributes.
STYLER

3 150 1.8 Python FastSpeech2 VS STYLER

Official repository of STYLER: Style Factor Modeling with Rapidity and Robustness via Speech Decomposition for Expressive and Controllable Neural Text to Speech, INTERSPEECH 2021 (by keonlee9420)
waveglow

2 2,218 0.0 Python FastSpeech2 VS waveglow

A Flow-based Generative Network for Speech Synthesis
Speech-Backbones

1 518 0.0 Jupyter Notebook FastSpeech2 VS Speech-Backbones

This is the main repository of open-sourced speech technology by Huawei Noah's Ark Lab.
DiffSinger

1 223 10.0 Python FastSpeech2 VS DiffSinger

PyTorch implementation of DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (focused on DiffSpeech) (by keonlee9420)
SaaSHub

www.saashub.com
sponsored

SaaSHub - Software Alternatives and Reviews. SaaSHub helps you find the best software and product alternatives

NOTE: The number of mentions on this list indicates mentions on common posts plus user suggested alternatives. Hence, a higher number means a better FastSpeech2 alternative or higher similarity.

Suggest an alternative to FastSpeech2

FastSpeech2 reviews and mentions

Posts with mentions or reviews of FastSpeech2. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-04-13.

[D] What is the best open source text to speech model?
15 projects | /r/MachineLearning | 13 Apr 2023

FastSpeech2 submitted: Jun 8, 2020 paper: https://arxiv.org/pdf/2006.04558.pdf github: https://github.com/ming024/FastSpeech2 (Not the official implementation but is the once cited the most)
What voice-changing apps are available right now?
4 projects | /r/artificial | 29 Jun 2022

We have the TorToiSe repo, the SV2TTS repo, and from here you have the other models like Tacotron 2, FastSpeech 2, and such. A there is a lot that goes into training a baseline for these models on the LJSpeech and LibriTTS datasets. Fine tuning is left up to the user.
I'm looking for something self-hosted, preferably linux-based (though win or mac will work too), that will allow me to train a 'voice model' with pre-recorded speech, and then replicate it from text of my choice.
3 projects | /r/VocalSynthesis | 20 Feb 2022
Voice-cloning library for conlangs?
3 projects | /r/conlangs | 9 Nov 2021

As for synthesis of text using your own voice - you can dig into Real Time Voice Cloning or maybe FastSpeech2, but I am not sure if you can use it with conlangs (and because of ML nature, you need many, many, many training data to get anything interesting).
A note from our sponsor - InfluxDB
www.influxdata.com | 18 Apr 2024

Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality. Learn more →

Stats

Basic FastSpeech2 repo stats

Mentions

Stars

1,604

Activity

0.0

Last Commit

6 months ago

ming024/FastSpeech2 is an open source project licensed under MIT License which is an OSI approved license.

The primary programming language of FastSpeech2 is Python.

FastSpeech2

FastSpeech2 Alternatives

Similar projects and alternatives to FastSpeech2

FastSpeech2 reviews and mentions

Stats

Popular Comparisons