edge-tts
vits
edge-tts | vits | |
---|---|---|
4 | 6 | |
3,701 | 6,324 | |
- | - | |
6.7 | 0.0 | |
6 days ago | 5 months ago | |
Python | Python | |
GNU General Public License v3.0 only | MIT License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
edge-tts
-
[discussion] text to voice generation for textbooks (non-math part)
i would very much like to use it to turn the text parts of a book into an audio where i could listen to it while reading. i used edge's tts for speech by giving a paragraph to clipboard and to edge-tts in order to listen the text but it causes two problems: 1. you need internet connection and have the book opened 2. can only do paragraph by paragraph, and is prone to errors or sometimes if you use it too much it wont convert the full text afterwards.
-
Building audiobooks for any documents (pdf, epub, doc, etc)
This sounds ambitious. I can recommend Edge TTS. Sounds good, albeit with some misspellings. You can see the code here: https://github.com/jing332/tts-server-android Here: https://github.com/rany2/edge-tts Edge TTS is not the only solution, but among the multilingual TTS you will hardly find something better.
-
[DEV] Super realistic voices in Tasker with Elevenlabs! Combine with Chat GPT to create a very impressive assistant!
I became aware of it through the Home Assistant integration of edge-tts. You can install edge-tts with pip in Termux. For the audio output I also installed 'mpv' in Termux.
- Running Edge TTS Directly With Python - Error Message
vits
-
[D] TTS systems to download & run offline
And the voice encapsulation system VITS https://github.com/jaywalnut310/vits
- [D] What is the best open source text to speech model?
- githubで公開されている音声自動生成AI、日本のアニメキャラ2890名分の音声を学習素材に超速度で進化中
- 日本語英語中国語を読み上げできる音声自動生成AIがgithubで公開され話題に
- Adversarial Learning for End-to-End Text-to-Speech
- [R] Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech
What are some alternatives?
tts-server-android - 这是一个Android系统TTS应用,内置微软演示接口,可自定义HTTP请求,可导入其他本地TTS引擎,以及根据中文双引号的简单旁白/对话识别朗读 ,还有自动重试,备用配置,文本替换等更多功能。| Microsoft TTS Android APP implementation (Use demo API)
tortoise-tts - A multi-voice TTS system trained with an emphasis on quality
tortoise-tts-fast - Fast TorToiSe inference (5x or your money back!)
crowdcast - Converts a subreddit into a podcast
tacotron2 - Tacotron 2 - PyTorch implementation with faster-than-realtime inference
TTS - 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
Parallel-Tacotron2 - PyTorch Implementation of Google's Parallel Tacotron 2: A Non-Autoregressive Neural TTS Model with Differentiable Duration Modeling
vall-e - An unofficial PyTorch implementation of the audio LM VALL-E
tacotron - A TensorFlow implementation of Google's Tacotron speech synthesis with pre-trained model (unofficial)
glow-tts - A Generative Flow for Text-to-Speech via Monotonic Alignment Search
w2v2-how-to - How to use our public wav2vec2 dimensional emotion model