edge-tts
EmotiVoice
edge-tts | EmotiVoice | |
---|---|---|
4 | 5 | |
3,701 | 6,369 | |
- | - | |
6.7 | 8.9 | |
6 days ago | 3 months ago | |
Python | Python | |
GNU General Public License v3.0 only | Apache License 2.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
edge-tts
-
[discussion] text to voice generation for textbooks (non-math part)
i would very much like to use it to turn the text parts of a book into an audio where i could listen to it while reading. i used edge's tts for speech by giving a paragraph to clipboard and to edge-tts in order to listen the text but it causes two problems: 1. you need internet connection and have the book opened 2. can only do paragraph by paragraph, and is prone to errors or sometimes if you use it too much it wont convert the full text afterwards.
-
Building audiobooks for any documents (pdf, epub, doc, etc)
This sounds ambitious. I can recommend Edge TTS. Sounds good, albeit with some misspellings. You can see the code here: https://github.com/jing332/tts-server-android Here: https://github.com/rany2/edge-tts Edge TTS is not the only solution, but among the multilingual TTS you will hardly find something better.
-
[DEV] Super realistic voices in Tasker with Elevenlabs! Combine with Chat GPT to create a very impressive assistant!
I became aware of it through the Home Assistant integration of edge-tts. You can install edge-tts with pip in Termux. For the audio output I also installed 'mpv' in Termux.
- Running Edge TTS Directly With Python - Error Message
EmotiVoice
- FLaNK Stack Weekly 12 February 2024
-
WhisperSpeech – An Open Source text-to-speech system built by inverting Whisper
Interested to see how it performs for Mandarin Chinese speech synthesis, especially with prosody and emotion. The highest quality open source model I've seen so far is EmotiVoice[0], which I've made a CLI wrapper around to generate audio for flashcards.[1] For EmotiVoice, you can apparently also clone your own voice with a GPU, but I have not tested this.[2]
[0] https://github.com/netease-youdao/EmotiVoice
[1] https://github.com/siraben/emotivoice-cli
[2] https://github.com/netease-youdao/EmotiVoice/wiki/Voice-Clon...
-
Microsoft releases Windows AI studio to run and fine tune models locally
Interesting. I'll have to check to be sure, but I think maybe something is happening automagically if you have reasonably up to date nvidia drivers on the host OS, because I was able to run the EmotiVoice TTS docker (which requires nvidia gpu) from WSL2.
https://github.com/netease-youdao/EmotiVoice
- FLaNK Stack Weekly for 13 November 2023
- EmotiVoice: A Multi-Voice and Prompt-Controlled TTS Engine
What are some alternatives?
tts-server-android - 这是一个Android系统TTS应用,内置微软演示接口,可自定义HTTP请求,可导入其他本地TTS引擎,以及根据中文双引号的简单旁白/对话识别朗读 ,还有自动重试,备用配置,文本替换等更多功能。| Microsoft TTS Android APP implementation (Use demo API)
Cgml - GPU-targeted vendor-agnostic AI library for Windows, and Mistral model implementation.
tortoise-tts - A multi-voice TTS system trained with an emphasis on quality
TTS - 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
crowdcast - Converts a subreddit into a podcast
draw-a-ui - Draw a mockup and generate html for it
MockingBird - 🚀AI拟声: 5秒内克隆您的声音并生成任意语音内容 Clone a voice in 5 seconds to generate arbitrary speech in real-time
vits - VITS: Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech
lhotse - Tools for handling speech data in machine learning projects.
voice100 - Voice100 includes neural TTS/ASR models. Inference of Voice100 is low cost as its models are tiny and only depend on CNN without autoregression.
clipea - 📎🟢 Like Clippy but for the CLI. A blazing fast AI helper for your command line