buzz
text-to-speech-ubuntu
buzz | text-to-speech-ubuntu | |
---|---|---|
21 | 1 | |
9,869 | 21 | |
- | - | |
8.5 | 4.3 | |
20 days ago | 2 months ago | |
Python | ||
MIT License | - |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
buzz
- Buzz: Transcribe and translate audio offline on your personal computer
- MacWhisper: Transcribe audio files on your Mac
-
Build Personal ChatGPT Using Your Data
Easiest 1-click way to install and use Stable Diffusion on your computer."
https://github.com/easydiffusion/easydiffusion
And while Whisper is OpenAI, it is trivial to use locally and extremely usefull
https://github.com/chidiwilliams/buzz
- automated transcription software that is HIPAA compliant?
- Question: Does anyone know of an AI or ChatGPT tool to create automatic SRT caption files by uploading a video?
- Brauchbare Speech-to-Text Lösungen für Windows?
-
I've pretty much had it with Premiere.
Install this for Resolve
-
As a Foreign student, I record letures alot so I can review it anytime. Thanks to Obsidian Audio Player its feels so effortless to review those records.
Just use this one: https://github.com/chidiwilliams/buzz
-
Whispers AI Modular Future
What utilities related to Whisper do you wish existed? What have you had to build yourself?
On the end user application side, I wish there was something that let me pick a podcast of my choosing, get it fully transcribed, and get an embeddings search plus answer q&a on top of that podcast or set of chosen podcasts. I've seen ones for specific podcasts, but I'd like one where I can choose the podcast. (Probably won't build it)
Also on the end user side, I wish there was an Otter alternative (still paid $30/mo, but unlimited minutes per month) that had longer transcription limits. (Started building this, not much interest from users though)
Things I've seen on the dev tool side:
Gladia (API call version of Whisper)
Whisper.cpp
Whisper webservice (https://github.com/ahmetoner/whisper-asr-webservice) - via this thread
Live microphone demo (not real time, it still does it in chunks) https://github.com/mallorbc/whisper_mic
Streamlit UI https://github.com/hayabhay/whisper-ui
Whisper playground https://github.com/saharmor/whisper-playground
Real time whisper https://github.com/shirayu/whispering
Whisper as a service https://github.com/schibsted/WAAS
Improved timestamps and speaker identification https://github.com/m-bain/whisperX
MacWhisper https://goodsnooze.gumroad.com/l/macwhisper
Crossplatform desktop Whisper that supports semi-realtime https://github.com/chidiwilliams/buzz
- Any suggestions for easy ways to add subtitles to YouTube videos?
text-to-speech-ubuntu
-
Ask HN: Are there any good open source Text-to-Speech tools?
https://github.com/gnat/text-to-speech-ubuntu
What are some alternatives?
whisper - Robust Speech Recognition via Large-Scale Weak Supervision
ollama-twice - Watch and hear endless conversations between two ollamas, hence the Two-Way Conversation Engine (TWICE)
openai-whisper-cpu - Improving transcription performance of OpenAI Whisper for CPU based deployment
StoryToolkitAI - An editing tool that uses AI to transcribe, understand content and search for anything in your footage, integrated with ChatGPT and other AI models
TTS - 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
audapolis - an editor for spoken-word audio with automatic transcription
timit - The DARPA TIMIT Acoustic-Phonetic Continuous Speech Corpus.
whisper-diarization - Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
tortoise-tts - A multi-voice TTS system trained with an emphasis on quality
opentts - Open Text to Speech Server