whisper-writer
LiveWhisper
whisper-writer | LiveWhisper | |
---|---|---|
2 | 2 | |
188 | 293 | |
- | - | |
6.6 | 0.0 | |
11 days ago | 5 months ago | |
Python | Python | |
GNU General Public License v3.0 only | MIT License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
whisper-writer
- Show HN: WhisperWriter β Speech-to-text using OpenAI's Whisper, coded by ChatGPT
-
Using ChatGPT to generate a GPT project end-to-end
I've also made six small apps completely coded by ChatGPT (with GitHub Copilot contributing a bit as well). Here are the two largest:
PlaylistGPT (https://github.com/savbell/playlist-gpt): A fun little web app that allows you to ask questions about your Spotify playlists and receive answers from Python code generated by OpenAI's models. I even added a feature where if the code written by GPT runs into errors, it can send the code and the error back to the model and ask it to fix it. It actually can debug itself quite often! One of the most impressive things for me was how it was able to model the UI after the Spotify app with little more than me asking it to do exactly that.
WhisperWriter (https://github.com/savbell/whisper-writer): A small speech-to-text app that uses OpenAI's Whisper API to auto-transcribe recordings from a user's microphone. It waits for a keyboard shortcut to be pressed, then records from the user's microphone until it detects a pause in their speech, and then types out the Whisper transcription to the active window. It only took me two hours to get a working prototype up and running, with additions such as graphic indicators taking a few more hours to implement.
I created the first for fun and the second to help me overcome a disability that impacts my ability to use a keyboard. I now use WhisperWriter literally every day (I'm even typing part of this comment with it), and I used it to prompt ChatGPT to write the code for a few additional personal projects that improve my quality-of-life in small ways. If people are interested, I may write up more about the prompting and pair programming process, since I definitely learned a lot as I worked through these, including some similar lessons to the article!
Personally, I am super excited about the possibilities these AI technologies open up for people like me, who may be facing small challenges that could be easily solved with a tiny app written in a few hours tailored specifically to their problem. I had been struggling to use my desktop computer because the Windows Dictation tool was very broken for me, but now I feel like I can use it to my full capacity again because I can type with WhisperWriter. Coding now takes a minimal amount of keyboard use thanks to these AI coding assistants -- and I am super grateful for that!
LiveWhisper
-
Speech Recognition module in Python
I've run into this EXACT SAME problem, and ended up creating my own SpeechRecognition alternative, using sounddevice (which unlike pyaudio IS compatible with my Linux Mint's audio drivers), and OpenAI's Whisper model.. Cause that was my only option, other than risking messing up my audio drivers.. heh
-
How to install and deploy OpenAI Whisper with Python
If anyone's interested, I took a wack at making Whisper transcribe semi-live, to the terminal: https://github.com/Nikorasu/LiveWhisper
What are some alternatives?
kaldi-active-grammar - Python Kaldi speech recognition with grammars that can be set active/inactive dynamically at decode-time
whisper-openai-gradio-implementation - Whisper is an automatic speech recognition (ASR) system Gradio Web UI Implementation
WhisperLive - A nearly-live implementation of OpenAI's Whisper.
whisper-standalone-win - Whisper & Faster-Whisper standalone executables for those who don't want to bother with Python.
AI-Waifu-Vtuber - AI Vtuber for Streaming on Youtube/Twitch
web-whisper - OpenAI's Whisper Audio to text transcription right into your web browser! An open source AI subtitling suite.
playlist-gpt - πΆπ©βπ» A fun little web app that analyzes your Spotify playlists with help from OpenAI's language models.
SwiftWhisper - π€ The easiest way to transcribe audio in Swift
easy-chat - A ChatGPT UI for young readers, written by ChatGPT
FlorenceBot - A fully interactive domain-specific chatbot implemented using Prolog and PySwip.
whisper-subtitles-webui - A gradio interface for making transcribed and translated subtitles for videos