WAAS
frogbase
WAAS | frogbase | |
---|---|---|
12 | 14 | |
1,743 | 754 | |
1.4% | - | |
7.0 | 4.3 | |
12 days ago | 7 months ago | |
JavaScript | Python | |
Apache License 2.0 | MIT License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
WAAS
-
Show HN: Minutes – Save up to 20% of salespeople time
This app does it locally on newer macs for free.
https://apps.apple.com/no/app/jojo-transcribe/id1659864300?m...
And open source: https://github.com/schibsted/WAAS
-
Only I don't use docker. WAAS. What alternatives to a docker install are there?
Clone the repo, then run the steps in the dockerfile: https://github.com/schibsted/WAAS/blob/main/Dockerfile
-
Whispers AI Modular Future
What utilities related to Whisper do you wish existed? What have you had to build yourself?
On the end user application side, I wish there was something that let me pick a podcast of my choosing, get it fully transcribed, and get an embeddings search plus answer q&a on top of that podcast or set of chosen podcasts. I've seen ones for specific podcasts, but I'd like one where I can choose the podcast. (Probably won't build it)
Also on the end user side, I wish there was an Otter alternative (still paid $30/mo, but unlimited minutes per month) that had longer transcription limits. (Started building this, not much interest from users though)
Things I've seen on the dev tool side:
Gladia (API call version of Whisper)
Whisper.cpp
Whisper webservice (https://github.com/ahmetoner/whisper-asr-webservice) - via this thread
Live microphone demo (not real time, it still does it in chunks) https://github.com/mallorbc/whisper_mic
Streamlit UI https://github.com/hayabhay/whisper-ui
Whisper playground https://github.com/saharmor/whisper-playground
Real time whisper https://github.com/shirayu/whispering
Whisper as a service https://github.com/schibsted/WAAS
Improved timestamps and speaker identification https://github.com/m-bain/whisperX
MacWhisper https://goodsnooze.gumroad.com/l/macwhisper
Crossplatform desktop Whisper that supports semi-realtime https://github.com/chidiwilliams/buzz
- [task] Write out the questions and responses from my podcast videos so I can turn it into a written article for my site - $5 per video x 15 videos = $75 - $80
- Self-host Whisper As a Service with GUI and queueing. Schibsted created a transcription service for our journalists to transcribe audio interviews and podcasts really quick.
- Show HN: Self-host Whisper As a Service with GUI and queueing
frogbase
-
For people who tried whisper
If you’re looking to use a local deployment and are comfortable with python projects, this deployment was fairly easy to use: https://github.com/hayabhay/whisper-ui
-
I have a two-step process for taking notes on transcripts that I'd like to share, but am also looking for feedback for the final step
I recommend using chat GPT to learn some very basic python/programming tho. I like this one https://github.com/hayabhay/whisper-ui which uses streamlit for easy to use UI for bulk transcriptions.
-
(Preferably) Self Hosted Podcasts with searchable transcripts
This tool 2 out of the 4 items you mentioned: https://github.com/hayabhay/whisper-ui
-
Whispers AI Modular Future
What utilities related to Whisper do you wish existed? What have you had to build yourself?
On the end user application side, I wish there was something that let me pick a podcast of my choosing, get it fully transcribed, and get an embeddings search plus answer q&a on top of that podcast or set of chosen podcasts. I've seen ones for specific podcasts, but I'd like one where I can choose the podcast. (Probably won't build it)
Also on the end user side, I wish there was an Otter alternative (still paid $30/mo, but unlimited minutes per month) that had longer transcription limits. (Started building this, not much interest from users though)
Things I've seen on the dev tool side:
Gladia (API call version of Whisper)
Whisper.cpp
Whisper webservice (https://github.com/ahmetoner/whisper-asr-webservice) - via this thread
Live microphone demo (not real time, it still does it in chunks) https://github.com/mallorbc/whisper_mic
Streamlit UI https://github.com/hayabhay/whisper-ui
Whisper playground https://github.com/saharmor/whisper-playground
Real time whisper https://github.com/shirayu/whispering
Whisper as a service https://github.com/schibsted/WAAS
Improved timestamps and speaker identification https://github.com/m-bain/whisperX
MacWhisper https://goodsnooze.gumroad.com/l/macwhisper
Crossplatform desktop Whisper that supports semi-realtime https://github.com/chidiwilliams/buzz
-
[P] Whisper-UI Update: You can now bulk-transcribe, save & search transcriptions with Streamlit & SQLAlchemy 2.0 [details in the comments]
Github Repo: https://github.com/hayabhay/whisper-ui
-
Self-host Whisper As a Service with GUI and queueing. Schibsted created a transcription service for our journalists to transcribe audio interviews and podcasts really quick.
People may also like this tool which is a bit more about searching the contents. https://github.com/hayabhay/whisper-ui
- Show HN: Self-host Whisper As a Service with GUI and queueing
-
Audio equivalent of paperless?
There is whisper ui to create meta data (speech 2 text).
- Whisper-UI: You can now bulk-transcribe, save & search transcriptions from YouTube with OpenAI's Whisper, Streamlit & SQLAlchemy 2.0
- Whisper-UI Update: You can now bulk-transcribe, save & search transcriptions with Streamlit & SQLAlchemy 2.0
What are some alternatives?
whisper - Robust Speech Recognition via Large-Scale Weak Supervision
whisper.cpp - Port of OpenAI's Whisper model in C/C++
whisper_mic - Project that allows one to use a microphone with OpenAI whisper.
transcribe-anything - Input a local file or url and this service will transcribe it using Whisper AI. Completely private and Free 🤯🤯🤯
whisper-playground - Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/
FlexGen - Running large language models on a single GPU for throughput-oriented scenarios.
openai-whisper-cpu - Improving transcription performance of OpenAI Whisper for CPU based deployment
whisper-asr-webservice - OpenAI Whisper ASR Webservice API
nlp