WAAS
transcribe-anything
WAAS | transcribe-anything | |
---|---|---|
12 | 11 | |
1,743 | 362 | |
1.4% | - | |
7.0 | 9.3 | |
12 days ago | 13 days ago | |
JavaScript | Python | |
Apache License 2.0 | MIT License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
WAAS
-
Show HN: Minutes – Save up to 20% of salespeople time
This app does it locally on newer macs for free.
https://apps.apple.com/no/app/jojo-transcribe/id1659864300?m...
And open source: https://github.com/schibsted/WAAS
-
Only I don't use docker. WAAS. What alternatives to a docker install are there?
Clone the repo, then run the steps in the dockerfile: https://github.com/schibsted/WAAS/blob/main/Dockerfile
-
Whispers AI Modular Future
What utilities related to Whisper do you wish existed? What have you had to build yourself?
On the end user application side, I wish there was something that let me pick a podcast of my choosing, get it fully transcribed, and get an embeddings search plus answer q&a on top of that podcast or set of chosen podcasts. I've seen ones for specific podcasts, but I'd like one where I can choose the podcast. (Probably won't build it)
Also on the end user side, I wish there was an Otter alternative (still paid $30/mo, but unlimited minutes per month) that had longer transcription limits. (Started building this, not much interest from users though)
Things I've seen on the dev tool side:
Gladia (API call version of Whisper)
Whisper.cpp
Whisper webservice (https://github.com/ahmetoner/whisper-asr-webservice) - via this thread
Live microphone demo (not real time, it still does it in chunks) https://github.com/mallorbc/whisper_mic
Streamlit UI https://github.com/hayabhay/whisper-ui
Whisper playground https://github.com/saharmor/whisper-playground
Real time whisper https://github.com/shirayu/whispering
Whisper as a service https://github.com/schibsted/WAAS
Improved timestamps and speaker identification https://github.com/m-bain/whisperX
MacWhisper https://goodsnooze.gumroad.com/l/macwhisper
Crossplatform desktop Whisper that supports semi-realtime https://github.com/chidiwilliams/buzz
- [task] Write out the questions and responses from my podcast videos so I can turn it into a written article for my site - $5 per video x 15 videos = $75 - $80
- Self-host Whisper As a Service with GUI and queueing. Schibsted created a transcription service for our journalists to transcribe audio interviews and podcasts really quick.
- Show HN: Self-host Whisper As a Service with GUI and queueing
transcribe-anything
-
Summarize audio recordings in text
transcribe-anything
-
$620,000 stolen from YouTuber Ethan Klein and the H3 Podcast by MCN BroadbandTV and their CEO Shahrzad Rafati
OpenAI whisper. Here is a tool that has it, a video downloader, and some other things bundled in with it: https://github.com/zackees/transcribe-anything
- 32 Open Source Libraries for Python's 32nd Birthday
-
Show HN: Self-host Whisper As a Service with GUI and queueing
People interested in this might also be interested in transcribe-anything [1].
It automates video fetching and uses whisper to generate .srt, .vtt and .txt files.
[1] https://github.com/zackees/transcribe-anything
-
[P] Free Youtube Subtitles Generator
Nice looks great, link broken but it just needed a hyphen https://github.com/zackees/transcribe-anything
-
Gpu accelerated ML apps will soon get a lot easier to deploy - Pytorch-cuda moving to 100% pypi hosting.
Right now the cuda accelerated whls are hosted outside of pypi which can only be accessed by using `--extra-index-url`, when installing from a requirements file (pip install -r requirements.txt). However pip install doesn't allow --extra-index-url for security reasons, which means deploying cuda accelerated ML apps on python is a complicated affair, see this [script](https://github.com/zackees/transcribe-anything/blob/main/install_cuda.py) as an example of what needs to be done to uninstall conflicting cpu only version of pytorch and replace it with cuda acceleration.
- Bro, listen: Interact with OpenAI using voice
- Convert YouTube to Text with OpenAI Whisper
-
Draw an owl
Transcribe Anything
-
Transcribe Video/Audio on the web using `transcribe-anything`, a front end to WhisperAI
Code Repo: https://github.com/zackees/transcribe-anything (please give my repo a like)
What are some alternatives?
whisper - Robust Speech Recognition via Large-Scale Weak Supervision
frogbase - Transform audio-visual content into navigable knowledge.
whisper.cpp - Port of OpenAI's Whisper model in C/C++
Hentai-Diffusion - The official place for the best A.I.
whisper_mic - Project that allows one to use a microphone with OpenAI whisper.
whisperX - WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
whisper-playground - Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/
subtitle-generator - Generate subtitles for youtube videos for free with https://text-generator.io
openai-whisper-cpu - Improving transcription performance of OpenAI Whisper for CPU based deployment
static_ffmpeg - Installs FFMPEG v5 On Win32/Ubuntu/MacOS
whisper-asr-webservice - OpenAI Whisper ASR Webservice API
ai-notes - notes for software engineers getting up to speed on new AI developments. Serves as datastore for https://latent.space writing, and product brainstorming, but has cleaned up canonical references under the /Resources folder.