openWakeWord
whisper-live-transcription
openWakeWord | whisper-live-transcription | |
---|---|---|
5 | 3 | |
457 | 95 | |
- | - | |
8.4 | 7.9 | |
about 1 month ago | 4 months ago | |
Jupyter Notebook | Python | |
Apache License 2.0 | - |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
openWakeWord
-
OpenAI releases Whisper v3, new generation open source ASR model
https://github.com/dscripka/openWakeWord
Balancing wake reliability vs false wake activation is a tricky balance. OWW is decent but could certainly be better.
It's used with Home Assistant now so I expect the training data and implementation overall to get significantly better fairly soon.
-
Distil-Whisper: distilled version of Whisper that is 6 times faster, 49% smaller
There's also OpenWakeWord[0]. The models are readily available in tflite and ONNX formats and are impressively "light" in terms of compute requirements and performance.
It should be possible.
[0] - https://github.com/dscripka/openWakeWord
-
Real-Time Noise Suppression for PipeWire writen in Rust
hey, quick question. do you mind if I use your stft function in the speech preprocessing library I've been working on? we've been trying to add support for doing mel spectrograms to build a runner for openwakeword, but progress is pretty slow because I've been soloing something I really don't have the right background for(I've never directly studied or worked with signal processing)
-
I'm new to Rust but want to contribute
potentially build another runner for open wakeword
-
I want to contribute in a big project
here's what's on the pipeline next: - finish mel-spectrogram implementation - publish initial version on crates - move python caching rust side - finish implementing in the precise rust port - potentially build another runner for (open wakeword)[https://github.com/dscripka/openWakeWord] - build an android app that supports user-defined wakewords and has some popular defaults to load. ps not a voice assistant, just the thing that activates the voice assistant.
whisper-live-transcription
-
OpenAI releases Whisper v3, new generation open source ASR model
I implemented a dummy real-time (tested on Mac M1) transcription approach with Whisper. You can find the project here: https://github.com/gaborvecsei/whisper-live-transcription
The idea was to provide transcription results as fast as you can, and you can refine it along the way by providing more and more context.
-
ChatGPT can now see, hear, and speak – openai.com
Here's a link to a project that claims half second latency for the transcription part: https://github.com/gaborvecsei/whisper-live-transcription
- Show HN: Live Transcription with Whisper in a client-server setup
What are some alternatives?
WhisperInput - Offline voice input panel & keyboard with punctuation for Android.
chatcraft.org - Developer-oriented ChatGPT clone
mfcc-rust
talk - Let's make sand talk
project-2501 - Project 2501 is an open-source AI assistant, written in C++.
whisper-dictation - Dictation app based on the OpenAI speed to text models
Clippy - A bunch of lints to catch common mistakes and improve your Rust code. Book: https://doc.rust-lang.org/clippy/
faster-whisper-dictation - Dictation app based on the Faster Whisper transcription with CTranslate2
TX-2-simulator - Simulator for the pioneering TX-2 computer
CTranslate2 - Fast inference engine for Transformer models
DeepFilterNet - Noise supression using deep filtering