rhasspy
kaldi-gstreamer-server
Our great sponsors
rhasspy | kaldi-gstreamer-server | |
---|---|---|
26 | 4 | |
2,263 | 1,054 | |
2.7% | - | |
2.3 | 0.0 | |
9 months ago | over 3 years ago | |
Shell | Python | |
MIT License | BSD 2-clause "Simplified" License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
rhasspy
-
New project: Grocy Rhasspy Skill
I've been working on this for a few months now and I think I have it to a point where I am ready to share. This is definitely a very niche solution but I am creating a new skill handler for Grocy for the Open Source Voice Assistant Rhasspy (https://github.com/rhasspy/rhasspy). My handler is here: https://github.com/MCHellspawn/hermes-app-grocy. It is not complete yet but getting there. With is skill and a working Rhasspy 2.5 setup you can do a lot of tasks in Grocy with your voice. So far you can create and delete shopping lists, create products and add and remove them from shopping lists, list chores, mark them complete or skipped, and more.
-
Ask HN: Home Voice Assistant Recommendations
It'll run on a cheap Ubuntu box if you can't get a Pi.
And lots of people seem to like Rhasspy too:
https://github.com/rhasspy/rhasspy
-
The failure of Amazon's Alexa shows Microsoft was right to kill Cortana
Here is one example https://community.rhasspy.org/
-
Someone has to say it: Voice assistants are not doing it for big tech
I tried Amazon's Alexa, the top end model with a display. Often it would taunt you about new/interesting things on the screen, but I could never get them to work. I'd had to memorize things to get even the basics working. Ended up unplugging it.
However Google's Assistant in comparison worked great, no memorization, and very useful. Sure time, weather, set timers, and alarms worked great with a very flexible set of natural language queries. Even more complex things like what will be the temperature tomorrow at 10pm, simple calculations and unit conversions. But also things like IMDB like queries about directors, actors, which movies someone was in, etc generally worked well. It seemed to really understand things, not just "A web search returned ...". Even more complex things like the wheelbase of a 2004 WRX would return an answer, not a search result.
With all that said I'm looking for a non-cloud/on site solution, even if it requires more work, most recently noticed https://github.com/rhasspy/rhasspy
- Rhasspy – Offline private voice assistant for many human languages
-
Google assistant alternatives?
I just found this one: https://github.com/rhasspy/rhasspy
kaldi-gstreamer-server
- Real-time full-duplex speech recognition server, based on Kaldi and GStreamer
- Ask HN: What problem are you close to solving and how can we help?
-
Open Source ASR with user-specific custom vocabularies?
Through my research, the most promising real-time transcription options appear to be Vosk or Kaldi Gstreamer. I’ve set them both up & they appear to work well for general transcription, but I’m not sure how to handle the user-specific custom vocabularies.
-
Speech to text software
It is kind of difficult to find something like this free of charge (and open source) since the ASR service needs to be hosted somewhere. If you are really interested in the topic then you could take a lit into kaldi and its pretrained models (but kaldi is kind of difficult to learn so I don't really recommend it if you want something quick) and then you could also combine that with kaldi-gstreamer in order to set up a server which you can turn on and off whenever you like.
What are some alternatives?
mycroft-core - Mycroft Core, the Mycroft Artificial Intelligence platform.
espnet - End-to-End Speech Processing Toolkit
ProjectAlice - Project Alice is a smart voice home assistant that is completely modular and extensible.
vosk-server - WebSocket, gRPC and WebRTC speech recognition server based on Vosk and Kaldi libraries
Kaldi Speech Recognition Toolkit - kaldi-asr/kaldi is the official location of the Kaldi project.
Home Assistant - :house_with_garden: Open source home automation that puts local control and privacy first.
bert-for-inference - A small repo showing how to easily use BERT (or other transformers) for inference
Leon - 🧠 Leon is your open-source personal assistant.
ChessPositionRanking - Software suite for ranking chess positions and accurately estimating the number of legal chess positions
rhino - On-device Speech-to-Intent engine powered by deep learning
mtpng - A parallelized PNG encoder in Rust