kaldi-active-grammar vs concise-encoding

kaldi-active-grammar

Python Kaldi speech recognition with grammars that can be set active/inactive dynamically at decode-time (by daanzu)

Source Code

Suggest alternative

Edit details

concise-encoding

The secure data format for a modern world (by kstenerud)

Encoding JSON Specification XML Data structures Data Visualization Documentation Parsing Security Datastructures

Source Code

concise-encoding.org

Suggest alternative

Edit details

Our great sponsors

InfluxDB - Power Real-Time Data Analytics at Scale

WorkOS - The modern identity platform for B2B SaaS

SaaSHub - Software Alternatives and Reviews

Our great sponsors

kaldi-active-grammar		concise-encoding
	Project
10	Mentions	22
329	Stars	255
-	Growth	-
0.0	Activity	7.2
10 months ago	Latest Commit	7 months ago
Python	Language	ANTLR
GNU Affero General Public License v3.0	License	GNU General Public License v3.0 or later

The number of mentions indicates the total number of mentions that we've tracked plus the number of user suggested alternatives.
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.

kaldi-active-grammar

Posts with mentions or reviews of kaldi-active-grammar. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-11-21.

Ask HN: How do you get started with adding voice commands to a computer system?
2 projects | news.ycombinator.com | 21 Nov 2023

https://github.com/dictation-toolbox/dragonfly
https://github.com/daanzu/kaldi-active-grammar
AMD Screws Gamers: Sponsorships Likely Block DLSS
4 projects | /r/Amd | 4 Jul 2023
Software I’m Thankful For
16 projects | news.ycombinator.com | 23 Sep 2022
Why, in 2022, is there no high quality method for voice control of a PC?
7 projects | news.ycombinator.com | 28 Jan 2022

With an open system/engine, you can train your own personal speech model. For kaldi-active-grammar (https://github.com/daanzu/kaldi-active-grammar), you can do so without all that much difficulty, although the process/documentation could certainly use improvement.
I bootstrapped my personal speech model by retaining the commands from me using WSR. My voice is quite abnormal, and it took only 10 hours of speech data to train a model orders of magnitude more accurate than any generic model I've ever used. And of course, I retain much of my usage now with Kaldi, so my model improves more and more over time. A virtuous flywheel!
Ask HN: Anyone voice code? I had a stroke and can't use my left side
2 projects | news.ycombinator.com | 16 Jan 2022

I have been coding entirely by voice for approximately 10 years now (by hand long before that). Most of that time I have been using the Dragonfly (https://github.com/dictation-toolbox/dragonfly) library to construct my own customized voice coding system. The library is highly flexible and open source, allowing you to easily customize everything to suit what you need to be productive. It is perhaps the power user analogue to Dragon Naturally Speaking. With it, you can certainly be highly productive coding by voice. In fact, I develop kaldi-active-grammar (https://github.com/daanzu/kaldi-active-grammar), a free and open source speech recognition backend usable by Dragonfly, itself entirely by voice. There's also a community of voice coders using Dragonfly and other tools that build on top of it, such as Caster (https://github.com/dictation-toolbox/Caster).
Ask HN: Who Wants to Collaborate?
58 projects | news.ycombinator.com | 1 Jan 2022

- Demo: https://www.youtube.com/watch?v=Qk1mGbIJx3s / Software: https://github.com/daanzu/kaldi-active-grammar
Far field audio is usually harder for any speech system to get correct, so having a good quality mic and using it nearby will _usually_ help with the transcription quality. As a long time Linux user, I would love to see it get some more powerful voice tools - really hope that this opens up over the next few years. Feel free to drop me an email (on my profile) happy to help with setup on any of the above.
How can I make Mycroft recognize non verbal audio sounds to command it?
3 projects | /r/Mycroftai | 29 Jul 2021
Linux Voice recognition/dictation/voice assistant/ one handed operation?
4 projects | /r/linuxquestions | 27 Jul 2021
Disabled computer science student ISO advice about single-handed keyboards
5 projects | /r/ErgoMechKeyboards | 21 Apr 2021

kaldi repo: https://github.com/daanzu/kaldi-active-grammar

concise-encoding

Posts with mentions or reviews of concise-encoding. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2024-03-07.

Ask HN: What Underrated Open Source Project Deserves More Recognition?
63 projects | news.ycombinator.com | 7 Mar 2024
It's Time for a Change: Datetime.utcnow() Is Now Deprecated
5 projects | news.ycombinator.com | 19 Nov 2023

"Local time" is time zone metadata. I've written a fair bit about timekeeping, because the context of what you're capturing becomes very important: https://github.com/kstenerud/concise-encoding/blob/master/ce...
RFC 3339 vs. ISO 8601
4 projects | news.ycombinator.com | 31 Aug 2023

This is basically why I ended up rolling my own text date format for Concise Encoding: https://github.com/kstenerud/concise-encoding/blob/master/ct...
ISO 8601 and RFC 3339 are fine for dates in the past, but they're not great as a general time format.
Ask HN: Please critique my metalanguage: “Dogma”
2 projects | news.ycombinator.com | 26 Feb 2023

This looks similar to https://concise-encoding.org/
Dogma was developed as a consequence of trying to describe Concise Binary Encoding. The CBE spec used to look like the preserves binary spec, full of hex values, tables and various ad-hoc illustrations: https://preserves.dev/preserves-binary.html
Now the CBE formal description looks like this: https://github.com/kstenerud/concise-encoding/blob/master/cb...
And the regular documentation looks like this: https://github.com/kstenerud/concise-encoding/blob/master/cb...
Dogma also does text formats (Concise Encoding has a text and binary format, so I needed a metalanguage that could do both in order to make it less jarring for a reader):
https://github.com/kstenerud/concise-encoding/blob/master/ct...
https://github.com/kstenerud/concise-encoding/blob/master/ct...
Concise Encoding Design Document
1 project | news.ycombinator.com | 7 Nov 2022
Keep ’Em Coming: Why Your First Ideas Aren’t Always the Best
2 projects | news.ycombinator.com | 7 Nov 2022

Hey thanks for taking the time to critique!
I actually do have an ANTLR file that is about 90% of the way there ( https://github.com/kstenerud/concise-encoding/tree/master/an... ), so I could use those as a basis...
One thing I'm not sure about is how to define a BNF rule that says for example: "An identifier is a series of characters from unicode categories Cf, L, M, N, and these specific symbol characters". BNF feels very ASCII-centric...
Working in the software industry, circa 1989 – Jim Grey
5 projects | news.ycombinator.com | 11 Jul 2022
It's still in the prerelease stage, but v1 will be released later this year. I'm mostly getting hits from China since they tend to be a lot more worried about security. I expect the rest of the world to catch on to the gaping security holes of JSON and friends in the next few years as the more sophisticated actors start taking advantage of them. For example https://github.com/kstenerud/concise-encoding/blob/master/ce...
There are still a few things to do:
- Update enctool (https://github.com/kstenerud/enctool) to integrate https://cuelang.org so that there's at least a command line schema validator for CE.
- Update the grammar file (https://github.com/kstenerud/concise-encoding/tree/master/an...) because it's a bit out of date.
- Revamp the compliance tests to be themselves written in Concise Encoding (for example https://github.com/kstenerud/go-concise-encoding/blob/master... but I'll be simplifying the format some more). That way, we can run the same tests on all CE implementations instead of everyone coming up with their own. I'll move the test definitions to their own repo when they're done and then you can just submodule it.
I'm thinking that they should look more like:
```
    c1
```
Breaking our Latin-1 assumptions
2 projects | news.ycombinator.com | 18 Jun 2022

Ugh Unicode has been the bane of my existence trying to write a text format spec. I started by trying to forbid certain characters to keep files editable and avoid Unicode rendering exploits (like hiding text, or making structured text behave differently than it looks), but in the end it became so much like herding cats that I had to just settle on https://github.com/kstenerud/concise-encoding/blob/master/ct...
Basically allow everything except some separators, most control chars, and some lookalike characters (which have to be updated as more characters are added to Unicode). It's not as clean as I'd like, but it's at least manageable this way.
I accidentally used YAML.parse instead of JSON.parse, and it worked?
8 projects | news.ycombinator.com | 23 Jan 2022

You might get a kick out of Concise Encoding then (shameless plug). It focuses on security and consistency of behavior.
https://concise-encoding.org/
In particular:
* How to deal with unrepresentable values: https://github.com/kstenerud/concise-encoding/blob/master/ce...
* Mandatory limits and security considerations: https://github.com/kstenerud/concise-encoding/blob/master/ce...
* Consistent error classification and processing: https://github.com/kstenerud/concise-encoding/blob/master/ce...
Ask HN: Who Wants to Collaborate?
58 projects | news.ycombinator.com | 1 Jan 2022

In the above example, `&a:` means mark the next object and give it symbolic identifier "a". `$a` means look up the reference to symbolic identifier "a". So this is a map whose "recusive link" key is a pointer to the map itself. How this data is represented internally by the receiver of such a document (a table, a struct, etc) is up to the implementation.
> - Time zones: ASN.1 supports ISO 8601 time types, including specification of local or UTC time.
Yes, this is the major failing of ISO 8601: They don't have true time zones. It only uses UTC offsets, which are a bad idea for so many reasons. https://github.com/kstenerud/concise-encoding/blob/master/ce...
> - Bin + txt: Again, I'm unclear on what you mean here, but ASN.1 has both binary and text-based encodings
Ah cool, didn't know about those.
> - Versioned: Also a little unclear to me
The intent is to specify the exact document formatting that the decoder can expect. For example we could in theory decide make CBE version 2 a bit-oriented format instead of byte-oriented in order to save space at the cost of processing time. It would be completely unreadable to a CBE 1 decoder, but since the document starts with 0x83 0x02 instead of 0x83 0x01, a CBE 1 decoder would say "I can't decode this" and a CBE 2 decoder would say "I can decode this".
With documents versioned to the spec, we can change even the fundamental structure of the format to deal with ANYTHING that might come up in future. Maybe a new security flaw in CBE 1 is discovered. Maybe a new data type becomes so popular that it would be crazy not to include it, etc. This avoids polluting the simpler encodings with deprecated types and bloating the format.

What are some alternatives?

When comparing kaldi-active-grammar and concise-encoding you can also consider the following projects:

silero-vad - Silero VAD: pre-trained enterprise-grade Voice Activity Detector

cue - The home of the CUE language! Validate and define text-based and dynamic configuration

nerd-dictation - Simple, hackable offline speech to text - using the VOSK-API.

joystick - A full-stack JavaScript framework for building stable, easy-to-maintain apps and websites.

pocketsphinx-python - Python interface to CMU Sphinxbase and Pocketsphinx libraries

postal-codes-json-xml-csv - Collection of postal codes in different formats, ready for importing.

mycroft-precise - A lightweight, simple-to-use, RNN wake word listener

futurecoder - 100% free and interactive Python course for beginners

Caster - Dragonfly-Based Voice Programming and Accessibility Toolkit

FrameworkBenchmarks - Source for the TechEmpower Framework Benchmarks project

dragonfly - Speech recognition framework allowing powerful Python-based scripting and extension of Dragon NaturallySpeaking (DNS), Windows Speech Recognition (WSR), Kaldi and CMU Pocket Sphinx

cue - CUE has moved to https://github.com/cue-lang/cue

kaldi-active-grammar vs silero-vad concise-encoding vs cue kaldi-active-grammar vs nerd-dictation concise-encoding vs joystick kaldi-active-grammar vs pocketsphinx-python concise-encoding vs postal-codes-json-xml-csv kaldi-active-grammar vs mycroft-precise concise-encoding vs futurecoder kaldi-active-grammar vs Caster concise-encoding vs FrameworkBenchmarks kaldi-active-grammar vs dragonfly concise-encoding vs cue

Compare kaldi-active-grammar vs concise-encoding and see what are their differences.

kaldi-active-grammar

concise-encoding

kaldi-active-grammar

concise-encoding

What are some alternatives?