LiteratureForEyesAndEars vs WhisperLive

LiteratureForEyesAndEars

By Yorwba

Suggest topics

Source Code

Suggest alternative

Edit details

WhisperLive

A nearly-live implementation of OpenAI's Whisper. (by collabora)

dictation obs openai text-to-speech Translation voice-recognition Whisper

Source Code

Suggest alternative

Edit details

Scout Monitoring - Free Django app performance insights with Scout Monitoring

Get Scout setup in minutes, and let us sweat the small stuff. A couple lines in settings.py is all you need to start monitoring your apps. Sign up for our free tier today.

www.scoutapm.com

featured

InfluxDB - Power Real-Time Data Analytics at Scale

Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.

www.influxdata.com

featured

LiteratureForEyesAndEars		WhisperLive
	Project
1	Mentions	4
0	Stars	1,350
-	Growth	12.6%
4.7	Activity	9.4
5 months ago	Latest Commit	7 days ago
Python	Language	Python
-	License	MIT License

The number of mentions indicates the total number of mentions that we've tracked plus the number of user suggested alternatives.
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.

LiteratureForEyesAndEars

Posts with mentions or reviews of LiteratureForEyesAndEars. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2024-01-17.

WhisperSpeech – An Open Source text-to-speech system built by inverting Whisper
9 projects | news.ycombinator.com | 17 Jan 2024

I have forced alignments, too.
E.g. for the True Story of Ah Q https://github.com/Yorwba/LiteratureForEyesAndEars/tree/mast... .align.json is my homegrown alignment format, .srt are standard subtitles, .txt is the text, but note that in some places I have [[original text||what it is pronounced as]] annotations to make the forced alignment work better. (E.g. the "." in LibriVox.org, pronounced as 點 "diǎn" in Mandarin.) Oh, and cmn-Hans is the same thing transliterated into Simplified Chinese.
The corresponding LibriVox URL is predictably https://librivox.org/the-true-story-of-ah-q-by-xun-lu/

WhisperLive

Posts with mentions or reviews of WhisperLive. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2024-01-29.

Show HN: WhisperFusion – Ultra-low latency conversations with an AI chatbot
7 projects | news.ycombinator.com | 29 Jan 2024

Everything runs locally, we use:
- WhisperLive for the transcription - https://github.com/collabora/WhisperLive
WhisperSpeech – An Open Source text-to-speech system built by inverting Whisper
9 projects | news.ycombinator.com | 17 Jan 2024

Check out WhisperLive: https://github.com/collabora/WhisperLive
If you're grappling with the slow march from cool tech demos to real-world language model apps, you might wanna check out WhisperLive. It's this rad open-source project that’s all about leveraging Whisper models for slick live transcription. Think real-time, on-the-fly translated captions for those global meetups. It's a neat example of practical, user-focused tech in action. Dive into the details on their GitHub page
Whisper: Nvidia RTX 4090 vs. M1 Pro with MLX
10 projects | news.ycombinator.com | 13 Dec 2023

https://github.com/collabora/WhisperLive
The is another one that uses huggingface's implementation, but I haven't tried it since my spec doesn't support flash-att2
Triple Threat: The Power of Transcription, Summary, and Translation
1 project | news.ycombinator.com | 3 Aug 2023

Curious to see how this works? Check out our demo page - https://col.la/transcription to generate your own transcription, summary, and translation, or use our browser extension - https://github.com/collabora/WhisperLive to get live transcriptions.

What are some alternatives?

When comparing LiteratureForEyesAndEars and WhisperLive you can also consider the following projects:

cog-whisper-diarization - Cog implementation of transcribing + diarization pipeline with Whisper & Pyannote

whisper-writer - 💬📝 A small dictation app using OpenAI's Whisper speech recognition model.

obs-zoom-and-follow - Dynamic zoom and mouse tracking script for OBS Studio

gpt_chatbot - This chatbot lets you use your microphone to communicate with GPT-4. It uses the OpenAI text to speech to respond with a voice. It uses Pinecone to store long term information and retrieves it to create context. API keys for OpenAI and Pinecone required. Tested on Windows

whisper_streaming - Whisper realtime streaming for long speech-to-text transcription and translation

gpt-voice-conversation-chatbot - Allows you to have an engaging and safely emotive spoken / CLI conversation with the AI ChatGPT / GPT-4 while giving you the option to let it remember things discussed.

WhisperFusion - WhisperFusion builds upon the capabilities of WhisperLive and WhisperSpeech to provide a seamless conversations with an AI.

PaddleSpeech - Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.

LiveWhisper - A nearly-live implementation of OpenAI's Whisper, using sounddevice. Requires existing Whisper install.

mlx - MLX: An array framework for Apple silicon

spokestack-python - Spokestack is a library that allows a user to easily incorporate a voice interface into any Python application with a focus on embedded systems.

espnet - End-to-End Speech Processing Toolkit

WhisperLive vs cog-whisper-diarization WhisperLive vs whisper-writer WhisperLive vs obs-zoom-and-follow WhisperLive vs gpt_chatbot WhisperLive vs whisper_streaming WhisperLive vs gpt-voice-conversation-chatbot WhisperLive vs WhisperFusion WhisperLive vs PaddleSpeech WhisperLive vs LiveWhisper WhisperLive vs mlx WhisperLive vs spokestack-python WhisperLive vs espnet

Scout Monitoring - Free Django app performance insights with Scout Monitoring

Get Scout setup in minutes, and let us sweat the small stuff. A couple lines in settings.py is all you need to start monitoring your apps. Sign up for our free tier today.

www.scoutapm.com

featured

InfluxDB - Power Real-Time Data Analytics at Scale

Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.

www.influxdata.com

featured