buzz
HandBrake
buzz | HandBrake | |
---|---|---|
21 | 981 | |
9,869 | 15,593 | |
- | 2.0% | |
8.5 | 9.8 | |
20 days ago | about 6 hours ago | |
Python | C | |
MIT License | GNU General Public License v3.0 or later |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
buzz
- Buzz: Transcribe and translate audio offline on your personal computer
- MacWhisper: Transcribe audio files on your Mac
-
Build Personal ChatGPT Using Your Data
Easiest 1-click way to install and use Stable Diffusion on your computer."
https://github.com/easydiffusion/easydiffusion
And while Whisper is OpenAI, it is trivial to use locally and extremely usefull
https://github.com/chidiwilliams/buzz
- automated transcription software that is HIPAA compliant?
- Question: Does anyone know of an AI or ChatGPT tool to create automatic SRT caption files by uploading a video?
- Brauchbare Speech-to-Text Lösungen für Windows?
-
I've pretty much had it with Premiere.
Install this for Resolve
-
As a Foreign student, I record letures alot so I can review it anytime. Thanks to Obsidian Audio Player its feels so effortless to review those records.
Just use this one: https://github.com/chidiwilliams/buzz
-
Whispers AI Modular Future
What utilities related to Whisper do you wish existed? What have you had to build yourself?
On the end user application side, I wish there was something that let me pick a podcast of my choosing, get it fully transcribed, and get an embeddings search plus answer q&a on top of that podcast or set of chosen podcasts. I've seen ones for specific podcasts, but I'd like one where I can choose the podcast. (Probably won't build it)
Also on the end user side, I wish there was an Otter alternative (still paid $30/mo, but unlimited minutes per month) that had longer transcription limits. (Started building this, not much interest from users though)
Things I've seen on the dev tool side:
Gladia (API call version of Whisper)
Whisper.cpp
Whisper webservice (https://github.com/ahmetoner/whisper-asr-webservice) - via this thread
Live microphone demo (not real time, it still does it in chunks) https://github.com/mallorbc/whisper_mic
Streamlit UI https://github.com/hayabhay/whisper-ui
Whisper playground https://github.com/saharmor/whisper-playground
Real time whisper https://github.com/shirayu/whispering
Whisper as a service https://github.com/schibsted/WAAS
Improved timestamps and speaker identification https://github.com/m-bain/whisperX
MacWhisper https://goodsnooze.gumroad.com/l/macwhisper
Crossplatform desktop Whisper that supports semi-realtime https://github.com/chidiwilliams/buzz
- Any suggestions for easy ways to add subtitles to YouTube videos?
HandBrake
- FFmpeg 7.0 Released
-
On Export, Left audio is delayed by about half a second, is there a fix?
https://handbrake.fr/ its very easy to use.
-
Linux GUI/Frontend for VMAF
nothing actually, i just mentioned it as an encoding frontend.
- Windows Photos saves really large video files
-
Vegas PRO18, my mov files.
At this point, use https://handbrake.fr/ to transcode into a higher res MP4 and use as proxy footage,
-
converting .avi files to .mp4
Maybe you could try HandBrake.
-
HandBrake 1.7.0 – The open source video transcoder
The release itself[0] would have been a better link.
0. https://github.com/HandBrake/HandBrake/releases/tag/1.7.0
- HandBrake 1.7.0 released with AMD & Nvidia AV1 hwenc, SVT-AV1 v1.7 & SVT-AV1 multi-pass ABR
- HandBrake 1.7.0 released
- HandBrake 1.7.0 Released
What are some alternatives?
whisper - Robust Speech Recognition via Large-Scale Weak Supervision
Tdarr - Tdarr - Distributed transcode automation using FFmpeg/HandBrake + Audio/Video library analytics + video health checking (Windows, macOS, Linux & Docker)
openai-whisper-cpu - Improving transcription performance of OpenAI Whisper for CPU based deployment
staxrip - 🎞 Video encoding GUI for Windows.
StoryToolkitAI - An editing tool that uses AI to transcribe, understand content and search for anything in your footage, integrated with ChatGPT and other AI models
Av1an - Cross-platform command-line AV1 / VP9 / HEVC / H264 encoding framework with per scene quality encoding
audapolis - an editor for spoken-word audio with automatic transcription
FFmpeg - Mirror of https://git.ffmpeg.org/ffmpeg.git
whisper-diarization - Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
obs-studio - OBS Studio - Free and open source software for live streaming and screen recording
text-to-speech-ubuntu - 🙊 Setup "selectable" text to speech / TTS on Ubuntu Linux 24.04 22.04 22.10 23.04 23.10 . Ideal for speed reading, programming, editing and writing.
lossless-cut - The swiss army knife of lossless video/audio editing