obsidian-audio-player
buzz
obsidian-audio-player | buzz | |
---|---|---|
3 | 21 | |
89 | 9,937 | |
- | - | |
2.1 | 8.5 | |
6 months ago | 24 days ago | |
JavaScript | Python | |
GNU General Public License v3.0 only | MIT License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
obsidian-audio-player
-
How can I do voice to text notes that go right into obsidian?
I have local code in Python I've played with for something like this, using OpenAI Whisper. I use the segment timestamps with the audio player plugin's bookmarks feature in order to manually (but fairly quickly) correct or check apparent errors. I put it aside a few weeks ago but have been meaning to get back to it but you'd need to follow command line instructions for Python to install and run it, I have no intention of building a UI beyond generating Markdown (which could include MOCs).
-
As a Foreign student, I record letures alot so I can review it anytime. Thanks to Obsidian Audio Player its feels so effortless to review those records.
M4A was sorted out in this pull request https://github.com/noonesimg/obsidian-audio-player/pull/8 Looks like it's already supported, you'd just need to add the file extension
buzz
- Buzz: Transcribe and translate audio offline on your personal computer
- MacWhisper: Transcribe audio files on your Mac
-
Build Personal ChatGPT Using Your Data
Easiest 1-click way to install and use Stable Diffusion on your computer."
https://github.com/easydiffusion/easydiffusion
And while Whisper is OpenAI, it is trivial to use locally and extremely usefull
https://github.com/chidiwilliams/buzz
- automated transcription software that is HIPAA compliant?
- Question: Does anyone know of an AI or ChatGPT tool to create automatic SRT caption files by uploading a video?
- Brauchbare Speech-to-Text Lösungen für Windows?
-
I've pretty much had it with Premiere.
Install this for Resolve
-
As a Foreign student, I record letures alot so I can review it anytime. Thanks to Obsidian Audio Player its feels so effortless to review those records.
Just use this one: https://github.com/chidiwilliams/buzz
-
Whispers AI Modular Future
What utilities related to Whisper do you wish existed? What have you had to build yourself?
On the end user application side, I wish there was something that let me pick a podcast of my choosing, get it fully transcribed, and get an embeddings search plus answer q&a on top of that podcast or set of chosen podcasts. I've seen ones for specific podcasts, but I'd like one where I can choose the podcast. (Probably won't build it)
Also on the end user side, I wish there was an Otter alternative (still paid $30/mo, but unlimited minutes per month) that had longer transcription limits. (Started building this, not much interest from users though)
Things I've seen on the dev tool side:
Gladia (API call version of Whisper)
Whisper.cpp
Whisper webservice (https://github.com/ahmetoner/whisper-asr-webservice) - via this thread
Live microphone demo (not real time, it still does it in chunks) https://github.com/mallorbc/whisper_mic
Streamlit UI https://github.com/hayabhay/whisper-ui
Whisper playground https://github.com/saharmor/whisper-playground
Real time whisper https://github.com/shirayu/whispering
Whisper as a service https://github.com/schibsted/WAAS
Improved timestamps and speaker identification https://github.com/m-bain/whisperX
MacWhisper https://goodsnooze.gumroad.com/l/macwhisper
Crossplatform desktop Whisper that supports semi-realtime https://github.com/chidiwilliams/buzz
- Any suggestions for easy ways to add subtitles to YouTube videos?
What are some alternatives?
obsidian-audio-notes - Easily take notes on podcasts and other audio files using Obsidian Audio Notes.
whisper - Robust Speech Recognition via Large-Scale Weak Supervision
bookmark_plugin - A better alternative to Chrome Bookmarks for Obsidian users. Customize templates on the fly and either append to existing notes, create new notes, or do both!
openai-whisper-cpu - Improving transcription performance of OpenAI Whisper for CPU based deployment
obsidian-system-dark-mode - Automatically use the operating system's setting to switch between light and dark mode.
StoryToolkitAI - An editing tool that uses AI to transcribe, understand content and search for anything in your footage, integrated with ChatGPT and other AI models
audapolis - an editor for spoken-word audio with automatic transcription
whisper-diarization - Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
text-to-speech-ubuntu - 🙊 Setup "selectable" text to speech / TTS on Ubuntu Linux 24.04 22.04 22.10 23.04 23.10 . Ideal for speed reading, programming, editing and writing.
opentts - Open Text to Speech Server
JCRequest - Another Guzzle wrapper
HTTPlug - HTTPlug, the HTTP client abstraction for PHP