ai-notes
WAAS
ai-notes | WAAS | |
---|---|---|
15 | 12 | |
4,554 | 1,738 | |
- | 1.2% | |
9.8 | 7.0 | |
8 days ago | 5 days ago | |
HTML | JavaScript | |
MIT License | Apache License 2.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
ai-notes
-
Minimal implementation of Mamba, the new LLM architecture, in 1 file of PyTorch
the field just moves fast. I have curated a list of non-hypey writers and youtubers who explain these things for a typical SWE audience if you are interested. https://github.com/swyxio/ai-notes/blob/main/Resources/Good%...
- SDXL Turbo: A Real-Time Text-to-Image Generation Model
-
DeepEval – Unit Testing for LLMs
added to my notes! https://github.com/swyxio/ai-notes/
- ChatGPT Code Interpreter Capabilities
-
Google just released a 100% free learning path on Generative AI with 9 Courses
and here are mine, organized by beginner/intermediate/advanced
https://github.com/swyxio/ai-notes/blob/main/README.md#top-a...
and then you can go into the individual modality specific notes for more reading
- Show HN: Self-host Whisper As a Service with GUI and queueing
-
Show HN: YouTube Summaries Using GPT
there's https://learnprompting.org/
i've also been keeping a popular series of notes https://github.com/sw-yx/ai-notes/blob/main/TEXT_PROMPTS.md
-
Show HN: I reverse prompt engineered every Notion AI feature
Direct link to the source prompts are here: https://github.com/sw-yx/ai-notes/blob/main/Resources/Notion...
- GitHub - sw-yx/prompt-eng: notes for prompt engineering
- My hand-curated list of major distros and forks of Stable Diffusion. Please suggest anything I missed!
WAAS
-
Show HN: Minutes – Save up to 20% of salespeople time
This app does it locally on newer macs for free.
https://apps.apple.com/no/app/jojo-transcribe/id1659864300?m...
And open source: https://github.com/schibsted/WAAS
-
Only I don't use docker. WAAS. What alternatives to a docker install are there?
Clone the repo, then run the steps in the dockerfile: https://github.com/schibsted/WAAS/blob/main/Dockerfile
-
Whispers AI Modular Future
What utilities related to Whisper do you wish existed? What have you had to build yourself?
On the end user application side, I wish there was something that let me pick a podcast of my choosing, get it fully transcribed, and get an embeddings search plus answer q&a on top of that podcast or set of chosen podcasts. I've seen ones for specific podcasts, but I'd like one where I can choose the podcast. (Probably won't build it)
Also on the end user side, I wish there was an Otter alternative (still paid $30/mo, but unlimited minutes per month) that had longer transcription limits. (Started building this, not much interest from users though)
Things I've seen on the dev tool side:
Gladia (API call version of Whisper)
Whisper.cpp
Whisper webservice (https://github.com/ahmetoner/whisper-asr-webservice) - via this thread
Live microphone demo (not real time, it still does it in chunks) https://github.com/mallorbc/whisper_mic
Streamlit UI https://github.com/hayabhay/whisper-ui
Whisper playground https://github.com/saharmor/whisper-playground
Real time whisper https://github.com/shirayu/whispering
Whisper as a service https://github.com/schibsted/WAAS
Improved timestamps and speaker identification https://github.com/m-bain/whisperX
MacWhisper https://goodsnooze.gumroad.com/l/macwhisper
Crossplatform desktop Whisper that supports semi-realtime https://github.com/chidiwilliams/buzz
- [task] Write out the questions and responses from my podcast videos so I can turn it into a written article for my site - $5 per video x 15 videos = $75 - $80
- Self-host Whisper As a Service with GUI and queueing. Schibsted created a transcription service for our journalists to transcribe audio interviews and podcasts really quick.
- Show HN: Self-host Whisper As a Service with GUI and queueing
What are some alternatives?
text2image-gui - Somewhat modular text2image GUI, initially just for Stable Diffusion
whisper - Robust Speech Recognition via Large-Scale Weak Supervision
diffusionbee-stable-diffusion-ui - Diffusion Bee is the easiest way to run Stable Diffusion locally on your M1 Mac. Comes with a one-click installer. No dependencies or technical knowledge needed.
whisper.cpp - Port of OpenAI's Whisper model in C/C++
m1_huggingface_diffusers_demo - Demo of how to get HuggingFace Diffusers working on an M1 Mac
whisper_mic - Project that allows one to use a microphone with OpenAI whisper.
stable-diffusion-ui - Easiest 1-click way to install and use Stable Diffusion on your computer. Provides a browser UI for generating images from text prompts and images. Just enter your text prompt, and see the generated image. [Moved to: https://github.com/easydiffusion/easydiffusion]
whisper-playground - Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/
perceiver-pytorch - Implementation of Perceiver, General Perception with Iterative Attention, in Pytorch
openai-whisper-cpu - Improving transcription performance of OpenAI Whisper for CPU based deployment
stable-diffusion - A latent text-to-image diffusion model
whisper-asr-webservice - OpenAI Whisper ASR Webservice API