distil-whisper
cog
distil-whisper | cog | |
---|---|---|
9 | 20 | |
3,225 | 7,224 | |
7.0% | 3.7% | |
8.9 | 9.4 | |
15 days ago | 4 days ago | |
Python | Python | |
MIT License | Apache License 2.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
distil-whisper
- FLaNK Stack 05 Feb 2024
-
Distil-Whisper: a distilled variant of Whisper that is 6x faster
Training code will be released in the Distil-Whisper repository this week, enabling anyone in the community to distill a Whisper model in their choice of language!
- FLaNK Stack Weekly 06 Nov 2023
-
AI — weekly megathread!
Hugging Face released Distil-Whisper, a distilled version of Whisper that is 6 times faster, 49% smaller, and performs within 1% word error rate (WER) on out-of-distribution evaluation sets [Details].
- Distil-Whisper: distilled version of Whisper that is 6 times faster, 49% smaller
- Distil-Whisper is up to 6x faster than Whisper while performing within 1% Word-Error-Rate on out-of-distribution eval sets
-
Distilling Whisper on 20,000 hours of open-sourced audio data
- GitHub page: https://github.com/huggingface/distil-whisper/tree/main
-
Talk-Llama
Is https://github.com/huggingface/distil-whisper on its way to whisper.cpp?
cog
-
AI Grant Traction in OSS Startups
View on GitHub
- Insanely Fast Whisper: Transcribe 300 minutes of audio in less than 98 seconds
-
Talk-Llama
I'm in the same situation. I found this cog project to dockerise ML https://github.com/replicate/cog : you write just one python class and a yaml file, and it takes care of the "CUDA hell" and deps. It even creates a flask app in front of your model.
That helps keep your system clean, but someone with big $s please rewrite pytorch to golang or rust or even nodejs / typescript.
-
Llama 2 – Meta AI
https://github.com/replicate/cog
Our thinking was just that a bunch of folks will want to fine-tune right away, then deploy the fine-tunes, so trying to make that easy... Or even just deploy the models-as-is on their own infra without dealing with CUDA insanity!
-
Handling concurrent requests to ML model API
I have used this tool before: https://github.com/replicate/cog/tree/main
-
Opinions on Cog: Containers for machine learning
Then I discovered Cog: Containers for Machine Learning. Looks like a way more flexible solution to plug in the existing infrastructure: you write your custom code and Cog plugs it in a Docker image with FastAPI, no extra ecosystem complexity added.
-
can someone teach me how to install the new stable diffusion repo?
Highly recommend using cog https://github.com/replicate/cog
- Run Stable Diffusion on Your M1 Mac’s GPU
- replicate/cog: Containers for machine learning
-
Why companies move off Heroku (besides the cost)
Dokku Maintainer here.
Dokku also supports Dockerfiles, Docker Images, Tarballs (similar to heroku slugs), and Cloud Native Buildpacks. I'm also actively working on AWS Lambda support (both for simple usage without much config as well as SAM-based usage) and investigating Replicate's Cog[1] and Railways Nixpacks[2] functionalities for building apps.
There are quite a few options in the OSS space (as well as Commercial offerings from new startups and popular incumbents). It's an interesting space to be in, and its always fun to see how new offerings innovate on existing solutions.
[1] https://github.com/replicate/cog
What are some alternatives?
WhisperInput - Offline voice input panel & keyboard with punctuation for Android.
nixpacks - App source + Nix packages + Docker = Image
pyvideotrans - Translate the video from one language to another and add dubbing. 将视频从一种语言翻译为另一种语言,并添加配音
pytorch_wavelets - Pytorch implementation of 2D Discrete Wavelet (DWT) and Dual Tree Complex Wavelet Transforms (DTCWT) and a DTCWT based ScatterNet
vid2cleantxt - Python API & command-line tool to easily transcribe speech-based video files into clean text
piku - The tiniest PaaS you've ever seen. Piku allows you to do git push deployments to your own servers.
faster-whisper - Faster Whisper transcription with CTranslate2
heroku-review-app-actions - GitHub action to automate managing review apps on your Heroku account
json-masker - High-performance JSON masker library in Java with no runtime dependencies
tvm - Open deep learning compiler stack for cpu, gpu and specialized accelerators
willow - Open source, local, and self-hosted Amazon Echo/Google Home competitive Voice Assistant alternative
memray - Memray is a memory profiler for Python