get-beam
store-sentry
get-beam | store-sentry | |
---|---|---|
9 | 2 | |
89 | 6 | |
- | - | |
7.9 | 4.7 | |
21 days ago | about 1 year ago | |
Shell | JavaScript | |
- | MIT License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
get-beam
-
Ask HN: Where to find an env with GPU for model training?
You should checkout https://beam.cloud (I'm the founder), it'll give you access to plenty of cloud GPU resources for training or inference.
Right now it's pretty hard to get GPU quota on AWS/GCP, so hopefully this is useful for you.
-
Cloudflare launches new AI tools to help customers deploy and run models
Cloudflare AI and Replicate are great for running off-the-shelf models, but anything custom is going to incur a 10+ minute cold start.
For running custom fine-tuned models on serverless, you could look into https://beam.cloud which is optimized for serving custom models with extremely fast cold start (I'm a little biased since I work there, but the numbers don't lie)
-
Workers AI: serverless GPU-powered inference on Cloudflare’s global network
Serverless only works if the cold boot is fast. For context, my company runs a serverless cloud GPU product called https://beam.cloud, which we've optimized for fast cold start. We see Whisper in production cold start in under 10s (across model sizes). A lot of our users are running semi-real time STT, and this seems to be working well for them.
-
Ultrafast serverless GPU runtime for custom SD models
I’m Eli, and my co-founder and I built Beam to run workloads on serverless cloud GPUs with hot reloading, autoscaling, and (of course) fast cold start. You don’t need Docker or AWS to use it, and everyone who signs up gets 10 hours of free GPU credit to try it out.
-
[D] We built Beam: An ultrafast serverless GPU runtime
Github with example apps and tutorials: https://github.com/slai-labs/get-beam/tree/main/examples
-
How to Finetune Llama 2: A Beginner's Guide
In this blog post, I want to make it as simple as possible to fine-tune the LLaMA 2 - 7B model, using as little code as possible. We will be using the Alpaca Lora Training script, which automates the process of fine-tuning the model and for GPU we will be using Beam.
- Run CodeLlama on a Serverless GPU
store-sentry
-
Cloudflare launches new AI tools to help customers deploy and run models
My entire App Store Server Notifications for iOS apps runs on Cloudflare Workers. I [open-sourced the code a while ago](https://github.com/workerforce/store-sentry) but it hasn’t gained much traction
- Store-Sentry: App Store Server Notification listening post
What are some alternatives?
discourse-ai
worker-template-postgres - Reference demo and modified PostgreSQL driver to connect Cloudflare Workers to a relational database.
whisper-turbo - Cross-Platform, GPU Accelerated Whisper 🏎️
serverless-dns - The RethinkDNS resolver that deploys to Cloudflare Workers, Deno Deploy, Fastly, and Fly.io
finetune-llama2
Google-Drive-Index - Index Google Drive Files Easily and Free
alpaca-lora - Instruct-tune LLaMA on consumer hardware
cloudflare-worker-router-template - A wrangler template for a super lightweight router (3.6 kB) with middleware support and ZERO dependencies for CloudFlare Workers, inspired by express.js syntax.