pyllama
serge
pyllama | serge | |
---|---|---|
5 | 40 | |
2,789 | 5,553 | |
- | 0.9% | |
5.0 | 9.8 | |
6 months ago | 3 days ago | |
Python | Svelte | |
GNU General Public License v3.0 only | Apache License 2.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
pyllama
-
Tech pioneers call for six-month pause of "out-of-control" AI development
Got u fam
-
Alpaca 7B Training - $75/Hour --> Bay Area? [P]
I am the author of pyllama https://github.com/juncongmoo/pyllama
- Integrate LLaMA into python code
- Together Releases The First Open-Source ChatGPT Alternative Called OpenChatKit
- pyllama - I just published a python library for LLaMA with Single GPU inference code
serge
- Show HN: I made an app to use local AI as daily driver
- chatgpt alternative
-
Show HN: LlamaGPT – Self-hosted, offline, private AI chatbot, powered by Llama 2
Very cool, this looks like a combination of chatbot-ui and llama-cpp-python? A similar project I've been using is https://github.com/serge-chat/serge. Nous-Hermes-Llama2-13b is my daily driver and scores high on coding evaluations (https://huggingface.co/spaces/mike-ravkine/can-ai-code-resul...).
-
LeCun: Qualcomm working with Meta to run Llama-2 on mobile devices
You might be pleased to hear that nothing really stops you from doing this today. If you ran Serge[0] on a Mac with Tailscale, you could hack together a decently-accelerated Llama chatbot.
[0] https://github.com/serge-chat/serge
-
Chatbot frontend library in Svelte?
Cannot help you with libraries specifically but both Serge and ChatUI are built using SvelteKit, so the code might be of some use to you.
- We’re back and…
-
Best way to use AMD CPU and GPU
Serge made it really easy for me to get started, but it all CPU-based.
-
Need Help
All that said this project probably solves your problem: https://github.com/serge-chat/serge
- Are you selfhosting a ChatGPT alternative?
-
What the hell??
You can play a little bit with more straightforward local models (the simplest to setup is https://github.com/nsarrazin/serge ), to see that any LLM is basically a party trick.
What are some alternatives?
playground - Play with neural networks!
gpt4all - gpt4all: run open-source LLMs anywhere
semantic-kernel - Integrate cutting-edge LLM technology quickly and easily into your apps
langflow - ⛓️ Langflow is a dynamic graph where each node is an executable unit. Its modular and interactive design fosters rapid experimentation and prototyping, pushing hard on the limits of creativity.
clip-interrogator - Image to prompt with BLIP and CLIP
llama.cpp - LLM inference in C/C++
tortoise-tts-fast - Fast TorToiSe inference (5x or your money back!)
text-generation-webui - A Gradio web UI for Large Language Models. Supports transformers, GPTQ, AWQ, EXL2, llama.cpp (GGUF), Llama models.
llm - An ecosystem of Rust libraries for working with large language models
FastChat - An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
GPTQ-for-LLaMa - 4 bits quantization of LLaMA using GPTQ
llama-gpt - A self-hosted, offline, ChatGPT-like chatbot. Powered by Llama 2. 100% private, with no data leaving your device. New: Code Llama support!