one-click-installers
llama-mps
one-click-installers | llama-mps | |
---|---|---|
18 | 4 | |
470 | 83 | |
- | - | |
8.9 | 3.8 | |
8 months ago | 8 months ago | |
Python | Python | |
GNU Affero General Public License v3.0 | GNU General Public License v3.0 only |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
one-click-installers
-
amd gpus on windows support?
AMD does not offer installation options for ROCm on Windows. I'm not familiar with the workarounds to make it work; if you find a solution, you can contribute it to https://github.com/oobabooga/one-click-installers/
-
Oobabooga for Windows
Running start_windows.bat should take care of everything.
-
Quant-Cude Error?
Had the same issue, turns out I was using an old 1 click installer / updater, you need to use https://github.com/oobabooga/one-click-installers and reinstall everything from scratch
-
Cant find the "start: file.
Are you sure you're looking at the right folder? start_windows.bat is there. It's listed in the source code: https://github.com/oobabooga/one-click-installers
- Any UI that allows Windows + AMD GPU ?
- WizardLM-30B-Uncensored
-
13b-4bit-128g - Trying to run compressed model without success. ( problem exist only with 13b models for some reason ) No error code has been displayed.
one-click-installers/INSTRUCTIONS.TXT
-
GPT4All: A little helper to get started
https://github.com/oobabooga/one-click-installers/issues/56 they explain it over here.
-
Visual Studio compile errors
I solved this by adding the Individual components 2019 Windows 10 SDK, C++ CMake tools for Windows, and MSVC v142 - VS 2019 C++ build tools. See https://github.com/oobabooga/one-click-installers/issues/56
-
python setup.py bdist_wheel did not run successfully.
It appears one of the extensions isn't pre-compiled on install. I believe you have the same problem as listed here. https://github.com/oobabooga/one-click-installers/issues/56
llama-mps
-
llama.cpp now officially supports GPU acceleration.
There are currently at least 3 ways to run llama on m1 with GPU acceleration. - mlc-llm (pre-built, only 1 model has been ported) - tinygrad (very memory efficient, not that easy to integrate into other projects) - llama-mps (original llama codebase + llama adapter support)
-
LLaMA-7B in Pure C++ with full Apple Silicon support
There is also a gpu-acelerated fork of the original repo
https://github.com/remixer-dec/llama-mps
- Llama-CPU: Fork of Facebooks LLaMa model to run on CPU
-
[D] Tutorial: Run LLaMA on 8gb vram on windows (thanks to bitsandbytes 8bit quantization)
I tried to port the llama-cpu version to a gpu-accelerated mps version for macs, it runs, but the outputs are not as good as expected and it often gives "-1" tokens. Any help and contributions on fixing it are welcome!
What are some alternatives?
GPTQ-for-LLaMa - 4 bits quantization of LLaMa using GPTQ
llama - Inference code for Llama models
gpt4all - gpt4all: run open-source LLMs anywhere
text-generation-webui - A Gradio web UI for Large Language Models. Supports transformers, GPTQ, AWQ, EXL2, llama.cpp (GGUF), Llama models.
gradio - Build and share delightful machine learning apps, all in Python. 🌟 Star to support our work!
awesome-ml - Curated list of useful LLM / Analytics / Datascience resources
KoboldAI
llama - Inference code for LLaMA models
WizardVicunaLM - LLM that combines the principles of wizardLM and vicunaLM
LLaMA_MPS - Run LLaMA inference on Apple Silicon GPUs.
micromamba-releases - Micromamba executables mirrored from conda-forge as Github releases
tinygrad - You like pytorch? You like micrograd? You love tinygrad! ❤️