MIOpen
mlc-llm
MIOpen | mlc-llm | |
---|---|---|
9 | 89 | |
983 | 17,053 | |
1.4% | 3.7% | |
9.7 | 9.9 | |
5 days ago | 3 days ago | |
Assembly | Python | |
MIT License | Apache License 2.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
MIOpen
- AI Libraries and AI Frameworks are "Not Available" for ROCm on Windows. Does that mean not yet, or never?
-
[Project] MLC LLM: Universal LLM Deployment with GPU Acceleration
More than three months behind schedule...
- Someone has run SD with a release candidate of ROCm 5.5 on RDNA 3 and gets 15 it/s
-
ROCm on a 7900 XTX
https://github.com/ROCmSoftwarePlatform/MIOpen/milestones Miopen for rocm 5.5 and 5.6 milestones are done.But you don't release those yet.I still don't understand, what is the point of selling AI/ML capable GPU without releasing the driver for that (rocm, etc).
- Sapphire Pulse 7900 xt (Techpowerup review)
- Man I wish I could do all this cool shit too
- Issues with Automatic1111 WebUI on Ubuntu 22.04.1 LTS with AMD GPU
- Miopen - AMD's Machine Intelligence Library
- Radeon ROCm 4.3 Released With HMM Allocations, Many Other Improvements
mlc-llm
- FLaNK 04 March 2024
-
Ai on a android phone?
This one uses gpu, it doesn't support Mistral yet: https://github.com/mlc-ai/mlc-llm
-
MLC vs llama.cpp
I have tried running mistral 7B with MLC on my m1 metal. And it kept crushing (git issue with description). Memory inefficiency problems.
-
[Project] Scaling LLama2 70B with Multi NVIDIA and AMD GPUs under 3k budget
Project: https://github.com/mlc-ai/mlc-llm
- Scaling LLama2-70B with Multi Nvidia/AMD GPU
-
AMD May Get Across the CUDA Moat
For LLM inference, a shoutout to MLC LLM, which runs LLM models on basically any API that's widely available: https://github.com/mlc-ai/mlc-llm
-
ROCm Is AMD's #1 Priority, Executive Says
One of your problems might be that gfx1032 is not supported by AMD's ROCm packages, which has a laughably short list of supported hardware: https://rocm.docs.amd.com/en/latest/release/gpu_os_support.h...
The normal workaround is to assign the closest architecture, eg gfx1030, so `HSA_OVERRIDE_GFX_VERSION=10.3.0` might help
Also, it looks like some of your tested projects are OpenCL? For me, I do something like: `yay -S rocm-hip-sdk rocm-ml-sdk rocm-opencl-sdk` to cover all the bases.
My recent interest has been LLMs and this is my general step by step for those (llama.cpp, exllama) for those interested: https://llm-tracker.info/books/howto-guides/page/amd-gpus
I didn't port the docs back in, but also here's a step-by-step w/ my adventures getting TVM/MLC working w/ an APU: https://github.com/mlc-ai/mlc-llm/issues/787
From my experience, ROCm is improving, but there's a good reason that Nvidia has 90% market share even at big price premiums.
-
Show HN: Ollama for Linux – Run LLMs on Linux with GPU Acceleration
Maybe they're talking about https://github.com/mlc-ai/mlc-llm which is used for web-llm (https://github.com/mlc-ai/web-llm)? Seems to be using TVM.
-
Show HN: Fine-tune your own Llama 2 to replace GPT-3.5/4
you already have TVM for the cross platform stuff
see https://tvm.apache.org/docs/how_to/deploy/android.html
or https://octoml.ai/blog/using-swift-and-apache-tvm-to-develop...
or https://github.com/mlc-ai/mlc-llm
- Ask HN: Are you training and running custom LLMs and how are you doing it?
What are some alternatives?
ROCm - AMD ROCm™ Software - GitHub Home [Moved to: https://github.com/ROCm/ROCm]
llama.cpp - LLM inference in C/C++
k-diffusion-directml - Karras et al. (2022) diffusion models for PyTorch
ggml - Tensor library for machine learning
stablediffusion-directml - High-Resolution Image Synthesis with Latent Diffusion Models
tvm - Open deep learning compiler stack for cpu, gpu and specialized accelerators
SillyTavern - LLM Frontend for Power Users. [Moved to: https://github.com/SillyTavern/SillyTavern]
text-generation-webui - A Gradio web UI for Large Language Models. Supports transformers, GPTQ, AWQ, EXL2, llama.cpp (GGUF), Llama models.
stable-diffusion-webui - Stable Diffusion web UI
llama-cpp-python - Python bindings for llama.cpp
fast-stable-diffusion - fast-stable-diffusion + DreamBooth
ollama - Get up and running with Llama 3, Mistral, Gemma, and other large language models.