llama-recipes
ollama
llama-recipes | ollama | |
---|---|---|
9 | 229 | |
10,065 | 72,781 | |
8.7% | 14.0% | |
9.8 | 9.9 | |
8 days ago | 4 days ago | |
Jupyter Notebook | Go | |
GNU General Public License v3.0 or later | MIT License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
llama-recipes
- Prompt Engineering with Llama2
-
Purple Llama by Meta AI
There are a whole bunch of prompts for this here: https://github.com/facebookresearch/llama-recipes/commit/109...
- [D] Recommendation for LLM fine-tuning codebase
- FLaNK Stack Weekly for 27 November 2023
- Finetune codellama for code completion task on specific programming language
-
[D] Tokenizers Truncation during Fine-tuning with Large Texts
Llama-recipes
-
How to fine tune llama2?
You can also try the recipe here https://github.com/facebookresearch/llama-recipes/blob/main/quickstart.ipynb
- Examples and recipes for Llama 2 model
- Llama Recipes
ollama
-
Ollama v0.1.45
I think the two main maintainers of Ollama have good intentions but suffer from a combination of being far too busy, juggling their forked llama.cpp server and not having enough automation/testing for PRs.
There is a new draft PR up to look at moving away from trying to juggle maintaining a llama.cpp fork to using llama.cpp with cgo bindings which I think will really help: https://github.com/ollama/ollama/pull/5034
-
SpringAI, llama3 and pgvector: bRAGging rights!
To support the exploration, I've developed a simple Retrieval Augmented Generation (RAG) workflow that works completely locally on the laptop for free. If you're interested, you can find the code itself here. Basically, I've used Testcontainers to create a Postgres database container with the pgvector extension to store text embeddings and an open source LLM with which I send requests to: Meta's llama3 through ollama.
-
RAG with OLLAMA
Note: Before proceeding further you need to download and run Ollama, you can do so by clicking here.
-
Ollama 0.1.42
`file://*` URLs are now allowed => ollama works with simple html files now
https://github.com/ollama/ollama/commit/1a29e9a879433fc55cf1...
-
How to setup a free, self-hosted AI model for use with VS Code
This guide assumes you have a supported NVIDIA GPU and have installed Ubuntu 22.04 on the machine that will host the ollama docker image. AMD is now supported with ollama but this guide does not cover this type of setup.
-
beginner guide to fully local RAG on entry-level machines
Nowadays, running powerful LLMs locally is ridiculously easy when using tools such as ollama. Just follow the installation instructions for your #OS. From now on, we'll assume using bash on Ubuntu.
- Codestral: Mistral's Code Model
- AIM Weekly 27 May 2024
-
Devoxx Genie Plugin : an Update
I focused on supporting Ollama, GPT4All, and LMStudio, all of which run smoothly on a Mac computer. Many of these tools are user-friendly wrappers around Llama.cpp, allowing easy model downloads and providing a REST interface to query the available models. Last week, I also added "👋🏼 Jan" support because HuggingFace has endorsed this provider out-of-the-box.
- Ask HN: Are companies self hosting LLMs?
What are some alternatives?
FLaNK-OpenAi - Chat
llama.cpp - LLM inference in C/C++
llm-toys - Small(7B and below) finetuned LLMs for a diverse set of useful tasks
gpt4all - gpt4all: run open-source LLMs anywhere
LLaVA - [NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
text-generation-webui - A Gradio web UI for Large Language Models. Supports transformers, GPTQ, AWQ, EXL2, llama.cpp (GGUF), Llama models.
CogVLM - a state-of-the-art-level open visual language model | 多模态预训练模型
private-gpt - Interact with your documents using the power of GPT, 100% privately, no data leaks
BakLLaVA
LocalAI - :robot: The free, Open Source OpenAI alternative. Self-hosted, community-driven and local-first. Drop-in replacement for OpenAI running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. It allows to generate Text, Audio, Video, Images. Also with voice cloning capabilities.
pymobiledevice3 - Pure python3 implementation for working with iDevices (iPhone, etc...).
llama - Inference code for Llama models