jarvis
ollama
jarvis | ollama | |
---|---|---|
1 | 229 | |
8 | 72,781 | |
- | 14.0% | |
5.0 | 9.9 | |
3 months ago | 4 days ago | |
PHP | Go | |
GNU General Public License v3.0 only | MIT License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
jarvis
-
PHP-powered OpenAI-assistant
I wanted to share a project I've been working on recently: Jarvis, a chatbot based on OpenAI API, with a particular focus on the new assistant-feature "function calling". It can easily run local code, including filesystem functions, sending emails, or calling APIs like DALL·E. It can eg. create whole websites with all files locally, even including (due to current api-limits) one generated image from dalle-3. Adding new local functions requires zero knowledge about python, symfony or the OpenAI-api, basic PHP knowledge is sufficient. The predefined skills include:- Filesystem operations (read, write, list, create) within the local container- Reading remote files via curl_get- Sending emails via SMTP- Creating images via Dalle-3- Accessing the Portainer APIIt is completely free(GPL), available on GitHub: https://github.com/Mugen0815/jarvis and DockerHub: https://hub.docker.com/r/mugen0815/jarvis
ollama
-
Ollama v0.1.45
I think the two main maintainers of Ollama have good intentions but suffer from a combination of being far too busy, juggling their forked llama.cpp server and not having enough automation/testing for PRs.
There is a new draft PR up to look at moving away from trying to juggle maintaining a llama.cpp fork to using llama.cpp with cgo bindings which I think will really help: https://github.com/ollama/ollama/pull/5034
-
SpringAI, llama3 and pgvector: bRAGging rights!
To support the exploration, I've developed a simple Retrieval Augmented Generation (RAG) workflow that works completely locally on the laptop for free. If you're interested, you can find the code itself here. Basically, I've used Testcontainers to create a Postgres database container with the pgvector extension to store text embeddings and an open source LLM with which I send requests to: Meta's llama3 through ollama.
-
RAG with OLLAMA
Note: Before proceeding further you need to download and run Ollama, you can do so by clicking here.
-
Ollama 0.1.42
`file://*` URLs are now allowed => ollama works with simple html files now
https://github.com/ollama/ollama/commit/1a29e9a879433fc55cf1...
-
How to setup a free, self-hosted AI model for use with VS Code
This guide assumes you have a supported NVIDIA GPU and have installed Ubuntu 22.04 on the machine that will host the ollama docker image. AMD is now supported with ollama but this guide does not cover this type of setup.
-
beginner guide to fully local RAG on entry-level machines
Nowadays, running powerful LLMs locally is ridiculously easy when using tools such as ollama. Just follow the installation instructions for your #OS. From now on, we'll assume using bash on Ubuntu.
- Codestral: Mistral's Code Model
- AIM Weekly 27 May 2024
-
Devoxx Genie Plugin : an Update
I focused on supporting Ollama, GPT4All, and LMStudio, all of which run smoothly on a Mac computer. Many of these tools are user-friendly wrappers around Llama.cpp, allowing easy model downloads and providing a REST interface to query the available models. Last week, I also added "👋🏼 Jan" support because HuggingFace has endorsed this provider out-of-the-box.
- Ask HN: Are companies self hosting LLMs?
What are some alternatives?
NexaAIOne - Arm you with all the essential tools to integrate AI seamlessly into your apps, regardless of the coding language you're comfortable
llama.cpp - LLM inference in C/C++
phpjelly - PHP Jelly
gpt4all - gpt4all: run open-source LLMs anywhere
text-generation-webui - A Gradio web UI for Large Language Models. Supports transformers, GPTQ, AWQ, EXL2, llama.cpp (GGUF), Llama models.
private-gpt - Interact with your documents using the power of GPT, 100% privately, no data leaks
LocalAI - :robot: The free, Open Source OpenAI alternative. Self-hosted, community-driven and local-first. Drop-in replacement for OpenAI running on consumer-grade hardware. No GPU required. Runs gguf, transformers, diffusers and many more models architectures. It allows to generate Text, Audio, Video, Images. Also with voice cloning capabilities.
llama - Inference code for Llama models
koboldcpp - A simple one-file way to run various GGML and GGUF models with KoboldAI's UI
exllama - A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.
text-generation-inference - Large Language Model Text Generation Inference
litellm - Call all LLM APIs using the OpenAI format. Use Bedrock, Azure, OpenAI, Cohere, Anthropic, Ollama, Sagemaker, HuggingFace, Replicate (100+ LLMs)