NeMo-Guardrails vs basaran

NeMo-Guardrails

NeMo Guardrails is an open-source toolkit for easily adding programmable guardrails to LLM-based conversational systems. (by NVIDIA)

Suggest topics

Source Code

Suggest alternative

Edit details

basaran

Basaran is an open-source alternative to the OpenAI text completion API. It provides a compatible streaming API for your Hugging Face Transformers-based text generation models. (by hyperonym)

generative-model Gpt huggingface language-model Natural Language Processing openai-api streaming-api text-generation chatgpt

DISCONTINUED

Suggest alternative

Edit details

Our great sponsors

WorkOS - The modern identity platform for B2B SaaS

InfluxDB - Power Real-Time Data Analytics at Scale

SaaSHub - Software Alternatives and Reviews

Our great sponsors

NeMo-Guardrails		basaran
	Project
13	Mentions	22
3,338	Stars	1,281
7.9%	Growth	-
9.9	Activity	10.0
6 days ago	Latest Commit	3 months ago
Python	Language	Python
GNU General Public License v3.0 or later	License	MIT License

The number of mentions indicates the total number of mentions that we've tracked plus the number of user suggested alternatives.
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.

NeMo-Guardrails

Posts with mentions or reviews of NeMo-Guardrails. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-06-07.

NeMO Guardrails from Nvidia
1 project | news.ycombinator.com | 4 Sep 2023
Run and create custom ChatGPT-like bots with OpenChat
15 projects | news.ycombinator.com | 7 Jun 2023

- https://github.com/NVIDIA/NeMo-Guardrails/
LangChain: The Missing Manual
5 projects | news.ycombinator.com | 19 May 2023
The Dual LLM pattern for building AI assistants that can resist prompt injection
2 projects | news.ycombinator.com | 14 May 2023

Here's "jailbreak detection", in the NeMo-Guardrails project from Nvidia:
https://github.com/NVIDIA/NeMo-Guardrails/blob/327da8a42d5f8...
I.e. they ask the llm if the prompt will break the llm. (I believe that more data /some evaluation on how well this performs is intended to be released. Probably fair to call this stuff "not battle tested".)
How To Setup a Model With Guardrails?
3 projects | /r/LocalLLaMA | 12 May 2023

I have been playing around with some models locally and creating a discord bot as a fun side project, and I wanted to setup some guardrails on inputs / outputs of the bot to make sure that it isn't violating any ethical boundaries. I was going to use Nvidia's Nemo guardrails, but they only support openai currently. Are there any other good ways to control inputs?
RasaGPT: First headless LLM chatbot built on top of Rasa, Langchain and FastAPI
13 projects | news.ycombinator.com | 8 May 2023

Thanks, I hadn't seen those. I did find https://github.com/NVIDIA/NeMo-Guardrails earlier but haven't looked into it yet.
I'm not sure it solves the problem of restricting the information it uses though. For example, as a proof of concept for a customer, I tried providing information from a vector database as context, but GPT would still answer questions that were not provided in that context. It would base its answers on information that was already crawled from the customer website and in the model. That is concerning because the website might get updated but you can't update the model yourself (among other reasons).
How do we prevent prompt injection in a GPT API app?
1 project | /r/OpenAI | 7 May 2023
Nvidia NeMo Guardrails – open-source guardrails to conversational systems
1 project | /r/CKsTechNews | 1 May 2023

1 project | news.ycombinator.com | 30 Apr 2023
Should LangChain be used in Prod?
2 projects | /r/LangChain | 27 Apr 2023

you can use guard rails with langchain - https://github.com/NVIDIA/NeMo-Guardrails

basaran

Posts with mentions or reviews of basaran. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-06-19.

OpenLLM
10 projects | news.ycombinator.com | 19 Jun 2023
Langchain and self hosted LLaMA hosted API
6 projects | /r/LocalLLaMA | 15 Jun 2023

What are the current best "no reinventing the wheel" approaches to have Langchain use an LLM through a locally hosted REST API, the likes of Oobabooga or hyperonym/basaran with streaming support for 4-bit GPTQ?
Run and create custom ChatGPT-like bots with OpenChat
15 projects | news.ycombinator.com | 7 Jun 2023

Disclaimer: I am curating LLM-tools on github [1]
A few thoughts:
* allow for custom endpoint URLs, this way people can use open source LLMs with a fake openAI API backend like basaran[2] or llama-api-server[3]
* look into better embedding methods for info-retrieval like InstructorEmbeddings or Document Summary Index
* Don't use a single embedding per content item, use multiple to increase retrieval quality
1 https://github.com/underlines/awesome-marketing-datascience/...
2 https://github.com/hyperonym/basaran
3 https://github.com/iaalm/llama-api-server
1-Jun-2023
2 projects | /r/dailyainews | 2 Jun 2023

open-source alternative to the OpenAI text completion API (https://github.com/hyperonym/basaran)
Introducing Basaran: self-hosted open-source alternative to the OpenAI text completion API
9 projects | /r/LocalLLaMA | 1 Jun 2023
Basaran is an open-source alternative to the OpenAI text completion API
1 project | news.ycombinator.com | 31 May 2023
Ask HN: What's the best self hosted/local alternative to GPT-4?
12 projects | news.ycombinator.com | 31 May 2023

Guanaco-65B[0] using Basaran[1] for your OpenAI compatible API. You can use any ChatGPT front-end which lets you change the OpenAI endpoint URL.
[0] An fp4 finetune of LLaMA-30B by Tim Dettmers
[1] https://github.com/hyperonym/basaran
Are all the finetunes stupid?
5 projects | /r/LocalLLaMA | 22 Apr 2023

For lm-eval, I think you'd either need to take GPTQ's inference script and shim it into a model: https://github.com/EleutherAI/lm-evaluation-harness/tree/master/lm_eval/models or you might be able to use a project like https://github.com/hyperonym/basaran and then you could use the gpt3 model...
Using the API in Node
3 projects | /r/Oobabooga | 11 Apr 2023

There are also: - Basaran repo: "Basaran is an open-source alternative to the OpenAI text completion API. It provides a compatible streaming API for your Hugging Face Transformers-based text generation models". "...Compatibility with OpenAI API and client libraries..."; - llama-cpp-python repo: "Simple Python bindings for @ggerganov's llama.cpp library...". "...OpenAI-like API...".
Researcher looking for help with how to prepare a finetuning dataset for models like Bloomz and Cerebras-GPT
2 projects | /r/ArtificialInteligence | 2 Apr 2023

I want to start with a totally freely available model, so again, that excludes things like LLaMA where the weights are only available through a wait list. The two models that most get my attention and (I think, and hope) fit my criteria of open availability are Cerebras-GPT (13b) and Bloomz (7b). The tools to process and fine-tune that seem most feasible to me, from my limit knowledge, are xturing and basaran.

What are some alternatives?

When comparing NeMo-Guardrails and basaran you can also consider the following projects:

guidance - A guidance language for controlling large language models. [Moved to: https://github.com/guidance-ai/guidance]

text-generation-inference - Large Language Model Text Generation Inference

langchainrb - Build LLM-powered applications in Ruby

openai-chatgpt-opentranslator - Python command that uses openai to perform text translations

guidance - A guidance language for controlling large language models.

AutoGPTQ - An easy-to-use LLMs quantization package with user-friendly apis, based on GPTQ algorithm.

lmql - A language for constraint-guided and efficient LLM programming.

llm-foundry - LLM training code for Databricks foundation models

text-generation-webui - A Gradio web UI for Large Language Models. Supports transformers, GPTQ, AWQ, EXL2, llama.cpp (GGUF), Llama models.

alpaca.cpp - Locally run an Instruction-Tuned Chat-Style LLM

pgvector - Open-source vector similarity search for Postgres

NeMo-Guardrails vs guidance basaran vs text-generation-inference NeMo-Guardrails vs langchainrb basaran vs openai-chatgpt-opentranslator NeMo-Guardrails vs guidance basaran vs AutoGPTQ NeMo-Guardrails vs lmql basaran vs llm-foundry NeMo-Guardrails vs text-generation-webui basaran vs alpaca.cpp NeMo-Guardrails vs pgvector basaran vs lmql

Compare NeMo-Guardrails vs basaran and see what are their differences.

NeMo-Guardrails

basaran

NeMo-Guardrails

basaran

What are some alternatives?