instructor-embedding VS llama_index

Compare instructor-embedding vs llama_index and see what are their differences.

InfluxDB - Power Real-Time Data Analytics at Scale
Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.
www.influxdata.com
featured
SaaSHub - Software Alternatives and Reviews
SaaSHub helps you find the best software and product alternatives
www.saashub.com
featured
instructor-embedding llama_index
4 75
1,703 31,184
3.1% 4.7%
5.9 10.0
10 days ago 4 days ago
Python Python
Apache License 2.0 MIT License
The number of mentions indicates the total number of mentions that we've tracked plus the number of user suggested alternatives.
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.

instructor-embedding

Posts with mentions or reviews of instructor-embedding. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-07-10.
  • My experience on starting with fine tuning LLMs with custom data
    7 projects | /r/LocalLLaMA | 10 Jul 2023
    If you li embeddings and vector DB, you should look into this: https://github.com/HKUNLP/instructor-embedding
  • Build Personal ChatGPT Using Your Data
    14 projects | news.ycombinator.com | 8 Jul 2023
    If you look at a embeddings leaderboard [1], one of the top competitors called InstructorXL [2] is just a pip install away. It's neck and neck with Ada v2 except for a shorter input length and half the dimensions, with the added benefit that you'll always have the model available.

    Most of the other options just work with the transformers library.

    [1] https://huggingface.co/spaces/mteb/leaderboard

    [2] https://github.com/HKUNLP/instructor-embedding

  • I've made a customisable SMS personal assistant which has infinite and persistent semantic memory.
    2 projects | /r/LocalLLaMA | 27 May 2023
    Use instructor-embedding to to make it 100% local and even maybe quick relationship lookup (embed relationship info with sentiment analysis instruction)
  • Whisper Transcription Formatting
    1 project | /r/artificial | 3 Feb 2023
    First.I believe having srt subtitles as whisper result would be better.Essentially you don't need just a list of words like YouTube does.You need something more structured.I don't remember what whisper outputs so I might be wrong.There is whisperx for that as example. And then maybe use gpt index over it.Or something like instructor model That can work.

llama_index

Posts with mentions or reviews of llama_index. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2024-04-01.

What are some alternatives?

When comparing instructor-embedding and llama_index you can also consider the following projects:

h2ogpt - Private chat with local GPT with document, images, video, etc. 100% private, Apache 2.0. Supports oLLaMa, Mixtral, llama.cpp, and more. Demo: https://gpt.h2o.ai/ https://codellama.h2o.ai/

langchain - ⚡ Building applications with LLMs through composability ⚡ [Moved to: https://github.com/langchain-ai/langchain]

openai-cookbook - Examples and guides for using the OpenAI API

langchain - 🦜🔗 Build context-aware reasoning applications

Nuggt - An Autonomous LLM Agent that runs on Wizcoder-15B

private-gpt - Interact with your documents using the power of GPT, 100% privately, no data leaks

vlite - fast vector database made in numpy

chatgpt-retrieval-plugin - The ChatGPT Retrieval Plugin lets you easily find personal or work documents by asking questions in natural language.

easydiffusion - Easiest 1-click way to create beautiful artwork on your PC using AI, with no tech knowledge. Provides a browser UI for generating images from text prompts and images. Just enter your text prompt, and see the generated image.

text-generation-webui - A Gradio web UI for Large Language Models. Supports transformers, GPTQ, AWQ, EXL2, llama.cpp (GGUF), Llama models.

lit-gpt - Hackable implementation of state-of-the-art open-source LLMs based on nanoGPT. Supports flash attention, 4-bit and 8-bit quantization, LoRA and LLaMA-Adapter fine-tuning, pre-training. Apache 2.0-licensed. [Moved to: https://github.com/Lightning-AI/litgpt]

gpt-llama.cpp - A llama.cpp drop-in replacement for OpenAI's GPT endpoints, allowing GPT-powered apps to run off local llama.cpp models instead of OpenAI.