Weaviate
Toshi
Our great sponsors
Weaviate | Toshi | |
---|---|---|
76 | 12 | |
9,524 | 4,117 | |
5.7% | 0.6% | |
10.0 | 6.1 | |
about 18 hours ago | 3 months ago | |
Go | Rust | |
BSD 3-clause "New" or "Revised" License | MIT License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
Weaviate
-
pgvecto.rs alternatives - qdrant and Weaviate
3 projects | 13 Mar 2024
- FLaNK Stack 29 Jan 2024
- Qdrant, the Vector Search Database, raised $28M in a Series A round
-
How to use Weaviate to store and query vector embeddings
In this tutorial, I introduce Weaviate, an open-source vector database, with the thenlper/gte-base embedding model from Alibaba, through Hugging Face's transformers library.
-
Choosing vector database: a side-by-side comparison
This will be solved in Weaviate https://github.com/weaviate/weaviate/issues/2424
-
Who's hiring developer advocates? (October 2023)
Link to GitHub -->
-
Do we think about vector dbs wrong?
Hey @rvrs, I work on Weaviate and we are doing some improvements around increasing write throughput:
1. gRPC. Using gRPC to write vectors has had a really nice performance boost. It is released in Weaviate core but here is still some work on do on the clients. Feel free to get in contact if you would like to try it out.
2. Parameter tuning. lowering `efConstruction` can speed up imports.
3. We are also working on async indexing https://github.com/weaviate/weaviate/issues/3463 which will further speed things up.
In comparison with pgvector, Weaviate has more flexible query options such as hybrid search and quantization to save memory on larger datasets.
- Weaviate vector database
- Weaviate 1.21: Support for ImageBind and GPT4all and more
- Weaviate Vector Database
Toshi
-
Tantivy 0.20 is released: Schemaless column store, Schemaless aggregations, Phrase prefix queries, Percentiles, and more...
I don't think you have an active project that addresses all those use cases. There was an attempt in Rust with Toshi that is built on top of tantivy, but the project seems to have stalled.
- An alternative to Elasticsearch that runs on a few MBs of RAM
-
Postgres Full Text Search vs. the Rest
I wish we had an extension like ZomboDB but using a lighter search engine like https://github.com/quickwit-oss/quickwit, https://github.com/toshi-search/Toshi and https://github.com/mosuka/bayard
Here I'm listing engines based on https://github.com/quickwit-oss/tantivy - tantivy is comparable to Lucene in its scope - but I'm sure there are other engines that could tackle ElasticSearch.
Another thing that could happen is maybe directly embed tantivy in Postgres using an extension, perhaps this could be an option too.
-
Ask HN: Does anybody still use bookmarking services?
I do something similar, though I index the page myself via a little browser extension I wrote. I click a button, the content gets POSTed to a server that throws it in Toshi[1]. I hacked it together on a Saturday, and it's been pretty handy; as you describe, much more useful than any bookmarking approach I've tried before.
[1] https://github.com/toshi-search/Toshi
-
*set Edge as default browser*
There is some incredible work being done in the web department, frameworks like rocket.rs and actix.rs are amazing. To get the latest info on web development in Rust, check arewewebyet.org. It doesn't list Toshi though, which is weird.
- Zinc Search engine. A lightweight alternative to elasticsearch that requires minimal resources, written in Go.
- Zinc Search engine. A lightweight alternative to Elasticsearch written in Go
- AWS releases forked Elasticsearch code. Announces new name: OpenSearc
What are some alternatives?
Milvus - A cloud-native vector database, storage for next generation AI applications
elasticsearch-rs - Official Elasticsearch Rust Client
faiss - A library for efficient similarity search and clustering of dense vectors.
MeiliSearch - A lightning-fast search API that fits effortlessly into your apps, websites, and workflow
pgvector - Open-source vector similarity search for Postgres
narg - A tool to generate LC/AP formulas for a given seed in Noita.
qdrant - Qdrant - High-performance, massive-scale Vector Database for the next generation of AI. Also available in the cloud https://cloud.qdrant.io/
sonic - 🦔 Fast, lightweight & schema-less search backend. An alternative to Elasticsearch that runs on a few MBs of RAM.
jina - ☁️ Build multimodal AI applications with cloud-native stack
lnx - ⚡ Insanely fast, 🌟 Feature-rich searching. lnx is the adaptable, typo tollerant deployment of the tantivy search engine.
vald - Vald. A Highly Scalable Distributed Vector Search Engine
OpenSearch - 🔎 Open source distributed and RESTful search engine.