gorilla-cli
unstructured
gorilla-cli | unstructured | |
---|---|---|
11 | 12 | |
1,162 | 6,515 | |
3.8% | 15.6% | |
5.5 | 9.8 | |
3 months ago | 7 days ago | |
Python | HTML | |
Apache License 2.0 | Apache License 2.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
gorilla-cli
- FLaNK 15 Jan 2024
-
Show HN: Shell-AI, run shell commands with natural language
Hello HN! I know this project is a super simple wrapper around LangChain/OpenAI but I just found myself wanting this badly myself: a super simple `pip install` package that I can use to get command suggestions within the terminal as I'm being productive doing other things.
The implementation is literally one short glue of LangChain and InquirerPy for interactive CLI.
I'm curious which ideas you all have to make this smarter/better. MIT licensed, if you're keen on contributing please feel free to do so. It's a pure hobby project for me.
Some key objectives: never automatically run shell code, I want to see what I run before I run it, present me with some alternatives, a simple path to using local models in the future (Llama 2 Code soon?).
Will add I was inspired by the great https://github.com/gorilla-llm/gorilla-cli project, but didn't like that it sent the prompt to some IP based endpoint.
-
Show HN: Poozle – open-source Plaid for LLMs
Very cool product! Have you consider relying on Gorilla for integrations?
https://github.com/gorilla-llm/gorilla-cli
- FLaNK Stack Weekly for 07August2023
- Show HN: Lemon AI – open-source Zapier NLA to empower agents
- GitHub - gorilla-llm/gorilla-cli: LLMs for your CLI (cum să faci operations doar în limba engleză)
-
30-Jun-2023
gorilla-cli: LLMs for your CLI (https://github.com/gorilla-llm/gorilla-cli)
- Gorilla-CLI: LLMs for CLI including K8s/AWS/GCP/Azure/sed and 1500 APIs
unstructured
-
LlamaCloud and LlamaParse
Be careful with unstructured:
https://github.com/Unstructured-IO/unstructured/blob/d11c70c...
from: https://github.com/open-webui/open-webui/issues/687
- FLaNK 15 Jan 2024
-
Bash One-Liners for LLMs
I’ve been looking at this
https://freeling-user-manual.readthedocs.io/en/v4.2/modules/...
at the freeling library in general, also spaCy and NLTK. The chunking algorithms being used in the likes of LangChain are remarkably bad surprisingly.
There is also
https://github.com/Unstructured-IO/unstructured
But I don’t like it, can’t explain why yet.
My intuition is that 1st step is clean sentences and paragraphs and titles/labels/headers. Then probably an LLM can handle outlining and table of contents generation using a stripped down list of objects in the text.
BRIO/BERT summarization could also have a role of some type.
Those are my ideas so far.
- Unstructured – OSS libraries and APIs to build custom preprocessing pipelines
-
More intelligent Pdf parsers
Unstructured is the best one I’ve used so far: https://www.unstructured.io
- Help extracting data from multiple PDF's
- Pre-processing text documents such as PDFs, HTML and Word Documents for LLMs
-
Using ChatGPT to read multiple PDFs and create writing using them as sources
https://www.unstructured.io/ can parse PDFs, then you can feed all of them to Claude, which has a 100k context window.
-
How can I convert restaurant’s traditional menu in pdf file to well structured list of menu items with prices in Excel file? Thank you
If the copy & pase method does not work: One approach is to use the functionality of Unstructured to parse the PDF. If need be, it can do OCR on the PDF too if you have Detectron2 installed. After conversion you would still have to save it as an excel file though.
-
PDF GPT allows you to chat with the contents of your PDF file
I would check out https://github.com/Unstructured-IO/unstructured (what lang chain uses) or https://github.com/axa-group/Parsr (probably what unstructured copied to get their startup off the ground lol)
What are some alternatives?
GPTCache - Semantic cache for LLMs. Fully integrated with LangChain and llama_index.
llmsherpa - Developer APIs to Accelerate LLM Projects
FLiPStackWeekly - FLaNK AI Weekly covering Apache NiFi, Apache Flink, Apache Kafka, Apache Spark, Apache Iceberg, Apache Ozone, Apache Pulsar, and more...
Parsr - Transforms PDF, Documents and Images into Enriched Structured Data
shell_gpt - A command-line productivity tool powered by AI large language models like GPT-4, will help you accomplish your tasks faster and more efficiently.
ragflow - RAGFlow is an open-source RAG (Retrieval-Augmented Generation) engine based on deep document understanding.
Transformers-Tutorials - This repository contains demos I made with the Transformers library by HuggingFace.
pdfGPT - PDF GPT allows you to chat with the contents of your PDF file by using GPT capabilities. The most effective open source solution to turn your pdf files in a chatbot!
CallCMLModel - An example on calling models deployed in CML
awesome-document-understanding - A curated list of resources for Document Understanding (DU) topic
examples - Analyze the unstructured data with Towhee, such as reverse image search, reverse video search, audio classification, question and answer systems, molecular search, etc.
llama_parse - Parse files for optimal RAG