CogVLM Alternatives

Similar projects and alternatives to CogVLM

llama.cpp

769 56,891 10.0 C++ CogVLM VS llama.cpp

LLM inference in C/C++
ollama

195 58,943 9.9 Go CogVLM VS ollama

Get up and running with Llama 3, Mistral, Gemma, and other large language models.
InfluxDB

www.influxdata.com sponsored

Power Real-Time Data Analytics at Scale. Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.
surge

170 2,913 9.5 C CogVLM VS surge

Synthesizer plug-in (previously released as Vember Audio Surge)
ComfyUI

125 33,352 9.9 Python CogVLM VS ComfyUI

The most powerful and modular stable diffusion GUI, api and backend with a graph/nodes interface.
FLiPStackWeekly

80 14 9.9 CogVLM VS FLiPStackWeekly

FLaNK AI Weekly covering Apache NiFi, Apache Flink, Apache Kafka, Apache Spark, Apache Iceberg, Apache Ozone, Apache Pulsar, and more...
llama_index

75 30,910 10.0 Python CogVLM VS llama_index

LlamaIndex is a data framework for your LLM applications
LLaVA

20 16,101 9.4 Python CogVLM VS LLaVA

[NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
WorkOS

workos.com sponsored

The modern identity platform for B2B SaaS. The APIs are flexible and easy-to-use, supporting authentication, user identity, and complex enterprise features like SSO and SCIM provisioning.
OpenAdapt

18 419 9.5 Python CogVLM VS OpenAdapt

AI-First Process Automation with Large [Language (LLMs) / Action (LAMs) / Multimodal (LMMs)] / Visual Language (VLMs)) Models
ollama-webui

14 5,789 9.8 Svelte CogVLM VS ollama-webui

Discontinued ChatGPT-Style WebUI for LLMs (Formerly Ollama WebUI) [Moved to: https://github.com/open-webui/open-webui]
FLaNK-Halifax

14 1 7.3 TypeScript CogVLM VS FLaNK-Halifax

Community over Code, Apache NiFi, Apache Kafka, Apache Flink, Python, GTFS, Transit, Open Source, Open Data
vimGPT

6 2,426 8.1 Python CogVLM VS vimGPT

Browse the web with GPT-4V and Vimium
krita-ai-diffusion

11 4,533 9.8 Python CogVLM VS krita-ai-diffusion

Streamlined interface for generating images with AI in Krita. Inpaint and outpaint with optional text prompt, no tweaking required.
Qwen-VL

4 3,667 8.7 Python CogVLM VS Qwen-VL

The official repo of Qwen-VL (通义千问-VL) chat & pretrained large vision language model proposed by Alibaba Cloud.
FLaNK-SaoPauloBrazil

10 2 6.8 HTML CogVLM VS FLaNK-SaoPauloBrazil

FLaNK-SaoPauloBrazil
llama-recipes

9 9,194 9.8 Jupyter Notebook CogVLM VS llama-recipes

Scripts for fine-tuning Meta Llama3 with composable FSDP & PEFT methods to cover single/multi-node GPUs. Supports default & custom datasets for applications such as summarization and Q&A. Supporting a number of candid inference solutions such as HF TGI, VLLM for local or cloud deployment. Demo apps to showcase Meta Llama3 for WhatsApp & Messenger.
vectorflow

9 635 8.4 Python CogVLM VS vectorflow

VectorFlow is a high volume vector embedding pipeline that ingests raw data, transforms it into vectors and writes it to a vector DB of your choice. (by dgarnitz)
vlite

7 686 6.2 Python CogVLM VS vlite

fast vector database made in numpy
awesome-api-security

5 2,730 6.6 CogVLM VS awesome-api-security

A collection of awesome API Security tools and resources. The focus goes to open-source tools and resources that benefit all the community.
uform

8 865 8.2 Python CogVLM VS uform

Pocket-Sized Multimodal AI for content understanding and generation across multilingual texts, images, and 🔜 video, up to 5x faster than OpenAI CLIP and LLaVA 🖼️ & 🖋️
LinkBERT

2 389 1.8 Python CogVLM VS LinkBERT

[ACL 2022] LinkBERT: A Knowledgeable Language Model 😎 Pretrained with Document Links
SaaSHub

www.saashub.com sponsored

SaaSHub - Software Alternatives and Reviews. SaaSHub helps you find the best software and product alternatives

NOTE: The number of mentions on this list indicates mentions on common posts plus user suggested alternatives. Hence, a higher number means a better CogVLM alternative or higher similarity.

Suggest an alternative to CogVLM

CogVLM reviews and mentions

Posts with mentions or reviews of CogVLM. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2024-01-08.

Mixtral: Mixture of Experts
8 projects | news.ycombinator.com | 8 Jan 2024

CogVLM is very good in my (brief) testing: https://github.com/THUDM/CogVLM
The model weights seem to be under a non-commercial license, not true open source, but it is "open access" as you requested.
IT Employment Grew by Just 700 Jobs in 2023, Down From 267,000 in 2022
2 projects | news.ycombinator.com | 8 Jan 2024

increasing growth most places in world
https://twitter.com/elonmusk/status/1743028102446408026
heres a total feature map of what was released in 2023:
https://twitter.com/enriquebrgn/status/1740950767325024387
I think thats definitely a signal that the B and C teams werent needed, considering they cut 90% of staff LOL.
As for the bots, AI is making it easier than ever to bypass those systems. CogVLM is just sitting there menacingly on github https://github.com/THUDM/CogVLM
Show HN: I built an open source AI video search engine to learn more about AI
2 projects | news.ycombinator.com | 19 Dec 2023
CogAgent-18B – visual-based GUI Agent capabilities
2 projects | news.ycombinator.com | 16 Dec 2023

Jump to heading for benchmarks and examples: https://github.com/THUDM/CogVLM/tree/main?tab=readme-ov-file...
What do you think. When should we expect the next SDXL version?
1 project | /r/StableDiffusion | 10 Dec 2023

Honestly at this point there is no need for human for captioning except maybe for NSFW content. Img2text is just good enough for nearly all images. GPTVision or open source equivalent (like CogVLM https://github.com/THUDM/CogVLM ) are just good enough.
shinning the spotlight on CogVLM
3 projects | /r/LocalLLaMA | 9 Dec 2023

A core Llama.cpp contributor, named cmp-nct, discovered stumbled upon what might be the next leap forward for vision/language models. CogVLM (which uses a Vicuna 7B language model combined with a 9B vision tower) excels particularly in OCR (Optical Character Recognition), detail detection, and minimal hallucinations. It effectively understands both handwritten and typed text, context, fine details, and background graphics. It even provides pixel coordinates for small visual targets. CovVLM surpasses other models like llava-1.5 and Qwen-VL in performance.
Image-to-Caption Generator
3 projects | /r/computervision | 7 Dec 2023

https://github.com/THUDM/CogVLM (really impressive)
Gemini: Google's most capable AI model yet
2 projects | news.ycombinator.com | 6 Dec 2023

I'm researching using LLMs for alt-text suggestion for forum users, can you share your finding so far?
Outside of GPT-4V I had good first results with https://github.com/THUDM/CogVLM
Open-source LLMs with Image Interpretation
1 project | /r/LocalLLaMA | 6 Dec 2023

I've got some decent results with CogVLM. Resolution kinda sucks at 490x490, though.
FLaNK Stack Weekly for 27 November 2023
28 projects | dev.to | 27 Nov 2023
A note from our sponsor - WorkOS
workos.com | 29 Apr 2024

The APIs are flexible and easy-to-use, supporting authentication, user identity, and complex enterprise features like SSO and SCIM provisioning. Learn more →