StableLM | AGIEval | |
---|---|---|
43 | 3 | |
15,853 | 649 | |
0.2% | 4.8% | |
5.0 | 5.2 | |
about 1 month ago | about 1 month ago | |
Jupyter Notebook | Python | |
Apache License 2.0 | MIT License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
StableLM
-
The Era of 1-bit LLMs: ternary parameters for cost-effective computing
https://github.com/Stability-AI/StableLM?tab=readme-ov-file#...
-
Stable LM 3B: Bringing Sustainable, High-Performance LMs to Smart Devices
https://mistral.ai/news/announcing-mistral-7b/
looking at the 3b results (here https://github.com/Stability-AI/StableLM#stablelm-alpha-v2 ?), it looks like Mistral (which outperforms Llama-2 13b) is far more powerful
-
FreeWilly 1 and 2, two new open-access LLMs
Does this mean Stability gave up on StableLM?
I notice that the repo hasn’t been updated since April, and a question asking for an update has been ignored for at least a month: https://github.com/Stability-AI/StableLM/issues/83
-
In five years, there will be no programmers left, believes Stability AI CEO
I'm not "ignoring" StableLM, if anything it's the impetus for my post. The alpha models were so bad and unusable that it seems they may have simply abandoned the project. It's clear they basically didn't know what they were doing, which is silly for a company of their size and specialization.
-
Losing the plot
1) StableLM released a checkpoint at 800B for their 3B and 7B at 800B tokens with 4096 context size, but perform very poorly on different benchmarks and finetuning is discouraged with such a weak base model
-
UAE's Technology Innovation Institute Launches Open-Source "Falcon 40B" Large Language Model for Research & Commercial Utilization
It is the best open-source model currently available. Falcon-40B outperforms LLaMA, StableLM, RedPajama, MPT, etc. See the OpenLLM Leaderboard.
- Consulta API GPT
- Google "We Have No Moat, And Neither Does OpenAI"
-
New to StableLM--is it possible to use this locally to fine-tune on a small subset of documents yet?
Someone shared this link on another recent post
-
[N] Stability AI releases StableVicuna: the world's first open source chatbot trained via RLHF
Github: https://github.com/Stability-AI/StableLM
AGIEval
- AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models
-
Google DeepMind CEO Says Some Form of AGI Possible in a Few Years
Microsoft defined a set of benchmarks in AGIEval.
-
[N] Stability AI announce their open-source language model, StableLM
AGIEval is an attempt at this... https://i.imgur.com/6RL9XNB.png
What are some alternatives?
text-generation-webui - A Gradio web UI for Large Language Models. Supports transformers, GPTQ, AWQ, EXL2, llama.cpp (GGUF), Llama models.
RedPajama-Data - The RedPajama-Data repository contains code for preparing large datasets for training large language models.
lm-evaluation-harness - A framework for few-shot evaluation of language models.
llama.cpp - LLM inference in C/C++
ggml - Tensor library for machine learning
LLM-Adapters - Code for our EMNLP 2023 Paper: "LLM-Adapters: An Adapter Family for Parameter-Efficient Fine-Tuning of Large Language Models"
Open-Assistant - OpenAssistant is a chat-based assistant that understands tasks, can interact with third-party systems, and retrieve information dynamically to do so.
alpaca_lora_4bit
llama - Inference code for Llama models
KoboldAI-Client