sru
allennlp
sru | allennlp | |
---|---|---|
1 | 13 | |
2,098 | 11,337 | |
0.0% | - | |
0.0 | 8.4 | |
over 2 years ago | over 1 year ago | |
Python | Python | |
MIT License | Apache License 2.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
sru
-
[D] Language Models for Smaller Languages
This one? Looks like it's easy to use, which is nice!
allennlp
-
How to solve ConfigurationError using HuggingFace Token Classifier
No clue. So what I did was google the error. Here's what I found: https://github.com/allenai/allennlp/issues/4319
- AllenNLP will be unmaintained in December
- AllenNLP Is EOL
- Any recommendation for the replacement of the toolkit jiant? [Research] [Discussion]
- Cedille, the largest French language model, open source with a freely accessible playground
-
[P] Cedille, the largest French language model (6b), released in open source
Another aspect we had fun with is dataset filtering. We have run the whole C4 French dataset through the Detoxify classifier to clean it up 🤬
-
Any allennlp users in this sub?
https://github.com/allenai/allennlp/discussions looks active
- Multilingual C4 (mC4) Dataset now released
- C4 dataset released (800GB Common Crawl-derived text; T5 training data)
What are some alternatives?
text - Models, data loaders and abstractions for language processing, powered by PyTorch
cedille-ai - ✒️ Cedille is a large French language model (6B), released under an open-source license
best-of-ml-python - 🏆 A ranked list of awesome machine learning Python libraries. Updated weekly.
fairseq - Facebook AI Research Sequence-to-Sequence Toolkit written in Python.
attention-is-all-you-need-pytorch - A PyTorch implementation of the Transformer model in "Attention is All You Need".
mesh-transformer-jax - Model parallel transformers in JAX and Haiku
datasets - 🤗 The largest hub of ready-to-use datasets for ML models with fast, easy-to-use and efficient data manipulation tools
lm-evaluation-harness - A framework for few-shot evaluation of language models.
chicksexer - A Python package for gender classification.
python-sutime - Python wrapper for Stanford CoreNLP's SUTime
liquid_time_constant_networks - Code Repository for Liquid Time-Constant Networks (LTCs)
PaddleHub - Awesome pre-trained models toolkit based on PaddlePaddle. (400+ models including Image, Text, Audio, Video and Cross-Modal with Easy Inference & Serving)