mamba
geospatial-data-lake
mamba | geospatial-data-lake | |
---|---|---|
15 | 5 | |
9,506 | 32 | |
15.3% | - | |
8.1 | 0.0 | |
8 days ago | about 1 year ago | |
Python | Python | |
Apache License 2.0 | MIT License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
mamba
-
Based: Simple linear attention language models
> how the recall can grow unbounded with no tradeoff
this? https://github.com/state-spaces/mamba/issues/175
-
Mamba: The Easy Way
If you want to learn this stuff as a computer engineer, you can read the code here [0]. I find the math quite helpful.
[0]: https://github.com/state-spaces/mamba
- FLaNK Stack 05 Feb 2024
- Introduction to State Space Models (SSM)
-
Fortran inference code for the Mamba state space language model
This model was discussed recently: https://news.ycombinator.com/item?id=38522428 It's a new kind of ML model architecture that can be used instead of a transformer in LLMs.
See also the original repo from the paper: https://github.com/state-spaces/mamba
-
Mamba outperforms transformers "everywhere we tried"
[2] - https://github.com/state-spaces/mamba
Out of curiosity, does anyone feel as though there's any benefit to linking to reddit when we can link to whatever the link is? I for one do not click the link and read discussion on reddit - if I wanted that sort of discussion, I would browse there, not HN.
- GitHub – State-Spaces/Mamba
-
Generate valid JSON with Mamba models
The library is compatible with any auto-regressive model, not transformers. To prove our point we integrated Mamba, a new state-space model architecture, to the library. Try it out!
-
[D] Thoughts on Mamba?
I ran the NanoGPT of Karparthy replacing Self-Attention with Mamba on his TinyShakespeare Dataset and within 5 minutes it started spitting out the following:
-
Mamba-Chat: A Chat LLM based on State Space Models
You might have come across the paper Mamba paper in the last days, which was the first attempt at scaling up state space models to 2.8B parameters to work on language data.
geospatial-data-lake
-
A curated list of questionable installation instructions
One option is to trust on first use, checksum the installation script and at least casually verify the diff each time the checksum changes[1].
Pros:
- Protects against simple hijacking.
- Reproducible as long as the installer doesn't also call out to a moving target, such as example.com/releases/latest.
Cons:
- Build breaks as soon as the installer is bumped. If it's bumped often (or just before an important release) this can cause pain.
- TOFU may not be acceptable, but of course you could review the code thoroughly before even the first use.
[1] https://github.com/linz/geostore/blob/b3cd162605109da8a3a688...
-
Ask HN: Good Python projects to read for modern Python?
I'd recommend a project from work, Geostore[1]. Highlights:
- 100% test coverage (with some typical exceptions like `if __name__ == "__main__":` blocks)
- Randomises test sequence and inputs reproducibly
- Passes Pylint with max McCabe complexity of 6
- Passes `mypy --strict`
- Formatted using Black and isort
[1] https://github.com/linz/geostore
-
Python Best Practices for a New Project in 2021
The current work project[1] has all of these: Pyenv, Poetry, Pytest, pytest-cov with 100% branch coverage, pre-commit, Pylint rather than Flake8, Black, mypy (with a stricter configuration than recommended here), and finally isort. These are all super helpful.
There's also a simpler template repo[2] with almost all of these.
[1] https://github.com/linz/geostore/
[2] https://github.com/linz/template-python-hello-world
- Codecov bash uploader was compromised
-
AWS CloudFormation Best Practices
As someone who's used CDK for a few months and never handcoded CF, that sounds completely correct. If you're comfortable with Python, here's a simple but non-trivial architecture you can check out: https://github.com/linz/geospatial-data-lake/blob/master/app....
What are some alternatives?
miniforge - A conda-forge distribution.
pydantic-factories - Simple and powerful mock data generation using pydantic or dataclasses
pip - The Python package installer
template-python-hello-world - :triangular_ruler: Python Hello World | Minimal template for Python development
llm.f90 - LLM inference in Fortran
asgi-correlation-id - Request ID propagation for ASGI apps
conda - A system-level, binary package and environment manager running on all major operating systems and platforms.
aws-cdk - The AWS Cloud Development Kit is a framework for defining cloud infrastructure in code
mamba-chat - Mamba-Chat: A chat LLM based on the state-space model architecture 🐍
dev-tasks - Automated development tasks for my own projects
spack - A flexible package manager that supports multiple versions, configurations, platforms, and compilers.