trax vs flax

Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.

www.influxdata.com

featured

SaaSHub - Software Alternatives and Reviews

SaaSHub helps you find the best software and product alternatives

www.saashub.com

featured

trax		flax
	Project
7	Mentions	10
7,957	Stars	5,538
0.4%	Growth	2.2%
4.7	Activity	9.7
3 months ago	Latest Commit	2 days ago
Python	Language	Python
Apache License 2.0	License	Apache License 2.0

The number of mentions indicates the total number of mentions that we've tracked plus the number of user suggested alternatives.
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.

trax

Posts with mentions or reviews of trax. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2024-04-23.

Maxtext: A simple, performant and scalable Jax LLM
10 projects | news.ycombinator.com | 23 Apr 2024

Is t5x an encoder/decoder architecture?
Some more general options.
The Flax ecosystem
https://github.com/google/flax?tab=readme-ov-file
or dm-haiku
https://github.com/google-deepmind/dm-haiku
were some of the best developed communities in the Jax AI field
Perhaps the “trax” repo? https://github.com/google/trax
Some HF examples https://github.com/huggingface/transformers/tree/main/exampl...
Sadly it seems much of the work is proprietary these days, but one example could be Grok-1, if you customize the details. https://github.com/xai-org/grok-1/blob/main/run.py
Replit's new Code LLM was trained in 1 week
12 projects | news.ycombinator.com | 3 May 2023

and the implementation https://github.com/google/trax/blob/master/trax/models/resea... if you are interested.
Hope you get to look into this!
RedPajama: Reproduction of Llama with Friendly License
4 projects | news.ycombinator.com | 17 Apr 2023

Thank you for developing the pipeline and amassing considerable compute for gathering and preprocessing this dataset!
I'm not sure if this is the right place to ask about this, but could you consider training an LLM using a more advanced, sparse transformer architecture (specifically, "Terraformer" from this paper https://arxiv.org/abs/2111.12763 and this codebase https://github.com/google/trax/blob/master/trax/models/resea... by Google Brain and OpenAI)? I understand the pressure to focus on training a straightforward LLaMA replication, but of course you see that it's a legacy dense architecture which limits its inference performance. This new architecture is not just an academic curiosity but is already validated at scale by Google, providing 10x+ inference performance boost on the same hardware.
Frankly, the community's compute budget - for training and for inference - isn't infinite, and neither is the public's interest in models that do not have advantage (at least in convenience) over closed-source ones; and so we should utilize both those resources as efficiently as possible. It could be a big step forward if you trained at least LLaMA-Terraformer-7B and 13B foundation models on the whole dataset.
The founder of Gmail claims that ChatGPT can “kill” Google in two years.
1 project | /r/Futurology | 31 Jan 2023

But a couple years later they came out with open source implementations yeah: https://github.com/google/trax/tree/master/trax/models/reformer
[D] Paper Explained - Sparse is Enough in Scaling Transformers (aka Terraformer) | Video Walkthrough
1 project | /r/MachineLearning | 1 Dec 2021

Code: https://github.com/google/trax/blob/master/trax/examples/Terraformer_from_scratch.ipynb
Why would I want to develop yet another deep learning framework?
4 projects | /r/learnmachinelearning | 16 Sep 2021
How to train large models on a normal laptop?
1 project | /r/LanguageTechnology | 14 Feb 2021

Training language models is expensive. Train the biggest model you can afford. I assume you've tried the colab from the reformer GitHub: https://github.com/google/trax/tree/master/trax/models/reformer

flax

Posts with mentions or reviews of flax. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2024-04-23.

Maxtext: A simple, performant and scalable Jax LLM
10 projects | news.ycombinator.com | 23 Apr 2024

Is t5x an encoder/decoder architecture?
Some more general options.
The Flax ecosystem
https://github.com/google/flax?tab=readme-ov-file
or dm-haiku
https://github.com/google-deepmind/dm-haiku
were some of the best developed communities in the Jax AI field
Perhaps the “trax” repo? https://github.com/google/trax
Some HF examples https://github.com/huggingface/transformers/tree/main/exampl...
Sadly it seems much of the work is proprietary these days, but one example could be Grok-1, if you customize the details. https://github.com/xai-org/grok-1/blob/main/run.py
What is the JAX/Flax equivalent of torch.nn.Parameter?
1 project | /r/JAX | 24 Apr 2023

https://github.com/google/flax/discussions/919 https://flax.readthedocs.io/en/latest/_modules/flax/linen/attention.html
Announcing flax 0.2 - A fully featured ECS
3 projects | /r/rust | 11 Sep 2022

Just as an FYI, you might be competing against another big open source project with the same name https://github.com/google/flax
Flax: How to use one linen module inside another for training?
1 project | /r/deeplearning | 28 Jun 2022

I have asked the same question on the Flax discussion page on Github as well.
[D] Should We Be Using JAX in 2022?
8 projects | /r/MachineLearning | 15 Feb 2022

What's your favorite Deep Learning API for JAX - Flax, Haiku, Elegy, something else?
PyTorch vs. TensorFlow in 2022
13 projects | news.ycombinator.com | 14 Dec 2021

As a researcher in RL & ML in a big industry lab, I would say most of my colleagues are moving to JAX 0https://github.com/google/jax], which this article kind of ignores. JAX is XLA-accelerated NumPy, it's cool beyond just machine learning, but only provides low-level linear algebra abstractions. However you can put something like Haiku [https://github.com/deepmind/dm-haiku] or Flax [https://github.com/google/flax] on top of it and get what the cool kids are using :)
[D] Getting Started with Deep Learning in JAX with Treex in 16 lines
1 project | /r/MachineLearning | 30 Oct 2021
[D] JAX learning resources?
4 projects | /r/JAX | 23 Sep 2021

- https://github.com/google/flax/tree/main/examples
Why would I want to develop yet another deep learning framework?
4 projects | /r/learnmachinelearning | 16 Sep 2021
[D] Why is tensorflow so hated on and pytorch is the cool kids framework?
4 projects | /r/MachineLearning | 12 Mar 2021

Any thoughts on Flax?

What are some alternatives?

When comparing trax and flax you can also consider the following projects:

dm-haiku - JAX-based neural network library

muzero-general - MuZero

Pytorch - Tensors and Dynamic neural networks in Python with strong GPU acceleration

ML-Optimizers-JAX - Toy implementations of some popular ML optimizers using Python/JAX

equinox - Elegant easy-to-use neural networks + scientific computing in JAX. https://docs.kidger.site/equinox/

extending-jax - Extending JAX with custom C++ and CUDA code

Flux.jl - Relax! Flux is the ML library that doesn't make you tensor

objax

numpyro - Probabilistic programming with NumPy powered by JAX for autograd and JIT compilation to GPU/TPU/CPU.

tf-transformers - State of the art faster Transformer with Tensorflow 2.0 ( NLP, Computer Vision, Audio ).

trax vs dm-haiku flax vs dm-haiku trax vs muzero-general flax vs Pytorch trax vs ML-Optimizers-JAX flax vs equinox trax vs extending-jax flax vs Flux.jl trax vs objax flax vs objax trax vs numpyro flax vs tf-transformers

Compare trax vs flax and see what are their differences.

trax

flax

trax

flax

What are some alternatives?