XMem
NUWA
Our great sponsors
XMem | NUWA | |
---|---|---|
10 | 23 | |
1,187 | 2,700 | |
- | 0.9% | |
7.9 | 4.7 | |
17 days ago | 8 days ago | |
Python | ||
GNU General Public License v3.0 only | - |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
XMem
-
Track-Anything: a flexible and interactive tool for video object tracking and segmentation, based on Segment Anything and XMem.
Nvm just found the occlusion video on https://github.com/hkchengrex/XMem holy shit
-
[D] Most important AI Paper´s this year so far in my opinion + Proto AGI speculation at the end
XMem: Long-Term Video Object Segmentation with an Atkinson-Shiffrin Memory Model ( Added because of the Atkinson-Shiffrin Memory Model ) Paper: https://arxiv.org/abs/2207.07115 Github: https://github.com/hkchengrex/XMem
- [D] Most Popular AI Research July 2022 pt. 2 - Ranked Based On GitHub Stars
- Most Popular AI Research July 2022 pt. 2 - Ranked Based On GitHub Stars
-
I trained a neural net to watch Super Smash Bros
Yeah MiVOS would speed up your tagging a lot. I also was curious if you saw XMem which just came out. I found that worked really well too.
-
[R] Unicorn: 🦄 : Towards Grand Unification of Object Tracking(Video Demo)
Have you check XMem?
NUWA
-
How long until we can create full length movies in ai ?
Github: https://github.com/microsoft/NUWA/tree/main/assets/nuwa_infinity/animation
- [D] Most Popular AI Research July 2022 pt. 2 - Ranked Based On GitHub Stars
- Most Popular AI Research July 2022 pt. 2 - Ranked Based On GitHub Stars
-
I'm building a timeline for generative image ML models. What's missing?
Microsoft NUWA: https://github.com/microsoft/NUWA
-
With so many new Text to Image "AI" emerging lately, is it not crazy to speculate about Text to Video?
Microsoft NUWA
-
Have any researchers in the field discussed anything about the prospect of 'text-to-video' - something that's a bit like DALL-E 2, but with a video as the finished output?
NÜWA from Microsoft.
-
Imagine this: complete "fake AI people" are coming, and you didn't even see this coming!
P.S., Lucidrains remade it! AND he's adding an audio transformer to it tomorrow he says! But he needs feedback and someone to train it, I don't think there is enough resources helping this project's training. You can reach him through: https://github.com/microsoft/NUWA
-
NÜWA - text to image
(from here)
What are some alternatives?
CogVideo - Text-to-video generation. The repo for ICLR2023 paper "CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers"
yolov7 - Implementation of paper - YOLOv7: Trainable bag-of-freebies sets new state-of-the-art for real-time object detectors
DALLE2-video - Direct application of DALLE-2 to video synthesis, using factored space-time Unet and Transformers
flash-attention - Fast and memory-efficient exact attention
deeplab2 - DeepLab2 is a TensorFlow library for deep labeling, aiming to provide a unified and state-of-the-art TensorFlow codebase for dense pixel labeling tasks.
min-dalle - min(DALL·E) is a fast, minimal port of DALL·E Mini to PyTorch
multiface - Hosts the Multiface dataset, which is a multi-view dataset of multiple identities performing a sequence of facial expressions.
theseus - A library for differentiable nonlinear optimization
stylegan2-pytorch - Simplest working implementation of Stylegan2, state of the art generative adversarial network, in Pytorch. Enabling everyone to experience disentanglement
Cream - This is a collection of our NAS and Vision Transformer work. [Moved to: https://github.com/microsoft/AutoML]
latent-diffusion - High-Resolution Image Synthesis with Latent Diffusion Models