Unicorn
NUWA
Our great sponsors
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
Unicorn
-
need help with object detection and object tracking using yolov4
Also check out Unicorn - https://github.com/MasterBin-IIAU/Unicorn
- [D] Most Popular AI Research July 2022 pt. 2 - Ranked Based On GitHub Stars
- Most Popular AI Research July 2022 pt. 2 - Ranked Based On GitHub Stars
-
Researchers from Bytedance and Dalian University Propose 🦄 ‘Unicorn’: a Unified Computer Vision Approach to Address Four Tracking Tasks Using a Single Model with the Same Model Parameters
Continue reading | Checkout the paper and github link
-
[R] Unicorn: 🦄 : Towards Grand Unification of Object Tracking(Video Demo)
Brief Overview We present a unified method, termed Unicorn, that can simultaneously solve four tracking problems (SOT, MOT, VOS, MOTS) with a single network using the same model parameters. For the first time, we accomplished the great unification of the tracking network architecture and learning paradigm. Unicorn performs on-par or better than its task-specific counterparts in 8 tracking datasets, including LaSOT, TrackingNet, MOT17, BDD100K, DAVIS16-17, MOTS20, and BDD100K MOTS. Our work is accepted to ECCV 2022 as an oral presentation ! Paper: https://arxiv.org/abs/2207.07078 Code: https://github.com/MasterBin-IIAU/Unicorn
-
[R] Unicorn: 🦄 : Towards Grand Unification of Object Tracking
Code for https://arxiv.org/abs/2207.07078 found: https://github.com/MasterBin-IIAU/Unicorn
NUWA
-
How long until we can create full length movies in ai ?
Github: https://github.com/microsoft/NUWA/tree/main/assets/nuwa_infinity/animation
-
[R] NUWA-Infinity, the first paper working on infinite visual synthesis!
Code for https://arxiv.org/abs/2207.09814 found: https://github.com/microsoft/NUWA
- [D] Most Popular AI Research July 2022 pt. 2 - Ranked Based On GitHub Stars
- Most Popular AI Research July 2022 pt. 2 - Ranked Based On GitHub Stars
-
I'm building a timeline for generative image ML models. What's missing?
Microsoft NUWA: https://github.com/microsoft/NUWA
- NUWA Infinity
-
With so many new Text to Image "AI" emerging lately, is it not crazy to speculate about Text to Video?
Microsoft NUWA
-
Have any researchers in the field discussed anything about the prospect of 'text-to-video' - something that's a bit like DALL-E 2, but with a video as the finished output?
NÃœWA from Microsoft.
- Art Student here. So about Dalle 2, am I in trouble or should I continue on with my studies? Moreover, what do you think the future holds in store for specific artists (ie comics as opposed to freelance writers as opposed to animators etc) in light of this announcement?
-
Imagine this: complete "fake AI people" are coming, and you didn't even see this coming!
P.S., Lucidrains remade it! AND he's adding an audio transformer to it tomorrow he says! But he needs feedback and someone to train it, I don't think there is enough resources helping this project's training. You can reach him through: https://github.com/microsoft/NUWA
What are some alternatives?
deeplab2 - DeepLab2 is a TensorFlow library for deep labeling, aiming to provide a unified and state-of-the-art TensorFlow codebase for dense pixel labeling tasks.
CogVideo - Text-to-video generation. The repo for ICLR2023 paper "CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers"
XMem - [ECCV 2022] XMem: Long-Term Video Object Segmentation with an Atkinson-Shiffrin Memory Model
DALLE2-video - Direct application of DALLE-2 to video synthesis, using factored space-time Unet and Transformers
theseus - A library for differentiable nonlinear optimization
latent-diffusion - High-Resolution Image Synthesis with Latent Diffusion Models
min-dalle - min(DALL·E) is a fast, minimal port of DALL·E Mini to PyTorch
yolov7 - Implementation of paper - YOLOv7: Trainable bag-of-freebies sets new state-of-the-art for real-time object detectors
hivemind - Decentralized deep learning in PyTorch. Built to train models on thousands of volunteers across the world.
Cream - This is a collection of our NAS and Vision Transformer work. [Moved to: https://github.com/microsoft/AutoML]