XMem vs mmdetection

Our great sponsors

WorkOS - The modern identity platform for B2B SaaS

InfluxDB - Power Real-Time Data Analytics at Scale

SaaSHub - Software Alternatives and Reviews

Our great sponsors

XMem		mmdetection
	Project
11	Mentions	23
1,584	Stars	27,658
-	Growth	2.0%
6.3	Activity	8.7
about 1 month ago	Latest Commit	4 days ago
Python	Language	Python
MIT License	License	Apache License 2.0

The number of mentions indicates the total number of mentions that we've tracked plus the number of user suggested alternatives.
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.

XMem

Posts with mentions or reviews of XMem. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-06-11.

[D] Which open source models can replicate wonder dynamics's drag'n'drop cg characters?
6 projects | /r/MachineLearning | 11 Jun 2023

Use Segmentation Model (SAM) combined with Inpainting model (E2FGVI) and Xmem to cut out the live action subject.
Track-Anything: a flexible and interactive tool for video object tracking and segmentation, based on Segment Anything and XMem.
3 projects | /r/StableDiffusion | 25 Apr 2023

Nvm just found the occlusion video on https://github.com/hkchengrex/XMem holy shit
XMem: Long-Term Video Object Segmentation with an Atkinson-Shiffrin Memory Model
1 project | news.ycombinator.com | 22 Aug 2022

1 project | news.ycombinator.com | 16 Jul 2022
[D] Most important AI Paper´s this year so far in my opinion + Proto AGI speculation at the end
10 projects | /r/MachineLearning | 14 Aug 2022

XMem: Long-Term Video Object Segmentation with an Atkinson-Shiffrin Memory Model ( Added because of the Atkinson-Shiffrin Memory Model ) Paper: https://arxiv.org/abs/2207.07115 Github: https://github.com/hkchengrex/XMem
[D] Most Popular AI Research July 2022 pt. 2 - Ranked Based On GitHub Stars
10 projects | /r/MachineLearning | 6 Aug 2022
Most Popular AI Research July 2022 pt. 2 - Ranked Based On GitHub Stars
10 projects | /r/learnmachinelearning | 2 Aug 2022
I trained a neural net to watch Super Smash Bros
3 projects | /r/computervision | 25 Jul 2022

Yeah MiVOS would speed up your tagging a lot. I also was curious if you saw XMem which just came out. I found that worked really well too.
University of Illinois Researchers Develop XMem; A Long-Term Video Object Segmentation Architecture Inspired By Atkinson-Shiffrin Memory Model
1 project | /r/computervision | 20 Jul 2022

Continue reading | Check out the paper and github link.
[R] Unicorn: 🦄 : Towards Grand Unification of Object Tracking(Video Demo)
2 projects | /r/MachineLearning | 18 Jul 2022

Have you check XMem?

mmdetection

Posts with mentions or reviews of mmdetection. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-04-12.

Semantic segementation
2 projects | /r/computervision | 12 Apr 2023

When I look for benchmarks I always start here https://paperswithcode.com/task/instance-segmentation/codeless it has the lists of datasets to measure models accross lots o papers. Many are very specific models with low support or community but it gives you a good idea of the state of the art. It also lists repositories related to good community. https://github.com/open-mmlab/mmdetection seems very active and the one that is being used the most, you could use the models that it has integrated in its model zoo, within the same repository. It has the benchmarks to compare those same models and some of them are from 2022
How to Convert Model Mask into Polygon and save JSON?
1 project | /r/deeplearning | 18 Jan 2023

MODEL: https://github.com/open-mmlab/mmdetection
Object Detection Model for Custom Dataset Training?
1 project | /r/learnmachinelearning | 11 Jan 2023

Would it make sense to work with OpenMMLab (https://github.com/open-mmlab/mmdetection) or Pytorch-image-models (https://github.com/rwightman/pytorch-image-models#models) since they offer a variety of models?
[P] Image search with localization and open-vocabulary reranking.
8 projects | /r/MachineLearning | 15 Dec 2022

I wanted to have a few choices getting localization into image search (index and search time). I immediately thought of using a region proposal network (rpn) from mask-rcnn to create patches that can also be indexed and searched (and add the localisation). I figured it might be somewhat agnostic to classes. I did not want to use mmdetection or detectron2 due to their dependencies and just getting the rpn was not worth it. I was encouraged by the PyTorch native implementations of detection/segmentation models but ended up finding yolox the best.
MMDeploy: Deploy All the Algorithms of OpenMMLab
22 projects | /r/u_Allent_pjlab | 21 Nov 2022

MMDetection: OpenMMLab detection toolbox and benchmark.
Removing the bounding box generated by OnnxRuntime segmentation
2 projects | /r/computervision | 4 Nov 2022

I have a semantic segmentation model trained using the mmdetection repo. Then it is converted to the ONNX format using the mmdeploy repo.
Keras vs Tensorflow vs Pytorch for a Final year Project
2 projects | /r/tensorflow | 10 Oct 2022

E.g. If you consider it an object detection problem it is: detect and localise all the pedestrians in a frame, and classify them by their (intended) action. IMO the easiest way to do this would be with mmdetection, which is built on top of pytorch. Just label your dataset, build a config, and boom you have a model. Inference with that model in only a few lines of code, you won't really need to learn too much to get started.
DeepSort with PyTorch(support yolo series)
13 projects | /r/u_No_Experience9104 | 20 Sep 2022

MMDetection
[D] Pre-trained networks and batch normalization
1 project | /r/MachineLearning | 15 Sep 2022

For example, in mmdetection, they expose options in their config & implementation to freeze batch norm layers in backbones and in this config, norm_eval is set to True meaning to freeze tracking of batch norm stats, while the ResNet backbone is frozen up to the 1st stage. Example of their backbone implementation can be found here.
Config files in plain Python
3 projects | /r/Python | 25 Aug 2022

MMDetection uses config Python scripting. It's easier to define nn.Module objects other than writing class name in a json config file

What are some alternatives?

When comparing XMem and mmdetection you can also consider the following projects:

yolov7 - Implementation of paper - YOLOv7: Trainable bag-of-freebies sets new state-of-the-art for real-time object detectors

detectron2 - Detectron2 is a platform for object detection, segmentation and other visual recognition tasks.

flash-attention - Fast and memory-efficient exact attention

yolov5 - YOLOv5 🚀 in PyTorch > ONNX > CoreML > TFLite

NAFNet - The state-of-the-art image restoration model without nonlinear activation functions.

pytorch-lightning - Build high-performance AI models with PyTorch Lightning (organized PyTorch). Deploy models with Lightning Apps (organized Python to build end-to-end ML systems). [Moved to: https://github.com/Lightning-AI/lightning]

deeplab2 - DeepLab2 is a TensorFlow library for deep labeling, aiming to provide a unified and state-of-the-art TensorFlow codebase for dense pixel labeling tasks.

PaddleDetection - Object Detection toolkit based on PaddlePaddle. It supports object detection, instance segmentation, multiple object tracking and real-time multi-person keypoint detection.

Cream - This is a collection of our NAS and Vision Transformer work. [Moved to: https://github.com/microsoft/AutoML]

mmdetection3d - OpenMMLab's next-generation platform for general 3D object detection.

EfficientZero - Open-source codebase for EfficientZero, from "Mastering Atari Games with Limited Data" at NeurIPS 2021.

sahi - Framework agnostic sliced/tiled inference + interactive ui + error analysis plots

XMem vs yolov7 mmdetection vs detectron2 XMem vs flash-attention mmdetection vs yolov5 XMem vs NAFNet mmdetection vs pytorch-lightning XMem vs deeplab2 mmdetection vs PaddleDetection XMem vs Cream mmdetection vs mmdetection3d XMem vs EfficientZero mmdetection vs sahi

Compare XMem vs mmdetection and see what are their differences.

XMem

mmdetection

XMem

mmdetection

What are some alternatives?