DiffusionDet. [ICCV2023 Best Paper Finalist] PyTorch implementation of DiffusionDet (https://arxiv.org/abs/2211.09788)
2.3kAdaptFormer. [NeurIPS 2022] Implementation of "AdaptFormer: Adapting Vision Transformers for Scalable Visual Recognition"
388PixelFlow. Pixel-Space Generative Models
316CycleMLP. [ICLR'22 Oral] Implementation of "CycleMLP: A MLP-like Architecture for Dense Prediction"
290Awesome-Diffusion-Transformers. https://www.shoufachen.com/Awesome-Diffusion-Transformers/
147WOO. [ICCV'21] Implementation of "Watch Only Once: An End-to-End Video Action Detection Framework"
45clone-anonymous4open. clone/download codes from https://anonymous.4open.science/
33gradio-box. Python
20Grounded-Segment-Anything-patch. Marrying Grounding DINO with Segment Anything & Stable Diffusion & BLIP - Automatically Detect , Segment and Generate Anything with Image and Text Inputs
13COMP3340_Transformer_MLP. Python
3img2dataset. Easily turn large sets of image urls to an image dataset. Can download, resize and package 100M urls in 20h on one machine.
3SlowFast. PySlowFast: video understanding codebase from FAIR for reproducing state-of-the-art video models.
1accelerate-patch. 🚀 A simple way to train and use PyTorch models with multi-GPU, TPU, mixed-precision
1Awesome-Video-Diffusion. A curated list of recent diffusion models for video generation, editing, restoration, understanding, etc.
1pytorch-grad-cam. Many Class Activation Map methods implemented in Pytorch for CNNs and Vision Transformers. Examples for classification, object detection, segmentation, embedding networks and more. Including Grad-CAM, Grad-CAM++, Score-CAM, Ablation-CAM and XGrad-CAM
1FaceRecognition. Face Recognition Using Python and MySQL
1Awesome-Anything-patch. AI methods for Anything: AnyObject, AnyGeneration, AnyModel, AnyTask
1