This is your work, valued
Dedicated to open-source communities.
TransUNet. This repository includes the official project of TransUNet, presented in our paper: TransUNet: Transformers Make Strong Encoders for Medical Image Segmentation.
3.2k3D-TransUNet. This is the official repository for the paper "3D TransUNet: Advancing Medical Image Segmentation through Vision Transformers"
315ViTamin. [CVPR 2024] Official implementation of "ViTamin: Designing Scalable Vision Models in the Vision-language Era"
211TransMix. [CVPR 2022] This repository includes the official project for the paper: TransMix: Attend to Mix for Vision Transformers.
158LLaVolta. [NeurIPS 2024] Efficient Large Multi-modal Models via Visual Context Compression
66spatialcode. Open studio for "Thinking with Spatial Code" (https://arxiv.org/pdf/2603.05591)
204D-Animal. PyTorch code for 4D-Animal.
3pytorch-image-models. PyTorch image models, scripts, pretrained weights -- ResNet, ResNeXT, EfficientNet, NFNet, Vision Transformer (ViT), MobileNet-V3/V2, RegNet, DPN, CSPNet, Swin Transformer, MaxViT, CoAtNet, ConvNeXt, and more
1MeWM. Medical World Model
1