This is your work, valued
TPT. Test-time Prompt Tuning (TPT) for zero-shot generalization in vision-language models (NeurIPS 2022))
★ 214AutoPoison. The official repository of the paper "On the Exploitability of Instruction Tuning".
★ 70AdvBN. Official PyTorch implementation of the NeurIPS 2021 paper: Encoding Robustness to Image Style via Adversarial Feature Perturbations.
★ 9adversarial_data_augmentation. This repository provides the official PyTorch implementation of the ICRA 2021 paper: Adversarial Differentiable Data Augmentation for Autonomous Systems.
★ 6RAGEN. RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.
★ 2.8kAwesome-ML-SYS-Tutorial. My learning notes for ML SYS.
★ 6.8kcambrian. Cambrian-1 is a family of multimodal LLMs with a vision-centric design.
★ 2kTensor-Puzzles. Solve puzzles. Improve your pytorch.
★ 4.3kRL4VLM. Official Repo for Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
★ 415verl. verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
★ 23kTinyZero. Minimal reproduction of DeepSeek R1-Zero
★ 13kOmAgent. [EMNLP-2024] Build multimodal language agents for fast prototype and production
★ 2.7kopen-instruct. AllenAI's post-training codebase
★ 3.8kOpenRLHF. An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
★ 9.9kLATTE. Python
★ 70ProVision. A instruction data generation system for multimodal language models.
★ 37attention_sinks. Extend existing LLMs way beyond the original training length with constant memory usage, without retraining
★ 735Hierarchical_Point_Attention. Python
★ 10MINT-1T. 🍃 MINT-1T: A one trillion token multimodal interleaved dataset.
★ 833SoM-LLaVA. [COLM-2024] List Items One by One: A New Data Source and Learning Paradigm for Multimodal LLMs
★ 146diffusion_memorization. Official repo for Detecting, Explaining, and Mitigating Memorization in Diffusion Models (ICLR 2024)
★ 80carving. Package to optimize Adversarial Attacks against (Large) Language Models with Varied Objectives
★ 71open_flamingo. An open-source framework for training large multimodal models.
★ 4.1kLAVIS. LAVIS - A One-stop Library for Language-Vision Intelligence
★ 11kxtuner. A Next-Generation Training Engine Built for Ultra-Large MoE Models
★ 5.2kconsistencydecoder. Consistency Distilled Diff VAE
★ 2.2knanoGPT. The simplest, fastest repository for training/finetuning medium-sized GPTs.
★ 62kMegatron-LM. Ongoing research training transformer models at scale
★ 17kNEFTune. Official repository of NEFTune: Noisy Embeddings Improves Instruction Finetuning
★ 412DeepSpeed. DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
★ 43kddim. Denoising Diffusion Implicit Models
★ 1.8kglide-text2im. GLIDE: a diffusion-based text-conditional image synthesis model
★ 3.7kFastChat. An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
★ 40klm-evaluation-harness. A framework for few-shot evaluation of language models.
★ 13kAutoPoison. The official repository of the paper "On the Exploitability of Instruction Tuning".
★ 70GOAT. Official implementation of GOAT model (ICML2023)
★ 38DCR. Official Pytorch repo of CVPR'23 and NeurIPS'23 papers on understanding replication in diffusion models.
★ 113tree-ring-watermark. Python
★ 366BYOD. The Official Repository for "Bring Your Own Data! Self-Supervised Evaluation for Large Language Models"
★ 108guidance. A guidance language for controlling large language models.
★ 22kevals. Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.
★ 19kTPT. Test-time Prompt Tuning (TPT) for zero-shot generalization in vision-language models (NeurIPS 2022))
★ 214dev. Press the . key on any repo
★ 1.5ksegment-anything. The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 55kContrastiveDecoding. contrastive decoding
★ 206lm-watermarking. Jupyter Notebook
★ 680open_clip. An open source implementation of CLIP.
★ 14ksnip-dedup. Python
★ 104textual_inversion. Jupyter Notebook
★ 3.1klatent-diffusion. High-Resolution Image Synthesis with Latent Diffusion Models
★ 14kNL-Augmenter. NL-Augmenter 🦎 → 🐍 A Collaborative Repository of Natural Language Transformations
★ 786stable-diffusion. A latent text-to-image diffusion model
★ 73kpytorch-lightning. Pretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.
★ 31kkornia. 🐍 Geometric Computer Vision Library for Spatial AI
★ 11kbyol-pytorch. Usable Implementation of "Bootstrap Your Own Latent" self-supervised learning, from Deepmind, in Pytorch
★ 1.9kIDPG. IDPG: An Instance-Dependent Prompt Generation Method
★ 53detr. Code & Models for 3DETR - an End-to-end transformer model for 3D object detection
★ 712federated. A collection of Google research projects related to Federated Learning and Federated Analytics.
★ 759vit-visualization. Shell
★ 197pytorch-image-models. The largest collection of PyTorch image encoders / backbones. Including train, eval, inference, export scripts, and pretrained weights -- ResNet, ResNeXT, EfficientNet, NFNet, Vision Transformer (ViT), MobileNetV4, MobileNet-V3 & V2, RegNet, DPN, CSPNet, Swin Transformer, MaxViT, CoAtNet, ConvNeXt, and more
★ 37kOpenPrompt. An Open-Source Framework for Prompt-Learning.
★ 4.9kPrefixTuning. Prefix-Tuning: Optimizing Continuous Prompts for Generation
★ 961prompt-tuning. Original Implementation of Prompt Tuning from Lester, et al, 2021
★ 701nerfies.github.io. JavaScript
★ 4.3kGroup-Free-3D. Group-Free 3D Object Detection via Transformers
★ 256AdvBN. Official PyTorch implementation of the NeurIPS 2021 paper: Encoding Robustness to Image Style via Adversarial Feature Perturbations.
★ 9dialoglue. DialoGLUE: A Natural Language Understanding Benchmark for Task-Oriented Dialogue
★ 288Compositional_Learning. Compositional Learning for Human Object Interaction
★ 13P-tuning-v2. An optimized deep prompt tuning strategy comparable to fine-tuning across scales and tasks
★ 2.1kpytorch-maml. PyTorch implementation of MAML: https://arxiv.org/abs/1703.03400
★ 567tent. ICLR21 Tent: Fully Test-Time Adaptation by Entropy Minimization
★ 472CoOp. Prompt Learning for Vision-Language Models (IJCV'22, CVPR'22)
★ 2.2kImageNetV2. A new test set for ImageNet
★ 267awesome-self-supervised-learning. A curated list of awesome self-supervised methods
★ 6.4kvit-pytorch. Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch
★ 25kGLIP. Grounded Language-Image Pre-training
★ 2.6kCLIP. CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
★ 34kadversarial_data_augmentation. This repository provides the official PyTorch implementation of the ICRA 2021 paper: Adversarial Differentiable Data Augmentation for Autonomous Systems.
★ 6pytorch-AdaIN. Unofficial pytorch implementation of 'Arbitrary Style Transfer in Real-time with Adaptive Instance Normalization' [Huang+, ICCV2017]
★ 1.2ksemseg. Semantic Segmentation in Pytorch
★ 1.4kFACT. Python
★ 188moco. PyTorch implementation of MoCo: https://arxiv.org/abs/1911.05722
★ 5.1kRobustNet. Official PyTorch implementation of RobustNet (CVPR 2021 Oral)
★ 225MetaAug. Python
★ 27pytorch-classification-advprop. Python
★ 106Deep-Learning-Papers-Reading-Roadmap. Deep Learning papers reading roadmap for anyone who are eager to learn this amazing tech!
★ 40kcomputer. C++
★ 408fad-texlive-talk. texlive talk materials in FAD Beijing 2011
★ 14vision. Datasets, Transforms and Models specific to Computer Vision
★ 18kpytorch-book. PyTorch tutorials and fun projects including neural talk, neural style, poem writing, anime generation (《深度学习框架PyTorch:入门与实战》)
★ 13kos_kernel_lab. OS kernel labs based on Rust/C Lang & RISC-V 64/X86-32
★ 4kChatterBot. ChatterBot is a machine learning, conversational dialog engine for creating chat bots
★ 15kchainer-DCGAN. Chainer implementation of Deep Convolutional Generative Adversarial Network
★ 932caffe. Caffe: a fast open framework for deep learning.
★ 35kneural-style. Torch implementation of neural style algorithm
★ 18kxv6-chinese. 中文版的 MIT xv6 文档
★ 3.5k