AI Researcher in AGI Center, Ant Research Institute.
DiffusionInst. This repo is the code of paper "DiffusionInst: Diffusion Model for Instance Segmentation" (ICASSP'24).
244Awesome-AI-Generated-Video-Detection. ✨✨Latest Papers on AI-Generated Video Detection and Related Areas
215DeMamba. This repository is the code of paper 'DeMamba: AI-Generated Video Detection on Million-Scale GenVideo Benchmark'.
194DiffUTE. This repository is the code of our paper "DiffUTE: Universal Text Editing Diffusion Model" (NeurIPS'2023).
143HDNet. This is the implementation of paper 'Hierarchical Dynamic Image Harmonization' (ACM MM'2023, Oral).
90SSFormers. This repository is the code of the paper "Sparse Spatial Transformers for Few-Shot Learning" (SCIENCE CHINA Information Sciences).
48M2L. This repository is the code of paper "Multi-level Metric Learning for Few-shot Image Recognition".(ICANN-2022))
34MATANet. This repository is the code of paper "Multi-scale Adaptive Task Attention Network for Few-Shot Learning (ICPR-2022)".
26ASL. This repository is the code of the paper "Shaping Visual Representations with Attributes for Few-Shot Learning (IEEE SPL)".
10ETL4Video. This is the official code of paper 'Efficient Transfer Learning for Video-language Foundation Models'. (CVPR'2025)
8MACL. This repo is the code of paper "Model-Aware Contrastive Learning: Towards Escaping Dilemmas" (ICML'2023).
8KDA. This repository is the code of paper 'Boosting Audio-visual Zero-shot Learning with Large Language Models'.
7CPR. This repository is the code of our paper "Conditional Prototype Rectification Prompt Learning".
6Awesome-Multi-modal-Diffusion-Large-Language-Model. ✨✨Latest Papers on Multimodal dLLM and Related Areas
5ShuffleMamba. Code of paper 'Stochastic Layer-Wise Shuffle: A Good Practice to Improve Vision Mamba Training'
1