Hangzhou, Zhejiang, China

chx_ant

Advanced
@chenhaoxing

AI Researcher in AGI Center, Ant Research Institute.

DiffusionInst. This repo is the code of paper "DiffusionInst: Diffusion Model for Instance Segmentation" (ICASSP'24).

244

Awesome-AI-Generated-Video-Detection. ✨✨Latest Papers on AI-Generated Video Detection and Related Areas

215

DeMamba. This repository is the code of paper 'DeMamba: AI-Generated Video Detection on Million-Scale GenVideo Benchmark'.

194

DiffUTE. This repository is the code of our paper "DiffUTE: Universal Text Editing Diffusion Model" (NeurIPS'2023).

143

HDNet. This is the implementation of paper 'Hierarchical Dynamic Image Harmonization' (ACM MM'2023, Oral).

90

SSFormers. This repository is the code of the paper "Sparse Spatial Transformers for Few-Shot Learning" (SCIENCE CHINA Information Sciences).

48

M2L. This repository is the code of paper "Multi-level Metric Learning for Few-shot Image Recognition".(ICANN-2022))

34

MATANet. This repository is the code of paper "Multi-scale Adaptive Task Attention Network for Few-Shot Learning (ICPR-2022)".

26

ASL. This repository is the code of the paper "Shaping Visual Representations with Attributes for Few-Shot Learning (IEEE SPL)".

10

ETL4Video. This is the official code of paper 'Efficient Transfer Learning for Video-language Foundation Models'. (CVPR'2025)

8

MACL. This repo is the code of paper "Model-Aware Contrastive Learning: Towards Escaping Dilemmas" (ICML'2023).

8

KDA. This repository is the code of paper 'Boosting Audio-visual Zero-shot Learning with Large Language Models'.

7

CPR. This repository is the code of our paper "Conditional Prototype Rectification Prompt Learning".

6

Awesome-Multi-modal-Diffusion-Large-Language-Model. ✨✨Latest Papers on Multimodal dLLM and Related Areas

5

ShuffleMamba. Code of paper 'Stochastic Layer-Wise Shuffle: A Good Practice to Improve Vision Mamba Training'

1
15
Apply