I‘m a PhD student at CUHK with the topic of multimodal learning. Feel free to contact me: bjzi@se.cuhk.edu.hk.
MiniMax-Remover. This is the official implementation of our paper: "MiniMax-Remover: Taming Bad Noise Helps Video Object Removal"
589COCOCO. Video-Inpaint-Anything: This is the inference code for our paper CoCoCo: Improving Text-Guided Video Inpainting for Better Consistency, Controllability and Compatibility.
325SENORITA. This is the official implementation of our Señorita-2M [Weights and Dataset] : A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists
112RSLAD. This is the official code for "Revisiting Adversarial Robustness Distillation: Robust Soft Labels Make Student Better"
45NLP_and_Multimodal_Distillation.
4Video_Language_Pretraining.
1Cross_Modal_Retrieval.
1-.
1