Research Fellow at MMLab@NTU on Large Multi-Modality Models for Perception and Generation.
Awesome_Matching_Pretraining_Transfering. The Paper List of Large Multi-Modality Model (Perception, Generation, Unification), Parameter-Efficient Finetuning, Vision-Language Pretraining, Conventional Image-Text Matching for Preliminary Insight.
446SGRAF. [AAAI2021] The code of “Similarity Reasoning and Filtration for Image-Text Matching”
219UniPT. [CVPR2024] The code of "UniPT: Universal Parallel Tuning for Transfer Learning with Efficient Parameter and Memory"
71RCAR. [TIP2023] The code of “Plug-and-Play Regulators for Image-Text Matching”
34DBL. [TIP2024] The code of “Deep Boosting Learning: A Brand-new Cooperative Approach for Image-Text Matching”
12SHERL. [ECCV2024] The code of "SHERL: Synthesizing High Accuracy and Efficient Memory for Resource-Limited Transfer Learning"
9GSSF. [TIP2024] The code of "GSSF: Generalized Structural Sparse Function for Deep Cross-modal Metric Learning"
8Awesome_Image_Text_Retrieval_Benchmark. The Unified Code of Image-Text Retrieval for Further Exploration.
2