PhD student @ University of Edinburgh.
awesome-self-supervised-multimodal-learning. [T-PAMI] A curated list of self-supervised multimodal learning resources.
278VLGuard. [ICML 2024] Safety Fine-Tuning at (Almost) No Cost: A Baseline for Vision Large Language Models.
90MEDFAIR. [ICLR 2023 spotlight] MEDFAIR: Benchmarking Fairness for Medical Imaging
74VL-ICL. [ICLR 2025] VL-ICL Bench: The Devil in the Details of Multimodal In-Context Learning
69conST. conST: an interpretable multi-modal contrastive learning framework for spatial transcriptomics
29FoolyourVLLMs. [ICML 2024] Fool Your (Vision and) Language Model With Embarrassingly Simple Permutations
15fpga-camera. camera OV2640 on FPGA Nexys4
12MIRB. Benchmarking Multi-Image Understanding in Vision and Language Models
11FPGA-CPU54. MIPS CPU on FPGA Nexys4 (54 intrs)
7FPGA-CPU. MIPS cpu on FPGA Nexys4 (31 instrs )
4