Sunnyvale, CA

Peng "Richard" Xia

Expert
@richard-peng-xia

awesome-multimodal-in-medical-imaging. A collection of resources on applications of multi-modal learning in medical imaging.

973

MMed-RAG. [ICLR'25] MMed-RAG: Versatile Multimodal RAG System for Medical Vision Language Models

337

RULE. [EMNLP'24] RULE: Reliable Multimodal RAG for Factuality in Medical Vision Language Models

98

CARES. [NeurIPS'24] CARES: A Comprehensive Benchmark of Trustworthiness in Medical Vision Language Models

79

LMPT. [ACLW'24] LMPT: Prompt Tuning with Class-Specific Embedding Loss for Long-tailed Multi-Label Visual Recognition

58

HGCLIP. [COLING'25] HGCLIP: Exploring Vision-Language Models with Graph Representations for Hierarchical Understanding

45

Chinese-Noisy-Text. This repository stores the code of the data augmentation method from Chinese word and character levels, which adds noise to words and characters in redundant, missing, selection and ordering respectively.

22

KD-CGEC. Code for Chinese grammatical error correction based on knowledge distillation

11

Harrypotter_Knowledge_Graph. 针对哈利波特这一主题在网上进行数据的搜集、筛选与清洗,并结合在知识表示课上所学到的知识,制作了一个哈利波特主题型知识图谱。附上了我们的多媒体资源,以及我们制作过程中的临时文件,以及最终的 json 文件。

4

Awesome-Med-Agent.

3

VLM_survey. Vision-Language Models for Vision Tasks: A Survey

1

DIGIX_Article_Quality_Discrimination. This repository stores the code for 2021 DIGIX全球校园AI算法精英大赛 赛题二 基于多模型迁移预训练文章质量判别.

1

LLMAgentPapers. Must-read Papers on LLM Agents.

1

Trigger-Identification. This repository stores the implementation of trigger identification task on Shanghai-HK Interdisciplinary Shared Tasks (2022).

1
14
Apply