Ph.D. Student at Fudan University, homepage: xinyu1205.github.io
recognize-anything. Open-source and strong foundation image recognition models.
3.7krobust-loss-mlml. Code for paper: Simple and Robust Loss Design for Multi-Label Learning with Missing Labels
52IDEA-pytorch. Code for paper: IDEA: Increasing Text Diversity via Online Multi-Label Recognition for Vision-Language Pre-training [ACM MM2022]
9verl. verl: Volcano Engine Reinforcement Learning for LLMs
6MGPO. High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning
5Implementation-and-hiding-analysis-of-reversible-steganography-algorithm-based-on-deep-learning. Jupyter Notebook
1LLaVA. [NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
1xinyu1205.github.io. HTML
1