Algorithm Engineer at Xiaohongshu (dots) | Ph.D. from MAC Lab, Xiamen University | Multimodal LLMs & Text-to-Image
External-Attention-pytorch. 🍀 Pytorch implementation of various Attention Mechanisms, MLP, Re-parameter, Convolution, which is helpful to further understand papers.⭐⭐⭐
12kFightingCV-Paper-Reading. ⭐⭐⭐FightingCV Paper Reading, which helps you understand the most advanced research work in an easier way 🍀 🍀 🍀
820X-Dreamer. A pytorch implementation of “X-Dreamer: Creating High-quality 3D Content by Bridging the Domain Gap Between Text-to-2D and Text-to-3D Generation”
75xmu-xiaoma666.
35RepMLP-pytorch. Pytorch implement ion of RepMLP
30LSTNet. Towards Local Visual Modeling for Image Captioning
30X-Mesh. A pytorch implementation of “ X-Mesh: Towards Fast and Accurate Text-driven 3D Stylization via Dynamic Textual Guidance”
29Multimodal-Open-O1. Multimodal Open-O1 (MO1) is designed to enhance the accuracy of inference models by utilizing a novel prompt-based approach. This tool works locally and aims to create inference chains akin to those used by OpenAI-o1, but with localized processing power.
28SDATR. Official Code for "Knowing what it is: Semantic-enhanced Dual Attention Transformer" (TMM2022)
19ECCV2022-Paper-List. ECCV2022-Paper-List
19ImageCaptionMetrics. This repository contains 2 tools: - A py3 Lib for NLP & image-caption metrics - Code for a two-tailed t-test with paired samples. It will reveals whether the difference of two results is significant. In this code, we complete evaluation code for Spice details(*i.e.*,Object, Relation, Attribute, Color, Count, and Size ).
18vMLLM. The official repository for “vMLLM: Boosting Multi-modal Large Language Model with Enhanced Visual Features”.
13CVAlgorithm. CV面试中的常见算法
8yoloair. 🔥🔥🔥YOLOAir:Including YOLOv5, YOLOv7, Transformer, YOLOX, YOLOR and other networks... Support to improve backbone, head, loss, IoU, NMS...The original version was created based on YOLOv5
7Visualizer. helper tools for attention visualization in deep learning
7MFM. An official implementation for "Knowing What to Learn: A Metric-Oriented Focal Mechanism for Image Captioning"
6MLP-Mixer-pytorch. Unofficial implementation of MLP-Mixer: An all-MLP Architecture for Vision
6Leetcode_diary. Leetcode is all you need
5Pytorch-Image-Classification. Pytorch-Image-Classification
4ECCV2022-Papers-with-Code. ECCV 2022 论文开源项目合集,同时欢迎各位大佬提交issue,分享ECCV 2020开源项目
4DTNet. The official repository for “Image Captioning via Dynamic Path Customization”.
3ECCV2022-Papers-with-Code-Demo. 收集 ECCV 最新的成果,包括论文、代码和demo视频等,欢迎大家推荐!
3Awesome-Model-Pytorch. pytorch implementation of deep learning models
2ECCV2022-Paper-Code-Interpretation. ECCV2022 论文/代码/解读合集,极市团队整理
2CoP. Python
1LLM-MPI. Python
1Beat. Python
1