Hangzhou, China

xmu-xiaoma666

Elite
@xmu-xiaoma666

Algorithm Engineer at Xiaohongshu (dots) | Ph.D. from MAC Lab, Xiamen University | Multimodal LLMs & Text-to-Image

External-Attention-pytorch. 🍀 Pytorch implementation of various Attention Mechanisms, MLP, Re-parameter, Convolution, which is helpful to further understand papers.⭐⭐⭐

12k

FightingCV-Paper-Reading. ⭐⭐⭐FightingCV Paper Reading, which helps you understand the most advanced research work in an easier way 🍀 🍀 🍀

820

X-Dreamer. A pytorch implementation of “X-Dreamer: Creating High-quality 3D Content by Bridging the Domain Gap Between Text-to-2D and Text-to-3D Generation”

75

xmu-xiaoma666.

35

RepMLP-pytorch. Pytorch implement ion of RepMLP

30

LSTNet. Towards Local Visual Modeling for Image Captioning

30

X-Mesh. A pytorch implementation of “ X-Mesh: Towards Fast and Accurate Text-driven 3D Stylization via Dynamic Textual Guidance”

29

Multimodal-Open-O1. Multimodal Open-O1 (MO1) is designed to enhance the accuracy of inference models by utilizing a novel prompt-based approach. This tool works locally and aims to create inference chains akin to those used by OpenAI-o1, but with localized processing power.

28

SDATR. Official Code for "Knowing what it is: Semantic-enhanced Dual Attention Transformer" (TMM2022)

19

ECCV2022-Paper-List. ECCV2022-Paper-List

19

ImageCaptionMetrics. This repository contains 2 tools: - A py3 Lib for NLP & image-caption metrics - Code for a two-tailed t-test with paired samples. It will reveals whether the difference of two results is significant. In this code, we complete evaluation code for Spice details(*i.e.*,Object, Relation, Attribute, Color, Count, and Size ).

18

vMLLM. The official repository for “vMLLM: Boosting Multi-modal Large Language Model with Enhanced Visual Features”.

13

CVAlgorithm. CV面试中的常见算法

8

yoloair. 🔥🔥🔥YOLOAir:Including YOLOv5, YOLOv7, Transformer, YOLOX, YOLOR and other networks... Support to improve backbone, head, loss, IoU, NMS...The original version was created based on YOLOv5

7

Visualizer. helper tools for attention visualization in deep learning

7

MFM. An official implementation for "Knowing What to Learn: A Metric-Oriented Focal Mechanism for Image Captioning"

6

MLP-Mixer-pytorch. Unofficial implementation of MLP-Mixer: An all-MLP Architecture for Vision

6

Leetcode_diary. Leetcode is all you need

5

Pytorch-Image-Classification. Pytorch-Image-Classification

4

ECCV2022-Papers-with-Code. ECCV 2022 论文开源项目合集,同时欢迎各位大佬提交issue,分享ECCV 2020开源项目

4

DTNet. The official repository for “Image Captioning via Dynamic Path Customization”.

3

ECCV2022-Papers-with-Code-Demo. 收集 ECCV 最新的成果,包括论文、代码和demo视频等,欢迎大家推荐!

3

Awesome-Model-Pytorch. pytorch implementation of deep learning models

2

ECCV2022-Paper-Code-Interpretation. ECCV2022 论文/代码/解读合集,极市团队整理

2

CoP. Python

1

LLM-MPI. Python

1

Beat. Python

1
27
Apply