This is your work, valued
PhD Student of East China University of Science and Technology
SEMat. Python
CMDA. [ICCV23] Official Implementation of CMDA: Cross-Modality Domain Adaptation for Nighttime Semantic Segmentation
MADM. [NIPS24] Official Implementation of Unsupervised Modality Adaptation with Text-to-Image Diffusion Models for Semantic Segmentation
TGVFM. Python
sam-3d-objects. SAM 3D Objects
3DEditFormer. Jupyter Notebook
SDMatte. Official code for our ICCV2025 paper "SDMatte: Grafting Diffusion Models for Interactive Matting"
dinov3. Reference PyTorch implementation and models for DINOv3
torch_flops. A library for calculating the FLOPs in the forward() process based on torch.fx
SoftCoT. ACL'2025: SoftCoT: Soft Chain-of-Thought for Efficient Reasoning with LLMs. and preprint: SoftCoT++: Test-Time Scaling with Soft Chain-of-Thought Reasoning
Step1X-Edit. A SOTA open-source image editing model, which aims to provide comparable performance against the closed-source models like GPT-4o and Gemini 2 Flash.
Direct3D-S2. [NeurIPS 2025] Direct3D‑S2: Gigascale 3D Generation Made Easy with Spatial Sparse Attention
ShapeLLM-Omni. [NeurIPS 2025 Spotlight] A Native Multimodal LLM for 3D Generation and Understanding
SplatFormer. [ICLR' 25] SplatFormer: Point Transformer for Robust 3D Gaussian Splatting
StableAnimator. [CVPR2025] We present StableAnimator, the first end-to-end ID-preserving video diffusion framework, which synthesizes high-quality videos without any post-processing, conditioned on a reference image and a sequence of poses.
VideoPainter. [SIGGRAPH2025] Official repo for paper "Any-length Video Inpainting and Editing with Plug-and-Play Context Control"
VACE. [ICCV 2025] Official implementations for paper: VACE: All-in-One Video Creation and Editing
DepthMaster. Official implementation of "DepthMaster: Taming Diffusion Models for Monocular Depth Estimation".
DSEC. Python
Depth-Anything-V2. [NeurIPS 2024] Depth Anything V2. A More Capable Foundation Model for Monocular Depth Estimation
Algorithm_Interview_Notes-Chinese. 2018/2019/校招/春招/秋招/自然语言处理(NLP)/深度学习(Deep Learning)/机器学习(Machine Learning)/C/C++/Python/面试笔记,此外,还包括创建者看到的所有机器学习/深度学习面经中的问题。 除了其中 DL/ML 相关的,其他与算法岗相关的计算机知识也会记录。 但是不会包括如前端/测试/JAVA/Android等岗位中有关的问题。
academicpages.github.io. Github Pages template based upon HTML and Markdown for personal, portfolio-based websites.
ControlNet. Let us control diffusion models!
ODISE. Official PyTorch implementation of ODISE: Open-Vocabulary Panoptic Segmentation with Text-to-Image Diffusion Models [CVPR 2023 Highlight]
rpg_vid2e. Open source implementation of CVPR 2020 "Video to Events: Recycling Video Dataset for Event Cameras"
ChuanhuChatGPT. GUI for ChatGPT API and many LLMs. Supports agents, file-based QA, GPT finetuning and query with web search. All with a neat UI.
SePiCo. [TPAMI 2023 ESI Highly Cited Paper] SePiCo: Semantic-Guided Pixel Contrast for Domain Adaptive Semantic Segmentation https://arxiv.org/abs/2204.08808
rtl8822bu. RTL8822BU Wireless Driver for Linux