This is your work, valued
Assistant Prof. @ CAU. Please refer to my personal homepage: https://sites.google.com/view/ozbro/
XVFI. [ICCV 2021, Oral 3%] Official repository of XVFI
★ 311DeMFI. [ECCV 2022] Official repository of DeMFI.
★ 84FISR. [AAAI 2020] Official repository of FISR.
★ 76JSI-GAN. [AAAI 2020] Official repository of JSI-GAN.
★ 60JihyongOh.
★ 3void-model. Python
★ 1.9kAuto-Resubmit. 🚀 Automated & lossless LaTeX paper migration tool. Instantly convert your Overleaf source between top-tier AI conference templates (NeurIPS, ICLR, ACL等). 一键无损转换顶会论文格式!解决转投时的排版折磨,完美保留公式、图表与引用,让科研人员专注研究而非排版。
★ 229awesome-mechanistic-interpretability. A carefully curated collection of high-quality libraries, projects, tutorials, research papers, and other essential resources focused on Mechanistic Interpretability, a growing subfield in machine learning interpretability research that aims to reverse-engineer neural networks into understandable computational components.
★ 1302025-Arxiv-Paper-List-Gaussian-Splatting. 📚 2025 Gaussian Splatting ArXiv Paper List — Updated Daily
★ 119Step1X-Edit. A SOTA open-source image editing model, which aims to provide comparable performance against the closed-source models like GPT-4o and Gemini 2 Flash.
★ 2.2kMineRender. Quick, Easy, Interactive 3D/2D Renders of Minecraft
★ 154Awesome-Controllable-Video-Generation. [ArXiv 2025] A survey about controllable video generation: This repo is the official awesome of "Controllable video generation: A survey"
★ 760WithAnyone. ✨ [ICLR'26] WithAnyone is capable of generating high-quality, controllable, and ID consistent images
★ 572Awesome-Multimodal-Large-Language-Models. :sparkles::sparkles:Latest Advances on Multimodal Large Language Models
★ 18kAll-in-One-Image-Restoration-Survey. [IEEE TPAMI 2025] A Survey on All-in-One Image Restoration: Taxonomy, Evaluation and Future Trends
★ 579map-anything. MapAnything: Universal Feed-Forward Metric 3D Reconstruction
★ 3.6kpdfdiff. Command-line tool to inspect the difference between (the text in) two PDF files
★ 260OminiControl. [ICCV 2025 Highlight] OminiControl: Minimal and Universal Control for Diffusion Transformer
★ 1.9kHuMo. HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning
★ 1.3kAwesome-Diffusion-Models-in-Medical-Imaging. Diffusion Models in Medical Imaging (Published in Medical Image Analysis Journal)
★ 2.1kWan2.1. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kBiM-VFI. [CVPR 2025] Official repository of BiM-VFI
★ 60fast-gaussian-rasterization. A geometry-shader-based, global CUDA sorted high-performance 3D Gaussian Splatting rasterizer. Can achieve a 5-10x speedup in rendering compared to the vanialla diff-gaussian-rasterization.
★ 1.2kMoDec-GS. [CVPR2025] MoDec-GS: Global-to-Local Motion Decomposition and Temporal Interval Adjustment for Compact Dynamic 3D Gaussian Splatting
★ 422024-NeurIPS-AverNet. Code for the paper "AverNet: All-in-one Video Restoration for Time-varying Unknown Degradations" (NeurIPS 2024)
★ 36cs231n.github.io. Public facing notes page
★ 11kbrush. 3D Reconstruction for all
★ 4.9kSana. SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
★ 8.6kttt-video-dit. Official PyTorch implementation of One-Minute Video Generation with Test-Time Training
★ 2.4kfvd-comparison. Comparison between Frechet Video Distance implementation from StyleGAN-V and the original paper
★ 129common_metrics_on_video_quality. You can easily calculate FVD, PSNR, SSIM, LPIPS for evaluating the quality of generated or predicted videos.
★ 583Awesome-CVPR2026-CVPR2025-CVPR2024-CVPR2021-CVPR2020-Low-Level-Vision. A Collection of Papers and Codes for CVPR2026/CVPR2025/CVPR2024/CVPR2021/CVPR2020 Low Level Vision
★ 1.7kawesome-diffusion-models-in-low-level-vision. A Repository for Diffusion-Model-related Papers in Low-level Vision
★ 556video-restoration-arxiv-daily. 【This project is no longer updated】
★ 47CogVideo. text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
★ 13kAwesome-Image-Editing. A Survey of Image Editing
★ 469IQA-PyTorch. 🔎 🖼️ 🔥PyTorch Toolbox for Image Quality Assessment, including PSNR, SSIM, LPIPS, FID, NIQE, NRQM(Ma), MUSIQ, TOPIQ, NIMA, DBCNN, BRISQUE, PI and more...
★ 3.3kDiffusion-Assignment1-DDPM. Jupyter Notebook
★ 46svd_keyframe_interpolation. Python
★ 297sapiens. High-resolution models for human tasks.
★ 5.4kAwesome-Mamba-in-Low-Level-Vision. A paper list of recent mamba efforts for low-level vision.
★ 469Mamba-in-CV. A paper list of some recent Mamba-based CV works.
★ 491IF. Python
★ 7.8kml-4m. 4M: Massively Multimodal Masked Modeling
★ 1.8kAwesome-LLMs-meet-Multimodal-Generation. 🔥🔥🔥 A curated list of papers on LLMs-based multimodal generation (image, video, 3D and audio).
★ 551eisai-anime-interpolator. ECCV2022: Improving the Perceptual Quality of 2D Animation Interpolation
★ 124frame-interpolation. FILM: Frame Interpolation for Large Motion, In ECCV 2022.
★ 3.1kUPR-Net. Official implementation of our CVPR2023 paper "A Unified Pyramid Recurrent Network for Video Frame Interpolation"
★ 126EMA-VFI. [CVPR 2023] Extracting Motion and Appearance via Inter-Frame Attention for Efficient Video Frame Interpolatio
★ 498AMT. Official code for "AMT: All-Pairs Multi-Field Transforms for Efficient Frame Interpolation" (CVPR2023)
★ 270Ranni. Python
★ 237flowmap. [3DV 2025] Code for "FlowMap: High-Quality Camera Poses, Intrinsics, and Depth via Gradient Descent" by Cameron Smith*, David Charatan*, Ayush Tewari, and Vincent Sitzmann
★ 981dust3r. DUSt3R: Geometric 3D Vision Made Easy
★ 7.3kSUPIR. SUPIR aims at developing Practical Algorithms for Photo-Realistic Image Restoration In the Wild. Our new online demo is also released at suppixel.ai.
★ 5.6kAwesome-Video-Diffusion. A curated list of recent diffusion models for video generation, editing, and various other applications.
★ 5.7kVGen. Official repo for VGen: a holistic video generation ecosystem for video generation building on diffusion models
★ 3.2kOpen-Sora. Open-Sora: Democratizing Efficient Video Production for All
★ 29kawesome-multimodal-ml. Reading list for research topics in multimodal machine learning
★ 6.9kAwesome-LLM. Awesome-LLM: a curated list of Large Language Model
★ 27kLAVIS. LAVIS - A One-stop Library for Language-Vision Intelligence
★ 11kawesome-3d-reconstruction-papers. A collection of 3D reconstruction papers in the deep learning era.
★ 910stable-diffusion. A latent text-to-image diffusion model
★ 73kcamp_zipnerf. Python
★ 726InstructIR. [ECCV 2024] InstructIR: High-Quality Image Restoration Following Human Instructions https://huggingface.co/spaces/marcosv/InstructIR
★ 742SEAL. ICLR 2024 (Spotlight) - SEAL: A Framework for Systematic Evaluation of Real-World Super-Resolution
★ 56FastChat. An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
★ 40kSegment-Everything-Everywhere-All-At-Once. [NeurIPS 2023] Official implementation of the paper "Segment Everything Everywhere All at Once"
★ 4.8kdot. Dense Optical Tracking: Connecting the Dots
★ 322omnimotion. Python
★ 2.3ktapnet. Tracking Any Point (TAP)
★ 2kTracking-Anything-with-DEVA. [ICCV 2023] Tracking Anything with Decoupled Video Segmentation
★ 1.5kKDSR. This project is the official implementation of 'Knowledge Distillation based Degradation Estimation for Blind Super-Resolution', ICLR2023
★ 130StableSR. [IJCV2024] Exploiting Diffusion Prior for Real-World Image Super-Resolution
★ 2.7kAGIN. [IEEE TCSVT'24] Study of Subjective and Objective Naturalness Assessment of AI-Generated Images
★ 38Q-Bench. ①[ICLR2024 Spotlight] (GPT-4V/Gemini-Pro/Qwen-VL-Plus+16 OS MLLMs) A benchmark for multi-modality LLMs (MLLMs) on low-level vision and visual quality assessment.
★ 287daclip-uir. [ICLR 2024] Controlling Vision-Language Models for Universal Image Restoration. 5th place in the NTIRE 2024 Restore Any Image Model in the Wild Challenge.
★ 816ReplaceAnything.
★ 2.4kSpacetimeGaussians. [CVPR 2024] Spacetime Gaussian Feature Splatting for Real-Time Dynamic View Synthesis
★ 826AnyText. Official implementation code of the paper <AnyText: Multilingual Visual Text Generation And Editing>
★ 4.9kOSRT. Official code of OSRT: Omnidirectional Image Super-Resolution with Distortion-aware Transformer
★ 80SeeSR. [CVPR2024] SeeSR: Towards Semantics-Aware Real-World Image Super-Resolution
★ 650Awesome-Text-to-Image. (ෆ`꒳´ෆ) A Survey on Text-to-Image Generation/Synthesis.
★ 2.4kDeepCache. [CVPR 2024] DeepCache: Accelerating Diffusion Models for Free
★ 969FollowYourHandle. [WACV 2025] Follow-Your-Handle: This repo is the official implementation of "MagicStick: Controllable Video Editing via Control Handle Transformations"
★ 99PromptSR. PyTorch code for our paper "Image Super-Resolution with Text Prompt Diffusion"
★ 126mip-splatting. [CVPR'24 Best Student Paper] Mip-Splatting: Alias-free 3D Gaussian Splatting
★ 1.5k4d-gaussian-splatting. [ICLR 2024] Real-time Photorealistic Dynamic Scene Representation and Rendering with 4D Gaussian Splatting
★ 1kDynamic3DGaussians. Python
★ 2.3kgaussian-splatting. Original reference implementation of "3D Gaussian Splatting for Real-Time Radiance Field Rendering"
★ 23k4DGaussians. [CVPR 2024] 4D Gaussian Splatting for Real-Time Dynamic Scene Rendering
★ 3.9kStableVSR. [ECCV 2024] Enhancing Perceptual Quality in Video Super-Resolution through Temporally-Consistent Detail Synthesis using Diffusion Models
★ 178EfficientSAM. EfficientSAM: Leveraged Masked Image Pretraining for Efficient Segment Anything
★ 2.5kTaleCrafter. [SIGGRAPH Asia 2023] An interactive story visualization tool that support multiple characters
★ 268InstaFlow. :zap: InstaFlow! One-Step Stable Diffusion with Rectified Flow (ICLR 2024)
★ 1.4kVideoCrafter. VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
★ 5.1kDemoFusion. Let us democratise high-resolution generation! (CVPR 2024)
★ 2kawesome-3D-gaussian-splatting. Curated list of papers and resources focused on 3D Gaussian Splatting, intended to keep pace with the anticipated surge of research in the coming months.
★ 8.8kgenerative-models. Generative Models by Stability AI
★ 27kRectifiedFlow. Official Implementation of Rectified Flow (ICLR2023 Spotlight)
★ 1.6kLoRA. Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"
★ 14kAwesome-diffusion-model-for-image-processing. one summary of diffusion-based image processing, including restoration, enhancement, coding, quality assessment
★ 956DiffBIR. [ECCV 2024] codes of DiffBIR: Towards Blind Image Restoration with Generative Diffusion Prior
★ 4.1klatent-consistency-model. Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference
★ 4.6kawesome-openai-vision-api-experiments. Must-have resource for anyone who wants to experiment with and build on the OpenAI vision API 🔥
★ 1.7kCVinW_Readings. A collection of papers on the topic of ``Computer Vision in the Wild (CVinW)''
★ 1.4kdenoising-diffusion-pytorch. Implementation of Denoising Diffusion Probabilistic Model in Pytorch
★ 11kmsth. NeurIPS2023 Spotlight: Masked Space-time Hashing
★ 92LIQE. [CVPR2023] Blind Image Quality Assessment via Vision-Language Correspondence: A Multitask Learning Perspective
★ 241zero123plus. Code repository for Zero123++: a Single Image to Consistent Multi-view Diffusion Base Model.
★ 2.1kStableVideo. [ICCV 2023] StableVideo: Text-driven Consistency-aware Diffusion Video Editing
★ 1.4kRerender_A_Video. [SIGGRAPH Asia 2023] Rerender A Video: Zero-Shot Text-Guided Video-to-Video Translation
★ 3kfactor-fields. [SIGGRAPH 2023] We provide a unified formula for neural fields (Factor Fields) and a novel dictionary factorization (Dictionary Fields)
★ 213RelayDiffusion. The official implementation of "Relay Diffusion: Unifying diffusion process across resolutions for image synthesis" [ICLR 2024 Spotlight]
★ 314textual_inversion. Jupyter Notebook
★ 3.1kStrivec. Official code for the paper: Strivec (ICCV2023)
★ 146S3IM-Neural-Fields. [ICCV 2023] Pytorch implementation of "S3IM: Stochastic Structural SIMilarity and Its Unreasonable Effectiveness for Neural Fields".
★ 241RDDM. CVPR 2024: Residual Denoising Diffusion Models
★ 579MPI-Flow. Official Code of "[ICCV 2023] MPI-Flow: Learning Realistic Optical Flow with Multiplane Images"
★ 65PCV. Open source Python module for computer vision
★ 2kdiffusionerf. [CVPR 2023] DiffusioNeRF: Regularizing Neural Radiance Fields with Denoising Diffusion Models
★ 308ResFields. [ICLR 2024 Spotlight ✨] ResFields: Residual Neural Fields for Spatiotemporal Signals
★ 168AvatarMAV. A PyTorch implementation of "AvatarMAV: Fast 3D Head Avatar Reconstruction Using Motion-Aware Neural Voxels"
★ 100LatentAvatar. A PyTorch implementation of "LatentAvatar: Learning Latent Expression Code for Expressive Neural Head Avatar"
★ 107fourier-feature-networks. Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional Domains
★ 1.4ktilted. Canonical Factors for Hybrid Neural Fields @ ICCV 2023
★ 107fc-clip. [NeurIPS 2023] This repo contains the code for our paper Convolutions Die Hard: Open-Vocabulary Segmentation with Single Frozen Convolutional CLIP
★ 345DeSAM. [MICCAI 2024] The official repository for DeSAM: Decoupled Segment Anything Model for Generalizable Medical Image Segmentation.
★ 134DiffIR. This project is the official implementation of 'Diffir: Efficient diffusion model for image restoration', ICCV2023
★ 611Awesome-ICCV2023-Low-Level-Vision. A Collection of Papers and Codes in ICCV2023/2021 about low level vision
★ 194Awesome-CVPR2024-Low-Level-Vision. A Collection of Papers and Codes in CVPR2023/2022 about low level vision
★ 660NeRFCapture. An iOS app that collects/streams posed images for NeRFs using ARKit
★ 303Semantic-SAM. [ECCV 2024] Official implementation of the paper "Semantic-SAM: Segment and Recognize Anything at Any Granularity"
★ 2.9kawesome-generative-ai. A curated list of modern Generative Artificial Intelligence projects and services
★ 12kawesome-generative-ai. A curated list of Generative AI tools, works, models, and references
★ 3.5knope-nerf. (CVPR 2023) NoPe-NeRF: Optimising Neural Radiance Field with No Pose Prior
★ 409segment-anything. The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 55ktorch_efficient_distloss. Efficient distortion loss with O(n) realization.
★ 132FastSAM. Fast Segment Anything
★ 8.4kdynamic-video-depth. Code for the SIGGRAPH 2021 paper "Consistent Depth of Moving Objects in Video".
★ 274HQ-50K. HQ-50K: A Large-scale, High-quality Dataset for Image Restoration
★ 95IDM. Python
★ 335CiaoSR. PyTorch implementation of "CiaoSR: Continuous Implicit Attention-in-Attention Network for Arbitrary-Scale Image Super-Resolution (CVPR2023)"
★ 136HI-Diff. [NeurIPS'23] Hierarchical Integration Diffusion Model for Realistic Image Deblurring
★ 251tiny-cuda-nn. Lightning fast C++/CUDA neural network framework
★ 4.5kDiffPIR. "Denoising Diffusion Models for Plug-and-Play Image Restoration", Yuanzhi Zhu, Kai Zhang, Jingyun Liang, Jiezhang Cao, Bihan Wen, Radu Timofte, Luc Van Gool.
★ 497image-restoration-sde. Image Restoration with Mean-Reverting Stochastic Differential Equations, ICML 2023. Winning solution of the NTIRE 2023 Image Shadow Removal Challenge.
★ 722latent-diffusion. High-Resolution Image Synthesis with Latent Diffusion Models
★ 14kdynibar. Implementation of DynIBaR Neural Dynamic Image-Based Rendering (CVPR 2023)
★ 817lietorch. Cuda
★ 851zipnerf-pytorch. Unofficial implementation of ZipNeRF
★ 853HexPlane. Official code for CVPR 2023 Paper, HexPlane: A Fast Representation for Dynamic Scenes
★ 311sdfstudio. A Unified Framework for Surface Reconstruction
★ 2.1kcloud_nerf. Python
★ 49DP-NeRF. [CVPR 2023 - Official] DP-NeRF: Deblurred Neural Radiance Field with Physical Scene Priors
★ 86TensoRF. [ECCV 2022] Tensorial Radiance Fields, a novel approach to model and reconstruct radiance fields
★ 1.2kpython-for-coding-test. [한빛미디어] "이것이 취업을 위한 코딩 테스트다 with 파이썬" 전체 소스코드 저장소입니다.
★ 2.4kai-tech-interview. 👩💻👨💻 AI 엔지니어 기술 면접 스터디 (⭐️ 2k+)
★ 2.3kgin-config. Gin provides a lightweight configuration framework for Python
★ 2.2kautoml. Google Brain AutoML
★ 6.5kpytorch-image-models. The largest collection of PyTorch image encoders / backbones. Including train, eval, inference, export scripts, and pretrained weights -- ResNet, ResNeXT, EfficientNet, NFNet, Vision Transformer (ViT), MobileNetV4, MobileNet-V3 & V2, RegNet, DPN, CSPNet, Swin Transformer, MaxViT, CoAtNet, ConvNeXt, and more
★ 37kLLFF. Code release for Local Light Field Fusion at SIGGRAPH 2019
★ 1.7khyperreel. Code release for HyperReel: High-Fidelity 6-DoF Video with Ray-Conditioned Sampling
★ 481RND. [AAAI 2023 Oral] GAN Prior based Null-Space Learning for Consistent Super-Resolution
★ 72multinerf. A Code Release for Mip-NeRF 360, Ref-NeRF, and RawNeRF
★ 3.8kdycheck. Official JAX Implementation of Monocular Dynamic View Synthesis: A Reality Check (NeurIPS 2022)
★ 217nerfstudio. A collaboration friendly studio for NeRFs
★ 12kFTVSR. [ECCV'22] FTVSR: Learning Spatiotemporal Frequency-Transformer for Compressed Video Super-Resolution
★ 175DAVSR. PyTorch implementation of "Towards Interpretable Video Super-Resolution via Alternating Optimization (ECCV2022)"
★ 74pytorch_forward_forward. Implementation of Hinton's forward-forward (FF) algorithm - an alternative to back-propagation
★ 1.5kDASS. An official PyTorch implementation of "Bi-directional Contrastive Learning for Domain Adaptive Semantic Segmentation", ECCV 2022.
★ 44ALIFE. An official implementation of "ALIFE: Adaptive Logit Regularizer and Feature Replay for Incremental Semantic Segmentation" (NeurIPS 2022) in PyTorch.
★ 49OIMNetPlus. An official PyTorch implementation of "OIMNet++: Prototypical Normalization and Localization-aware Learning for Person Search", ECCV 2022.
★ 46DKD. An official implementation of "Decomposed Knowledge Distillation for Class-incremental Semantic Segmentation" (NeurIPS 2022) in PyTorch.
★ 68gradient-descent-the-ultimate-optimizer. Code for our NeurIPS 2022 paper
★ 370DiffusionDet. [ICCV2023 Best Paper Finalist] PyTorch implementation of DiffusionDet (https://arxiv.org/abs/2211.09788)
★ 2.3kstarlight_denoising. Jupyter Notebook
★ 107Awesome-Diffusion-Models. A collection of resources and papers on Diffusion Models
★ 12ktech-interview-handbook. Curated coding interview preparation materials for busy software engineers
★ 141kAppearanceFreeActionRecognition. [ECCV 2022] Is Appearance Free Action Recognition Possible?
★ 59pytorch-meta. A collection of extensions and data-loaders for few-shot learning & meta-learning in PyTorch
★ 2.1kresume. Software developer resume in Latex
★ 6.9knerf-factory. An awesome PyTorch NeRF library
★ 1.3kpips. Particle Video Revisited
★ 603nsff_pl. Neural Scene Flow Fields using pytorch-lightning, with potential improvements
★ 224d2nerf. Jupyter Notebook
★ 192NeRF-SR. NeRF-SR: High-Quality Neural Radiance Fields using Supersampling
★ 147SCNeRF. [ICCV21] Self-Calibrating Neural Radiance Fields
★ 471Deblur-NeRF. Python
★ 285StyleNeRF. This is the open source implementation of the ICLR2022 paper "StyleNeRF: A Style-based 3D-Aware Generator for High-resolution Image Synthesis"
★ 949NeuRay. [CVPR2022] Neural Rays for Occlusion-aware Image-based Rendering
★ 419Neural_3D_Video. The repository for CVPR 2022 Paper "Neural 3D Video Synthesis"
★ 361awesome-neural-rendering. Resources of Neural Rendering
★ 2.4kRSTT. Official pytorch implementation of paper "RSTT: Real-time Spatial Temporal Transformer for Space-Time Video Super-Resolution"
★ 143TTVSR. [CVPR'22 Oral] TTVSR: Learning Trajectory-Aware Transformer for Video Super-Resolution
★ 222ELAN. [ECCV2022] Efficient Long-Range Attention Network for Image Super-resolution
★ 239M2M_VFI. Many-to-many Splatting for Efficient Video Frame Interpolation
★ 93google-research. Google Research
★ 38kpoolformer. PoolFormer: MetaFormer Is Actually What You Need for Vision (CVPR 2022 Oral)
★ 1.4kSinNeRF. [ECCV 2022] "SinNeRF: Training Neural Radiance Fields on Complex Scenes from a Single Image", Dejia Xu, Yifan Jiang, Peihao Wang, Zhiwen Fan, Humphrey Shi, Zhangyang Wang
★ 330einops. Flexible and powerful tensor operations for readable and reliable code (for pytorch, jax, TF and others)
★ 9.6kBasicVSR_PlusPlus. Official repository of "BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and Alignment"
★ 812InfoNeRF. Python
★ 144DSNeRF. Code release for DS-NeRF (Depth-supervised Neural Radiance Fields)
★ 785vit-tensorflow. Vision Transformer Cookbook with Tensorflow
★ 342pytorch_diffusion. PyTorch reimplementation of Diffusion Models
★ 585