This is your work, valued
Anime2Sketch. A sketch extractor for anime/illustration.
★ 2.1kZooming-Slow-Mo-CVPR-2020. Fast and Accurate One-Stage Space-Time Video Super-Resolution (accepted in CVPR 2020)
★ 932AODA. Official implementation of "Adversarial Open Domain Adaptation for Sketch-to-Photo Synthesis"(WACV 2022/CVPRW 2021)
★ 86mukosame.github.io. HTML
★ 78Face-Annotation-Tool. Help you manually label facial landmarks.
★ 44Print-Defect-Detection-and-Quality-Assessment. A collection of print defect detection and quality assessment papers with code (constantly updating)
★ 29haha. 长者语录生成器WP8.1应用
★ 9Chinese-Colors. Chinese traditional color card app for WP8.1
★ 6ImageToolbox. Matlab toolbox of Digital Color Imaging Systems
★ 4video-enhancement. A list of resources for video enhancement, including video super-resolutio, interpolation, denoising, compression artifact removal et al..
★ 4How-to-build-DL-BOX. A tutorial for Deep Learning Beginners: How to build a linux deep learning work station from scratch
★ 3BME595A_DeepLearning. Weekly project for BME595A(Deep Learning) 2018 Fall, Purdue University.
★ 2make-a-texture. Website for Make-A-Texture
★ 1EdgeGAN. Code for paper "SketchyCOCO: Image Generation from Freehand Scene Sketches"
★ 1ucore_os_lab. os kernel labs for operating systems course in Tsinghua University.
★ 1pulse. PULSE: Self-Supervised Photo Upsampling via Latent Space Exploration of Generative Models
★ 1style2paints. sketch + style = paints :art:
★ 1Boosting-High-Level-Vision-with-Joint-Compression-Artifacts-Reduction-and-Super-Resolution. ICPR-2020
★ 1RIP-Peak. A matlab script to find the peak and calculate NA from Preform RIP file, PK2600
★ 1animeGAN. A simple PyTorch Implementation of Generative Adversarial Networks, focusing on anime face drawing.
★ 1Time_App. Timer App for windows phone 8.1(finished)
★ 1Deformable-Convolution-V2-PyTorch. Deformable ConvNets V2 (DCNv2) in PyTorch
★ 1Zooming-SlowMo-CVPR-2020. Link to https://github.com/Mukosame/Zooming-Slow-Mo-CVPR-2020
★ 13dcodebench. Benchmarking Agentic Procedural 3D Modeling Via Code
★ 70GLM-5. GLM-5: From Vibe Coding to Agentic Engineering
★ 6.8katlas-lean. ATLAS Autoformalized Textbook Library At Scale
★ 276autoform-bot. Autoform Bot
★ 95repoprover. Research code base for Automatic Textbook Formalization
★ 157Awesome_Think_With_Images. Resources and paper list for "Thinking with Images for LVLMs". This repository accompanies our survey on how LVLMs can leverage visual information for complex reasoning, planning, and generation.
★ 1.5kcube. Roblox Foundation Model for 3D Intelligence
★ 1.2krecastnavigation. Industry-standard navigation-mesh toolset for games
★ 7.8kPartPacker. Efficient Part-level 3D Object Generation via Dual Volume Packing
★ 824FastNoiseLite. Fast Portable Noise Library - C# C++ C Java HLSL GLSL JavaScript Rust Go
★ 3.5kgenesis-world. Simulation platform for general-purpose robotics & embodied AI learning.
★ 30kOctree-GS. [TPAMI 2025] Octree-GS: Towards Consistent Real-time Rendering with LOD-Structured 3D Gaussians
★ 856MeshAnything. [ICLR 2025] From anything to mesh like human artists. Official impl. of "MeshAnything: Artist-Created Mesh Generation with Autoregressive Transformers"
★ 2.3kinfinigen. Infinite Photorealistic Worlds using Procedural Generation
★ 7.2ksd-forge-layerdiffuse. [WIP] Layer Diffusion for WebUI (via Forge)
★ 4.1kPlatoNeRF. PlatoNeRF: 3D Reconstruction in Plato's Cave via Single-View Two-Bounce Lidar
★ 82CAD. CAD: Photorealistic 3D Generation via Adversarial Distillation (CVPR 2024)
★ 129Point-UV-Diffusion. (ICCV2023) This is the official PyTorch implementation of ICCV2023 paper: Texture Generation on 3D Meshes with Point-UV Diffusion
★ 218revisiting-sepconv. an implementation of Revisiting Adaptive Convolutions for Video Frame Interpolation using PyTorch
★ 93SyncDreamer. [ICLR 2024 Spotlight] SyncDreamer: Generating Multiview-consistent Images from a Single-view Image
★ 1kAll-In-One-Deflicker. [CVPR2023] Blind Video Deflickering by Neural Filtering with a Flawed Atlas
★ 761CoDeF. [CVPR'24 Highlight] Official PyTorch implementation of CoDeF: Content Deformation Fields for Temporally Consistent Video Processing
★ 4.8kAnimateDiff. Official implementation of AnimateDiff.
★ 12kgaussian-splatting. Original reference implementation of "3D Gaussian Splatting for Real-Time Radiance Field Rendering"
★ 23kControlNet-v1-1-nightly. Nightly release of ControlNet 1.1
★ 5.2kShadowNeuS. ShadowNeuS: Neural SDF Reconstruction by Shadow Ray Supervision (CVPR 2023)
★ 87NeTF_public. Neural transient field for non-line-of-sight imaging
★ 41DragGAN. Unofficial Implementation of DragGAN - "Drag Your GAN: Interactive Point-based Manipulation on the Generative Image Manifold" (DragGAN 全功能实现,在线Demo,本地部署试用,代码、模型已全部开源,支持Windows, macOS, Linux)
★ 5kfastcomposer. [IJCV] FastComposer: Tuning-Free Multi-Subject Image Generation with Localized Attention
★ 715TemporalKit. An all in one solution for adding Temporal Stability to a Stable Diffusion Render via an automatic1111 extension
★ 2kIF. Python
★ 7.8kTEXTurePaper. Official Implementation for "TEXTure: Text-Guided Texturing of 3D Shapes"
★ 799GRL-Image-Restoration. Python
★ 417ImageReward. [NeurIPS 2023] ImageReward: Learning and Evaluating Human Preferences for Text-to-image Generation
★ 1.7kxatlas. Mesh parameterization / UV unwrapping library
★ 2.5kAutoGPT. AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters.
★ 186kworldsheet. Code release for the paper “Worldsheet Wrapping the World in a 3D Sheet for View Synthesis from a Single Image”, ICCV 2021.
★ 83permuto_sdf. Code for our CVPR'23 paper - "PermutoSDF: Fast Multi-View Reconstruction with Implicit Surfaces using Permutohedral Lattices"
★ 448segment-anything. The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 55keg3d. Python
★ 3.3kMoStGAN-V. [CVPR 2023] Official PyTorch implementation of MoStGAN-V
★ 25360SR-Challenge. NTIRE 2023 Challenge on 360° Omnidirectional Image and Video Super-Resolution
★ 87gigagan-pytorch. Implementation of GigaGAN, new SOTA GAN out of Adobe. Culmination of nearly a decade of research into GANs
★ 1.9kChatGPT-Paper-Reader. This repo offers a simple interface that helps you to read&summerize research papers in pdf format. You can ask some questions after reading. This interface is developed based on openai API and using GPT-3.5-turbo model.
★ 756panic3d-anime-reconstruction. CVPR 2023: PAniC-3D Stylized Single-view 3D Reconstruction from Portraits of Anime Characters
★ 826zero123. Zero-1-to-3: Zero-shot One Image to 3D Object (ICCV 2023)
★ 3.1klayered-neural-atlases. Python
★ 612Tune-A-Video. [ICCV 2023] Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation
★ 4.4kMiDaS. Code for robust monocular depth estimation described in "Ranftl et. al., Towards Robust Monocular Depth Estimation: Mixing Datasets for Zero-shot Cross-dataset Transfer, TPAMI 2022"
★ 5.4kcross-domain-compositing. Jupyter Notebook
★ 186GLIGEN. Open-Set Grounded Text-to-Image Generation
★ 2.2kPerspective-and-Equirectangular. A tool to transfer Perspective between Equirectangular
★ 194MultiDiffusion. Official Pytorch Implementation for "MultiDiffusion: Fusing Diffusion Paths for Controlled Image Generation" presenting "MultiDiffusion" (ICML 2023)
★ 1.1kstable-diffusion-webui. Stable Diffusion web UI
★ 164kk-diffusion. Karras et al. (2022) diffusion models for PyTorch
★ 2.6kT2I-Adapter. T2I-Adapter
★ 3.8kControlNet. Let us control diffusion models!
★ 34kpix2pix3D. pix2pix3D: Generating 3D Objects from 2D User Inputs
★ 1.7kpix2pix-zero. Zero-shot Image-to-Image Translation [SIGGRAPH 2023]
★ 1.1kstable-diffusion. Jupyter Notebook
★ 1.5kAttend-and-Excite. Official Implementation for "Attend-and-Excite: Attention-Based Semantic Guidance for Text-to-Image Diffusion Models" (SIGGRAPH 2023)
★ 770imaginAIry. Pythonic AI generation of images and videos
★ 8.2kX-Decoder. [CVPR 2023] Official Implementation of X-Decoder for generalized decoding for pixel, image and language
★ 1.3kmixbox. Mixbox is a library for natural color mixing based on real pigments.
★ 3.5kinstruct-pix2pix. Python
★ 6.9kSINE. This respository contains the code for the CVPR 2023 paper SINE: SINgle Image Editing with Text-to-Image Diffusion Models.
★ 190CogVideo. text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
★ 13kCogView. Text-to-Image generation. The repo for NeurIPS 2021 paper "CogView: Mastering Text-to-Image Generation via Transformers".
★ 1.8kdiffusers. 🤗 Diffusers: State-of-the-art diffusion models for image and audio generation in PyTorch
★ 1.9kframeintIFE. Official PyTorch implementation of "Frame Interpolation for Dynamic Scenes with Implicit Flow Encoding"
★ 41prompt-to-prompt. Jupyter Notebook
★ 3.5kml-stable-diffusion. Stable Diffusion with Core ML on Apple Silicon
★ 18kVersatile-Diffusion. Versatile Diffusion: Text, Images and Variations All in One Diffusion Model, arXiv 2022 / ICCV 2023
★ 1.3kSCUNet. Practical Blind Denoising via Swin-Conv-UNet and Data Synthesis (Machine Intelligence Research 2023)
★ 817InvokeAI-DiffusionCraftAI. This version of Stable Diffusion features a slick WebGUI, an interactive command-line script that combines text2img and img2img functionality in a "dream bot" style interface, and multiple features and other enhancements. For more info, see the website link below.
★ 56torchdynamo. A Python-level JIT compiler designed to make unmodified PyTorch programs faster.
★ 1.1kgoogle-research. Google Research
★ 38kDreambooth-Stable-Diffusion. Implementation of Dreambooth (https://arxiv.org/abs/2208.12242) with Stable Diffusion
★ 7.7kstable-dreamfusion. Text-to-3D & Image-to-3D & Mesh Exportation with NeRF + Diffusion.
★ 8.8kText2LIVE. Official Pytorch Implementation for "Text2LIVE: Text-Driven Layered Image and Video Editing" (ECCV 2022 Oral)
★ 886stable-diffusion. Jupyter Notebook
★ 1kstable-diffusion. A latent text-to-image diffusion model
★ 73kdiffusers. 🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
★ 34ktext2mesh. 3D mesh stylization driven by a text input in PyTorch
★ 978AITemplate. AITemplate is a Python framework which renders neural network into high performance CUDA/HIP C++ code. Specialized for FP16 TensorCore (NVIDIA GPU) and MatrixCore (AMD GPU) inference.
★ 4.7kbit. Code repo for the paper BiT Robustly Binarized Multi-distilled Transformer
★ 115mmdeploy. OpenMMLab Model Deployment Framework
★ 3.1kLoFTR. Code for "LoFTR: Detector-Free Local Feature Matching with Transformers", CVPR 2021, T-PAMI 2022
★ 2.9kDFM. Python (Pytorch) and Matlab (MatConvNet) implementations of CVPR 2021 Image Matching Workshop paper DFM: A Performance Baseline for Deep Feature Matching
★ 206DGC-Net. A PyTorch implementation of "DGC-Net: Dense Geometric Correspondence Network"
★ 206affnet. Code and weights for local feature affine shape estimation paper "Repeatability Is Not Enough: Learning Discriminative Affine Regions via Discriminability"
★ 279image-matching-toolbox. This is a toolbox repository to help evaluate various methods that perform image matching from a pair of images.
★ 596vivict. An easy to use in-browser tool for subjective comparison of the visual quality of different encodings of the same video source.
★ 138matlab_imresize. Python implementation of MatLab imresize function
★ 158jax3d. Python
★ 763RVRT. Recurrent Video Restoration Transformer with Guided Deformable Attention (NeurlPS2022, official repository)
★ 451mayavi. 3D visualization of scientific data in Python
★ 1.4kcolmap. COLMAP - Structure-from-Motion and Multi-View Stereo
★ 12kPREF. PREF: Phasorial Embedding Fields for Compact Neural Representations
★ 112youbit. Host any type of file on YouTube
★ 682AdelaiDepth. This repo contains the projects: 'Virtual Normal', 'DiverseDepth', and '3D Scene Shape'. They aim to solve the monocular depth estimation, 3D scene reconstruction from single image problems.
★ 1.1kCatPapers. Cool vision, learning, and graphics papers on Cats!
★ 1.2kTiNeuVox. TiNeuVox: Fast Dynamic Radiance Fields with Time-Aware Neural Voxels (SIGGRAPH Asia 2022)
★ 351DeltaCNN. Cuda
★ 13DeltaCNN. DeltaCNN End-to-End CNN Inference of Sparse Frame Differences in Videos
★ 59IBRNet. Python
★ 5223d-moments. Code for CVPR 2022 paper '3D Moments from Near-Duplicate Photos'
★ 251liif. Learning Continuous Image Representation with Local Implicit Image Function, in CVPR 2021 (Oral)
★ 1.4kVideoINR-Continuous-Space-Time-Super-Resolution. [CVPR 2022] VideoINR: Learning Video Implicit Neural Representation for Continuous Space-Time Super-Resolution
★ 302STDAN-CVPRW-2022. Space-Time Video Super-Resolution Using Deformable Attention Network
★ 4MulimgViewer. MulimgViewer is a multi-image viewer that can open multiple images in one interface, which is convenient for image comparison and image stitching.
★ 1.4kNAFNet. The state-of-the-art image restoration model without nonlinear activation functions.
★ 3.1kmetaseq. Repo for external large-scale work
★ 6.6kpytorch-optimizer. torch-optimizer -- collection of optimizers for Pytorch
★ 3.2kvvenc. VVenC, the Fraunhofer Versatile Video Encoder
★ 1.2ke2style. [TIP 2022] E2Style: Improve the Efficiency and Effectiveness of StyleGAN Inversion
★ 151Neighborhood-Attention-Transformer. Neighborhood Attention Transformer, arxiv 2022 / CVPR 2023. Dilated Neighborhood Attention Transformer, arxiv 2022
★ 1.2kFalcor. Real-Time Rendering Framework
★ 3.2klatent-diffusion. High-Resolution Image Synthesis with Latent Diffusion Models
★ 14kEDSC-pytorch. Code for Multiple Video Frame Interpolation via Enhanced Deformable Separable Convolution
★ 76Awesome-Diffusion-Models. A collection of resources and papers on Diffusion Models
★ 12kgeometry-free-view-synthesis. Is a geometric model required to synthesize novel views from a single image?
★ 381videowalk. Repository for "Space-Time Correspondence as a Contrastive Random Walk" (NeurIPS 2020)
★ 281video-diffusion-pytorch. Implementation of Video Diffusion Models, Jonathan Ho's new paper extending DDPMs to Video Generation - in Pytorch
★ 1.4kTimeCycle. Learning Correspondence from the Cycle-consistency of Time (CVPR 2019)
★ 723Torch-Encoding-Layer. Deep Texture Encoding Network
★ 92kubric. A data generation pipeline for creating semi-realistic synthetic multi-object videos with rich annotations such as instance segmentation masks, depth maps, and optical flow.
★ 2.8knerfies. This is the code for Deformable Neural Radiance Fields, a.k.a. Nerfies.
★ 2kRepLKNet-pytorch. Scaling Up Your Kernels to 31x31: Revisiting Large Kernel Design in CNNs (CVPR 2022)
★ 941ECCV2022-RIFE. ECCV2022 - Real-Time Intermediate Flow Estimation for Video Frame Interpolation
★ 5.5kDSQ. pytorch implementation of "Differentiable Soft Quantization: Bridging Full-Precision and Low-Bit Neural Networks"
★ 131Explorable-Super-Resolution. Python
★ 63BRCN. A unofficial implementation of paper method that 'Bidirectional Recurrent Convolutional Networks for Multi-Frame Super-Resolution'
★ 39antialiased-cnns. pip install antialiased-cnns to improve stability and accuracy
★ 1.7kmobile-vision. Mobile vision models and code
★ 922frame-interpolation. FILM: Frame Interpolation for Large Motion, In ECCV 2022.
★ 3.1keinops. Flexible and powerful tensor operations for readable and reliable code (for pytorch, jax, TF and others)
★ 9.6kVRT. VRT: A Video Restoration Transformer (official repository)
★ 1.5kAwesome-Incremental-Learning. Awesome Incremental Learning
★ 4.5kslow-motion. Python
★ 32Learning-to-Extract-a-Video-Sequence-from-a-Single-Motion-Blurred-Image. Python
★ 32BIN. Blurry Video Frame Interpolation (CVPR20)
★ 216DeFMO. [CVPR 2021] DeFMO: Deblurring and Shape Recovery of Fast Moving Objects
★ 177ConvNeXt. Code release for ConvNeXt model
★ 6.4klama. 🦙 LaMa Image Inpainting, Resolution-robust Large Mask Inpainting with Fourier Convolutions, WACV 2022
★ 10kJoJoGAN. Official PyTorch repo for JoJoGAN: One Shot Face Stylization
★ 1.4kFovVideoVDP. FovVideoVDP: A visible difference predictor for wide field-of-view video
★ 84mmflow. OpenMMLab optical flow toolbox and benchmark
★ 1.1kawesome-neural-rendering. Resources of Neural Rendering
★ 2.4kstylegan3. Official PyTorch implementation of StyleGAN3
★ 6.9kpytorch-pwc. a reimplementation of PWC-Net in PyTorch that matches the official Caffe version
★ 660Neural-Scene-Flow-Fields. PyTorch implementation of paper "Neural Scene Flow Fields for Space-Time View Synthesis of Dynamic Scenes"
★ 740RobustVideoMatting. Robust Video Matting in PyTorch, TensorFlow, TensorFlow.js, ONNX, CoreML!
★ 9.5kNeighbor2Neighbor. Neighbor2Neighbor: Self-Supervised Denoising from Single Noisy Images
★ 302focal-frequency-loss. [ICCV 2021] Focal Frequency Loss for Image Reconstruction and Synthesis
★ 709EGVSR. Efficient & Generic Video Super-Resolution
★ 952simsiam. PyTorch implementation of SimSiam https//arxiv.org/abs/2011.10566
★ 1.2kGPEN. Jupyter Notebook
★ 2.6kKAIR. Image Restoration Toolbox (PyTorch). Training and testing codes for DPIR, USRNet, DnCNN, FFDNet, SRMD, DPSR, BSRGAN, SwinIR
★ 3.5kNerfingMVS. [ICCV 2021 Oral] NerfingMVS: Guided Optimization of Neural Radiance Fields for Indoor Multi-view Stereo
★ 436bolt. 10x faster matrix and vector operations
★ 2.5kFastformer. A pytorch &keras implementation and demo of Fastformer.
★ 192omnimatte. Python
★ 804SCNeRF. [ICCV21] Self-Calibrating Neural Radiance Fields
★ 471Video-Swin-Transformer. This is an official implementation for "Video Swin Transformers".
★ 1.7kpt_ir. Image Restoration Toolkit in PyTorch
★ 37StyleGAN-anime. StyleGAN and StyleGAN2 implementation for generating anime faces.
★ 78ABME. Asymmetric Bilateral Motion Estimation for Video Frame Interpolation, ICCV2021
★ 104SwinIR. SwinIR: Image Restoration Using Swin Transformer (official repository)
★ 5.6kAwesome-Space-Time-Video-Super-Resolution. A List of Recent Space-Time Video Super-Resolution methods
★ 20Deep-SloMo. Official PyTorch implementation of "Deep Slow Motion Video Reconstruction with Hybrid Imaging System" (TPAMI)
★ 92meta-interpolation. Source code for CVPR 2020 paper "Scene-Adaptive Video Frame Interpolation via Meta-Learning"
★ 80XVFI. [ICCV 2021, Oral 3%] Official repository of XVFI
★ 311FeatureFlow. A state-of-the-art Video Frame Interpolation Method using feature flows blending. (CVPR 2020)
★ 151FLAVR. Code for FLAVR: A fast and efficient frame interpolation technique.
★ 516mocogan. MoCoGAN: Decomposing Motion and Content for Video Generation
★ 601SPDNet. [ICCV2021] Official code for Structure-Preserving Deraining with Residue Channel Prior Guidance
★ 57MoCoGAN-HD. [ICLR 2021 Spotlight] A Good Image Generator Is What You Need for High-Resolution Video Synthesis
★ 246fairscale. PyTorch extensions for high performance and large scale training.
★ 3.4kVideo-Frame-Interpolation-Collections. A collection of state-of-the-art video frame interpolation (VFI) methods.
★ 653db. A framework for analyzing computer vision models with simulated data
★ 126dynamic-video-depth. Code for the SIGGRAPH 2021 paper "Consistent Depth of Moving Objects in Video".
★ 274rpg_timelens. Repository relating to the CVPR21 paper TimeLens: Event-based Video Frame Interpolation
★ 626CaFM-Pytorch-ICCV2021. iccv2021-paper3163
★ 109Real-ESRGAN. Real-ESRGAN aims at developing Practical Algorithms for General Image/Video Restoration.
★ 36kStyleCLIP. Official Implementation for "StyleCLIP: Text-Driven Manipulation of StyleGAN Imagery" (ICCV 2021 Oral)
★ 4.1kDouZero. [ICML 2021] DouZero: Mastering DouDizhu with Self-Play Deep Reinforcement Learning | 斗地主AI
★ 4.6kxcit. Official code Cross-Covariance Image Transformer (XCiT)
★ 681GFPGAN. GFPGAN aims at developing Practical Algorithms for Real-world Face Restoration.
★ 38kSPADE-AnimeFace. Python
★ 4OPT. Implementation for <Orthogonal Over-Parameterized Training> in CVPR'21.
★ 22Barbershop. Barbershop: GAN-based Image Compositing using Segmentation Masks (SIGGRAPH Asia 2021)
★ 1.4kco-separation. Co-Separating Sounds of Visual Objects (ICCV 2019)
★ 98Awesome-Super-Resolution. Collect super-resolution related papers, data, repositories
★ 3.1ktorchdistill. A coding-free framework built on PyTorch for reproducible deep learning studies. PyTorch Ecosystem. 🏆26 knowledge distillation methods presented at TPAMI, CVPR, ICLR, ECCV, NeurIPS, ICCV, AAAI, etc are implemented so far. 🎁 Trained models, training logs and configurations are available for ensuring the reproducibiliy and benchmark.
★ 1.6kAnimeDrawingsDataset. A dataset for 2D pose estimation of anime/manga images.
★ 126gradio. Build and share delightful machine learning apps, all in Python. 🌟 Star to support our work!
★ 43kpytorchvideo. A deep learning library for video understanding research.
★ 3.6kzoom-learn-zoom. computational zoom from raw sensor data
★ 333SupER. Super-Resolution Erlangen (SupER): Benchmarking Super-Resolution Algorithms on Real Data
★ 87