This is your work, valued
Postdoc in ML&CV at the University of Edinburgh.
LDMVFI. [AAAI 2024] "LDMVFI: Video Frame Interpolation with Latent Diffusion Models", Duolikun Danier, Fan Zhang, David Bull
★ 193ST-MFNet. [CVPR 2022] "ST-MFNet: A Spatio-Temporal Multi-Flow Network for Frame Interpolation", Duolikun Danier, Fan Zhang, David Bull
★ 75depthcues. [CVPR 2025] "DepthCues: Evaluating Monocular Depth Perception in Large Vision Models", Duolikun Danier, Mehmet Aygün, Changjian Li, Hakan Bilen, Oisin Mac Aodha
★ 23flolpips. [IEEE PCS 2022 best paper finalist] "FloLPIPS: A Bespoke Video Quality Metric for Frame Interpoation", Duolikun Danier, Fan Zhang, David Bull
★ 22BVI-VFI-database. [IEEE TIP 2023] "BVI-VFI: A Video Quality Database for Video Frame Interpolation", Duolikun Danier, Fan Zhang, David Bull
★ 8EDC. [IEEE ICIP 2022] "Enhancing Deformable Convolution based Video Frame Interpolation with Coarse-to-fine 3D CNN", Duolikun Danier, Fan Zhang, David Bull
★ 5CFKD. Python
★ 1Claude-Code-OrangeBook. Claude Code从入门到精通橙皮书 by 花叔
★ 71hello-agents. 📚 《从零开始构建智能体》——从零开始的智能体原理与实践教程
★ 70kMole. 🐹 Clean, uninstall, analyze, optimize, and monitor your Mac from the terminal.
★ 61kefficientsam3. EfficientSAM3 compresses SAM3 into lightweight, edge-friendly models via progressive knowledge distillation for fast promptable concept segmentation and tracking.
★ 647GeometryForcing. [ICLR26] Official implementation of Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling
★ 214HumanSystemOptimization. 健康学习到150岁 - 人体系统调优不完全指南
★ 22kIndoorInverseRendering. [SIGGRAPH Asia'22] Learning-based Inverse Rendering of Complex Indoor Scenes with Differentiable Monte Carlo Raytracing
★ 112RealCam-I2V. Python
★ 60vipe. ViPE: Video Pose Engine for Geometric 3D Perception
★ 2.1kmatcha. [CVPR 2025 Highlight] MATCHA: Towards Matching Anything.
★ 77sv3d-diffusers. Stability-AI's SV3D (ECCV 2024 oral, Voleti et al.) in the diffusers convention.
★ 33MVGBench. A comprehensive benchmark suite for multi-view generation models
★ 21latent-radiance-field. [ICLR 2025] Latent Radiance Fields with 3D-aware 2D Representations
★ 125HunyuanWorld-1.0. Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels with Hunyuan3D World Model
★ 2.9kGenDeg. Official code for GenDeg: Diffusion-Based Degradation Synthesis for Generalizable All-in-One Image Restoration
★ 68zero-academic-page. Zero Academic Homepage is a clean, modern and responsive theme for academic personal websites.
★ 40Awesome-Video-Diffusion-Models. [CSUR] A Survey on Video Diffusion Models
★ 2.3kPoseTraj. [CVPR 2025] PoseTraj: Pose-Aware Trajectory Control in Video Diffusion
★ 23Go-with-the-Flow. The official implementation of CVPR'25 Oral paper "Go-with-the-Flow: Motion-Controllable Video Diffusion Models Using Real-Time Warped Noise"
★ 1.1kpvrobo. Code for the "When Pre-trained Visual Representations Fall Short: Limitations in Visuo-Motor Robot Learning" paper.
★ 12FiT3D. [ECCV 2024] Improving 2D Feature Representations by 3D-Aware Fine-Tuning
★ 329perception_models. State-of-the-art Image & Video CLIP, Multimodal Large Language Models, and More!
★ 2.3kCameraCtrl. Python
★ 658GEN3C. [CVPR 2025 Highlight] GEN3C: 3D-Informed World-Consistent Video Generation with Precise Camera Control
★ 1.4klatentsplat. [ECCV 2024] Implementation of latentSplat: Autoencoding Variational Gaussians for Fast Generalizable 3D Reconstruction
★ 239ESAM. [ICLR 2025, Oral] EmbodiedSAM: Online Segment Any 3D Thing in Real Time
★ 634neural-isometries. Official JAX implementation of neural isometries - taming transformations for equivariant ML
★ 36CUT3R. Official implementation of Continuous 3D Perception Model with Persistent State
★ 1.5kzeroverse. Official code for NeurIPS 2024 paper LRM-Zero: Training Large Reconstruction Models with Synthesized Data
★ 155DINO-Foresight. [NeurIPS 2025] Official Implementation of DINO-Foresight: Looking into the Future with DINO
★ 167radfoam. Original implementation of "Radiant Foam: Real-Time Differentiable Ray Tracing"
★ 665ViewDiff. ViewDiff generates high-quality, multi-view consistent images of a real-world 3D object in authentic surroundings. (CVPR2024).
★ 3823DCorrEnhance. Python
★ 37AASSL. Python
★ 2cosmos. NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.
★ 11k3D-Gaussian-Splatting-Papers. 3D高斯论文,持续更新,欢迎交流讨论。
★ 3.1kMVImgNet. CVPR2023 | MVImgNet: A Large-scale Dataset of Multi-view Images
★ 488nerfies.github.io. JavaScript
★ 4.3karxiv-latex-cleaner. arXiv LaTeX Cleaner: Easily clean the LaTeX code of your paper to submit to arXiv
★ 7kexpert_readed_books. 2021年最新总结,推荐工程师合适读本,计算机科学,软件技术,创业,思想类,数学类,人物传记书籍
★ 12kSurgicalDINO. [IPCAI 2024 (IJCARS special issue)] Surgical-DINO: Adapter Learning of Foundation Models for Depth Estimation in Endoscopic Surgery
★ 70ViT-Adapter. [ICLR 2023 Spotlight] Vision Transformer Adapter for Dense Predictions
★ 1.5kAwesome-PVRobotics. A curated list of papers, regarding the deployment of pre-trained visual representations (PVR) for robot learning tasks.
★ 14GeoMeter. Codebase of GeoMeter: Understanding Depth and Height Perception of Large Visual Language Models
★ 1croco. Python
★ 508dust3r. DUSt3R: Geometric 3D Vision Made Easy
★ 7.3kamodalAPI. Jupyter Notebook
★ 40Awesome-LLM-3D. Awesome-LLM-3D: a curated list of Multi-modal Large Language Model in 3D world Resources
★ 2.2kSVD_Xtend. Stable Video Diffusion Training Code and Extensions.
★ 731phy-sd. Python
★ 6G3DR. [CVPR 2024] G3DR: Generative 3D Reconstruction in ImageNet
★ 38Era3D. Python
★ 644pytorch-image-models. The largest collection of PyTorch image encoders / backbones. Including train, eval, inference, export scripts, and pretrained weights -- ResNet, ResNeXT, EfficientNet, NFNet, Vision Transformer (ViT), MobileNetV4, MobileNet-V3 & V2, RegNet, DPN, CSPNet, Swin Transformer, MaxViT, CoAtNet, ConvNeXt, and more
★ 37kVPT. Dataset for visual perspective taking
★ 11GeoBench. A toolbox for benchmarking SOTA discriminative and generative geometry estimation models.
★ 65V3D. [T-PAMI 2025] V3D: Video Diffusion Models are Effective 3D Generators
★ 522OpenLRM. An open-source impl. of Large Reconstruction Models
★ 1.2kGeneral-World-Models-Survey.
★ 527probe3d. [CVPR 2024] Probing the 3D Awareness of Visual Foundation Models
★ 357splatter-image. Official implementation of `Splatter Image: Ultra-Fast Single-View 3D Reconstruction' CVPR 2024
★ 1.1kdepth-fm. [AAAI 2025, Oral] DepthFM: Fast Monocular Depth Estimation with Flow Matching
★ 755dino. PyTorch code for Vision Transformers training with the Self-Supervised learning method DINO
★ 7.6kal-folio. A beautiful, simple, clean, and responsive Jekyll theme for academics
★ 16kllama3-from-scratch. llama3 implementation one matrix multiplication at a time
★ 15kbooks. 各类闲书分享(equb版本,ipad可直接打开阅读)
★ 739Self-supervised-Learning.
★ 162dinov2. PyTorch code and models for the DINOv2 self-supervised learning method.
★ 13kawesome_lists. Awesome Lists for Tenure-Track Assistant Professors and PhD students. (助理教授/博士生生存指南)
★ 1.6kAwesome-Mamba-Papers. Awesome Papers related to Mamba.
★ 1.4kNeRCo. [ICCV 2023] Implicit Neural Representation for Cooperative Low-light Image Enhancement
★ 267recurrent-memory-transformer. [NeurIPS 22] [AAAI 24] Recurrent Transformer-based long-context architecture.
★ 779STDiffProject. [AAAI'24] "STDiff: Spatio-temporal Diffusion for Continuous Stochastic Video Prediction". Xi Ye, Guillaume-Alexandre Bilodeau
★ 31NPVP. [CVPRW'23] "A unified model for continuous conditional video prediction". Xi Ye, Guillaume-Alexandre Bilodeau.
★ 14VPTR. The repository for paper VPTR: Efficient Transformers for Video Prediction
★ 102Diffusion-Low-Light. Official pytorch implementation for "Low-light Image Enhancement with Wavelet-based Diffusion Models"
★ 313LoRA. Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"
★ 14kvariational-diffusion. Unofficial implementation of Variational Diffusion Models in PyTorch (Lightning)
★ 12LLDiffusion. The code of paper "LLDiffusion: Learning Degradation Representations in Diffusion Models for Low-Light Image Enhancement", PR 2025
★ 39CLEDiffusion. (MM2023)CLE Diffusion: Controllable Light Enhancement Diffusion Model. Authors: Yuyang Yin, Dejia Xu, Chuangchuang Tan, Ping Liu, Yao Zhao, Yunchao Wei
★ 76Lighting-the-Darkness-in-the-Deep-Learning-Era-Open.
★ 815EnlightenGAN. [IEEE TIP] "EnlightenGAN: Deep Light Enhancement without Paired Supervision" by Yifan Jiang, Xinyu Gong, Ding Liu, Yu Cheng, Chen Fang, Xiaohui Shen, Jianchao Yang, Pan Zhou, Zhangyang Wang
★ 1.1kawesome-low-light-image-enhancement. This is a resouce list for low light image enhancement
★ 1.8kDID.
★ 14consistencydecoder. Consistency Distilled Diff VAE
★ 2.2kshortest-path-diffusion. Official code for the paper "Image generation with shortest path diffusion" accepted at ICML 2023.
★ 24Explicit-Visual-Prompt. [CVPR 2023 & TPAMI 2025] Explicit Visual Prompting for Low-Level Structure Segmentations
★ 231Diffusion-DEIS-SN. Repo for NeurIPS 23 Diffusion workshop paper titled "Score Normalization for a Faster Diffusion Exponential Integrator Sampler" from MediaTek Research UK
★ 4ST-MFNet-Mini. PyTorch implementation of ST-MFNet Mini
★ 5ExplainableVQA. [ACMMM Oral, 2023] "Towards Explainable In-the-wild Video Quality Assessment: A Database and a Language-Prompted Approach"
★ 87diffusers. 🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
★ 34kVideo-Frame-Interpolation-Summary. Video Frame Interpolation Summary and Infer
★ 135IFRNet. IFRNet: Intermediate Feature Refine Network for Efficient Frame Interpolation (CVPR 2022)
★ 343UPR-Net. Official implementation of our CVPR2023 paper "A Unified Pyramid Recurrent Network for Video Frame Interpolation"
★ 126EMA-VFI. [CVPR 2023] Extracting Motion and Appearance via Inter-Frame Attention for Efficient Video Frame Interpolatio
★ 498mlp_maps. Code for "Representing Volumetric Videos as Dynamic MLP Maps" CVPR 2023
★ 238frameintIFE. Official PyTorch implementation of "Frame Interpolation for Dynamic Scenes with Implicit Flow Encoding"
★ 41AMT. Official code for "AMT: All-Pairs Multi-Field Transforms for Efficient Frame Interpolation" (CVPR2023)
★ 270CURE. Official Implementation of CURE, ECCV 2022
★ 14Awesome-Dataset-Distillation. A curated list of awesome papers on dataset distillation and related applications.
★ 2kVIDUE. Joint Video Multi-Frame Interpolation and Deblurring under Unknown Exposure Time (CVPR2023)
★ 74awesome-ai-art-image-synthesis. A list of awesome tools, ideas, prompt engineering tools, colabs, models, and helpers for the prompt designer playing with aiArt and image synthesis. Covers Dalle2, MidJourney, StableDiffusion, and open source tools.
★ 1.8knr-vqa-consumervideo. No-Reference Video Quality Model for consumer video
★ 49tuning_playbook. A playbook for systematically maximizing the performance of deep learning models.
★ 30kFAST-VQA-and-FasterVQA. [ECCV2022, TPAMI2023] FAST-VQA, and its extended version FasterVQA.
★ 348vector-quantize-pytorch. Vector (and Scalar) Quantization, in Pytorch
★ 4klatent-diffusion. High-Resolution Image Synthesis with Latent Diffusion Models
★ 14kx-transformers. A concise but complete full-attention transformer with a set of promising experimental features from various papers
★ 5.9kpytorch-deep-video-prior. [NeurIPS 2020] Blind Video Temporal Consistency via Deep Video Prior
★ 127bjontegaard. This repository contains scripts to calculate the famous Bjontegaard-Delta metric with different interpolation functions.
★ 41Cold-Diffusion-Models. Official implementation of Cold-Diffusion for different transformations in pytorch.
★ 1.1kAwesome-Diffusion-Models. A collection of resources and papers on Diffusion Models
★ 12kUnityARCourseraSolutions. Partial solutions for the Coursera course in Handheld AR App Development with Unity. Feel free to contribute! 🎓
★ 16Production-Ready-Applied-Deep-Learning. Production-Ready Applied Deep Learning
★ 92denoising-diffusion-pytorch. Implementation of Denoising Diffusion Probabilistic Models in PyTorch
★ 405mcvd-pytorch. Official implementation of MCVD: Masked Conditional Video Diffusion for Prediction, Generation, and Interpolation (https://arxiv.org/abs/2205.09853)
★ 370DBVI. Official code for Deep Bayesian Video Frame Interpolation (ECCV2022)
★ 18video-diffusion-pytorch. Implementation of Video Diffusion Models, Jonathan Ho's new paper extending DDPMs to Video Generation - in Pytorch
★ 1.4kdenoising-diffusion-pytorch. Implementation of Denoising Diffusion Probabilistic Model in Pytorch
★ 11kscrcpy. Display and control your Android device
★ 147kVideo-Frame-Interpolation-Rankings-and-Video-Deblurring-Rankings. ABME AdaFNIO AMT BiM-VFI BiT CBBD CDFI CtxSyn DBVI DC-BVFI DQBC DRVI DvP EAFI EBME EDC EDEN EDENVFI EDSC EMA-VFI FGDCN FILM FLAVR GIMM-VFI HFD HiFI H-VFI IFRNet InterpAny-Clearer IQ-VFI JNMR LADDER M2M MA-GCSPA MoMo PerVFI RIFE RN-VFI SepConv SoftSplat SSR ST-MFNet Swin-VFI TDPNet TLB-VFI TTVFI UGFI UPR-Net VFIformer VFIFT VFIMamba VFIT VRT XVFI
★ 156nerf. Code release for NeRF (Neural Radiance Fields)
★ 1VFIPS. The PyTorch implementation for A Perceptual Quality Metric for Video Frame Interpolation
★ 45stable-diffusion. A latent text-to-image diffusion model
★ 73kexamples. A set of examples around pytorch in Vision, Text, Reinforcement Learning, etc.
★ 24kECCV2022-RIFE. Official MegEngine Implementation of Real-Time Intermediate Flow Estimation for Video Frame Interpolation
★ 30UTI-VFI. Video Frame Interpolation without Temporal Priors (a general method for blurry video interpolation)
★ 35BIN. Blurry Video Frame Interpolation (CVPR20)
★ 216frame-interpolation. FILM: Frame Interpolation for Large Motion, In ECCV 2022.
★ 3.1kcv-barbetti. :briefcase: Repository for various academic and professional CV versions
★ 10ECCV-2022-Papers.
★ 135LaViSE. Explaining Deep Convolutional Neural Networks via Unsupervised Visual-Semantic Filter Attention (CVPR 2022)
★ 20gmflow. [CVPR'22 Oral] GMFlow: Learning Optical Flow via Global Matching
★ 796randd. Video codec comparison toolkit
★ 6VRT. VRT: A Video Restoration Transformer (official repository)
★ 1.5kVideoINR-Continuous-Space-Time-Super-Resolution. [CVPR 2022] VideoINR: Learning Video Implicit Neural Representation for Continuous Space-Time Super-Resolution
★ 302vit-pytorch. Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch
★ 25kRSTT. Official pytorch implementation of paper "RSTT: Real-time Spatial Temporal Transformer for Space-Time Video Super-Resolution"
★ 143Video-Swin-Transformer. This is an official implementation for "Video Swin Transformers".
★ 1.7kVideo-Frame-Interpolation-Transformer. Python
★ 103TTSR. [CVPR'20] TTSR: Learning Texture Transformer Network for Image Super-Resolution
★ 792oneflow. OneFlow is a deep learning framework designed to be user-friendly, scalable and efficient.
★ 9.4kcomputer-book. collection of Computer Book
★ 437VIDEVAL. [IEEE TIP'2021] "UGC-VQA: Benchmarking Blind Video Quality Assessment for User Generated Content", Zhengzhong Tu, Yilin Wang, Neil Birkbeck, Balu Adsumilli, Alan C. Bovik
★ 135maxim. [CVPR 2022 Oral] Official repository for "MAXIM: Multi-Axis MLP for Image Processing". SOTA for denoising, deblurring, deraining, dehazing, and enhancement.
★ 1.1kdeep-video-prior. [NeurIPS 2020] Blind Video Temporal Consistency via Deep Video Prior
★ 329Neural-IMage-Assessment. A PyTorch Implementation of Neural IMage Assessment
★ 585SinNeRF. [ECCV 2022] "SinNeRF: Training Neural Radiance Fields on Complex Scenes from a Single Image", Dejia Xu, Yifan Jiang, Peihao Wang, Zhiwen Fan, Humphrey Shi, Zhangyang Wang
★ 330VFIformer. Video Frame Interpolation with Transformer (CVPR2022)
★ 132deepIQA. NR and FR IQA models based on Deep Convolutional Neural Networks
★ 209vbliinds. Fast Blind Natural Video Quality (V-BLIINDS)
★ 27DISTS. IQA: Deep Image Structure and Texture Similarity Metric
★ 486M2M_VFI. Many-to-many Splatting for Efficient Video Frame Interpolation
★ 93CONTRIQUE. Official implementation for "Image Quality Assessment using Contrastive Learning"
★ 140GST_VQA.
★ 1Inter4K. Official repository for downloading and using Inter4K video interpolation dataset
★ 50OpticalFlow_Visualization. Python optical flow visualization following Baker et al. (ICCV 2007) as used by the MPI-Sintel challenge
★ 458CVPR-2022-Papers.
★ 630PerceptualSimilarity. LPIPS metric. pip install lpips
★ 4.3kclic2021-devkit. Development Kit for the CLIC compression challenge 2021.
★ 18HFR-BVQA. Python
★ 17WeightStandardization. Standardizing weights to accelerate micro-batch training
★ 547XVFI. [ICCV 2021, Oral 3%] Official repository of XVFI
★ 311siti-tools. SI TI calculation tools
★ 58EQVI. Winning solution of AIM2020 VTSR Challenge (video interpolation). EQVI-Enhanced Quadratic Video Interpolation.
★ 119fucking-algorithm. Crack LeetCode, not only how, but also why.
★ 135kinteresting-python. 有趣的Python爬虫和Python数据分析小项目(Some interesting Python crawlers and data analysis projects)
★ 5kR_Unet. Video prediction using lstm and unet
★ 31pytorch-parallelism.
★ 4CDFI. [CVPR 2021] CDFI: Compression-Driven Network Design for Frame Interpolation
★ 114Visual-Information-Fidelity---Python. Visual Information Fidelity Code - Python
★ 40awesome-optical-flow-algorithm. A curated list of resources dedicated to optical flow algorithms. Feel free to make PRs to contribute.
★ 159pyflow. Fast, accurate and easy to run dense optical flow with python wrapper
★ 661OF_DIS. Fast Optical Flow using Dense Inverse Search (DIS)
★ 325GDConvNet. Python
★ 15EDSC-pytorch. Code for Multiple Video Frame Interpolation via Enhanced Deformable Separable Convolution
★ 76BMBC. BMBC: Bilateral Motion Estimation with Bilateral Cost Volume for Video Interpolation, ECCV 2020
★ 86pytorch-voxel-flow. Video Frame Synthesis using Deep Voxel Flow
★ 141CAIN. Source code for AAAI 2020 paper "Channel Attention Is All You Need for Video Frame Interpolation"
★ 350DAIN. Depth-Aware Video Frame Interpolation (CVPR 2019)
★ 8.3kmeta-interpolation. Source code for CVPR 2020 paper "Scene-Adaptive Video Frame Interpolation via Meta-Learning"
★ 80personal-website. Code that'll help you kickstart a personal website that showcases your work as a software developer.
★ 7.6k