This is your work, valued
Tenure-track associate professor at Shenzhen MSU-BIT University, China.
Awesome-Temporal-Action-Localization. A curated list of temporal action localization/detection and related area (e.g. temporal action proposal) resources.
★ 588PGCN. Graph Convolutional Networks for Temporal Action Localization (ICCV2019)
★ 322DRN. Dense Regression Network for Video Grounding (CVPR2020)
★ 53temporal-robustness-benchmark. Python
★ 20GCM. Graph Convolutional Module for Temporal Action Localization in Videos
★ 10tsn-pytorch. Python
★ 2Video2Reward.
★ 1Awesome_Long_Form_Video_Understanding. Awesome papers & datasets specifically focused on long-term videos.
★ 381egtr. [CVPR 2024 Best paper award candidate] EGTR: Extracting Graph from Transformer for Scene Graph Generation
★ 149GAR-bi-causal. Python
★ 5AVION. [arXiv:2309.16669] Code release for "Training a Large Video Model on a Single Machine in a Day"
★ 138awesome-video-domain-adaptation. A comprehensive collection of awesome research and other items about video domain adaptation
★ 114awesome-test-time-adaptation. Collection of awesome test-time (domain/batch/instance) adaptation methods
★ 1.3kmist. Jupyter Notebook
★ 37peft. 🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
★ 21kLaCLIP. [NeurIPS 2023] Text data, code and pre-trained models for paper "Improving CLIP Training with Language Rewrites"
★ 291Swin-Unet. [ECCVW 2022] The codes for the work "Swin-Unet: Unet-like Pure Transformer for Medical Image Segmentation"
★ 2.4kAwesome-Multimodal-Large-Language-Models. :sparkles::sparkles:Latest Advances on Multimodal Large Language Models
★ 18kDiffusionTAD. [ICCV 2023] Official PyTorch implementation of the paper "DiffTAD: Temporal Action Detection with Proposal Denoising Diffusion"
★ 37open_clip. An open source implementation of CLIP.
★ 14kXPretrain. Multi-modality pre-training
★ 511PLRC. Code for Point-Level Regin Contrast (https//arxiv.org/abs/2202.04639)
★ 353D-SPS. Python
★ 64EMB. Pytorch Implementation of ECCV'22 paper: Video Activity Localisation with Uncertainties in Temporal Boundary
★ 17eamat. Entity-Aware and Motion-Aware Transformers for Language-driven Action Localization(IJCAI-22)
★ 12STAN. Official PyTorch implementation of the paper "Revisiting Temporal Modeling for CLIP-based Image-to-Video Knowledge Transferring"
★ 107SelfDZSR. [ECCV 2022] Self-Supervised Learning for Real-World Super-Resolution from Dual Zoomed Observations
★ 78ECCV2026-Papers-with-Code. ECCV 2026 论文和开源项目合集,同时欢迎各位大佬提交issue,分享ECCV 2026论文和开源项目
★ 2.3kDiffusionInst. This repo is the code of paper "DiffusionInst: Diffusion Model for Instance Segmentation" (ICASSP'24).
★ 244ViTTA. Video Test-Time Adaptation for Action Recognition (CVPR 2023)
★ 53MaskCLIP. Official PyTorch implementation of "Extract Free Dense Labels from CLIP" (ECCV 22 Oral)
★ 481singularity. [ACL 2023] Official PyTorch code for Singularity model in "Revealing Single Frame Bias for Video-and-Language Learning"
★ 136stochastic-backpropagation.
★ 17FrozenBiLM. [NeurIPS 2022] Zero-Shot Video Question Answering via Frozen Bidirectional Language Models
★ 159event-based_vision_resources. Event-based Vision Resources. Community effort to collect knowledge on event-based vision technology (papers, workshops, datasets, code, videos, etc)
★ 3.6kawesome-3dcv-papers-daily. 主要记录计算机视觉、VSLAM、点云、结构光、机械臂抓取、三维重建、深度学习、自动驾驶等前沿paper与文章。
★ 667awesome-source-free-test-time-adaptation. Test-time Adaptation, Test-time Training and Source-free Domain Adaptation
★ 549TCGL. [IEEE T-IP 2022] TCGL: Temporal Contrastive Graph for Self-supervised Video Representation Learning
★ 24revisiting-spatial-temporal-layouts. Codebase for "Revisiting spatio-temporal layouts for compositional action recognition" (Oral at BMVC 2021).
★ 27OCRTOC_software_package. Python
★ 75easy_handeye. Automated, hardware-independent Hand-Eye Calibration
★ 1.2kAwesome_Underwater_Datasets. Pointers to large-scale underwater datasets and relevant resources.
★ 706UWKFusion. Front end GUI for 3D scene reconstruction using Kinect v2 on underwater depth data.
★ 24cotta. [CVPR 2022] Official CoTTA Code for our paper Continual Test-Time Domain Adaptation
★ 324Look-Outside-Room. [CVPR 2022] Look Outside the Room: Synthesizing A Consistent Long-Term 3D Scene Video from A Single Image
★ 82CACL. [CVPR 2022] Cross-Architecture Self-supervised Video Representation Learning
★ 24MMN. [AAAI 2022] Negative Sample Matters: A Renaissance of Metric Learning for Temporal Grounding
★ 91awesome-multimodal-ml. Reading list for research topics in multimodal machine learning
★ 6.9kViGA. "Video Moment Retrieval from Text Queries via Single Frame Annotation" in SIGIR 2022
★ 68graph-based-deep-learning-literature. links to conference publications in graph-based deep learning
★ 5.1kVideoINR-Continuous-Space-Time-Super-Resolution. [CVPR 2022] VideoINR: Learning Video Implicit Neural Representation for Continuous Space-Time Super-Resolution
★ 302MAE-pytorch. Unofficial PyTorch implementation of Masked Autoencoders Are Scalable Vision Learners
★ 2.7kE2E-TAD. [CVPR 2022] An Empirical Study of End-to-end Temporal Action Detection
★ 86TemporalAlignNet. [CVPR'22 Oral] Temporal Alignment Networks for Long-term Video. Tengda Han, Weidi Xie, Andrew Zisserman.
★ 122activitygraph_transformer. Python
★ 70auditory-slow-fast. Implementation of "Slow-Fast Auditory Streams for Audio Recognition, ICASSP, 2021" in PyTorch
★ 73Underwater-Color-Correction. Using GANs to correct color distortion in underwater images.
★ 327awesome-audio-visual. A curated list of different papers and datasets in various areas of audio-visual processing
★ 775xbmc-addons-chinese. Addon scripts, plugins, and skins for XBMC Media Center. Special for chinese laguage.
★ 3.2kawesome-vision-language-pretraining-papers. Recent Advances in Vision and Language PreTrained Models (VL-PTMs)
★ 1.2kSP-TAD. Towards High-Quality Temporal Action Detection with Sparse Proposals
★ 10DCAN. [AAAI 2022] DCAN: Improving Temporal Action Detection via Dual Context Aggregation
★ 17Video-Swin-Transformer. This is an official implementation for "Video Swin Transformers".
★ 1.7kpytorch_violet. A PyTorch implementation of VIOLET
★ 138ActionCLIP. This is the official implement of paper "ActionCLIP: A New Paradigm for Action Recognition"
★ 615Dynamic-Vision-Transformer. Accelerating T2t-ViT by 1.6-3.6x.
★ 260vtpack. code base for vision transformers
★ 36SSTAP. Code for our CVPR 2021 Paper "Self-Supervised Learning for Semi-Supervised Temporal Action Proposal".
★ 72STAM. Official implementation of "An Image is Worth 16x16 Words, What is a Video Worth?" (2021 paper)
★ 221VisualTransformers. A Pytorch Implementation of the following paper "Visual Transformers: Token-based Image Representation and Processing for Computer Vision"
★ 181TadTR. [TIP 2022] End-to-end Temporal Action Detection with Transformer
★ 167Lighting-the-Darkness-in-the-Deep-Learning-Era-Open.
★ 815WaterGAN. Source code for "WaterGAN: Unsupervised Generative Network to Enable Real-time Color Correction of Monocular Underwater Images"
★ 171Underwater-image-restoration. Implementation of Chongyi Li et.al algorithm for underwater image restoration.
★ 35UnderwaterImage-Enhancement. 水下图像增强算法,三个工程均为MATLAB版本,作者北京大学张文浩
★ 42UWCNN. Code and Datasets for "Underwater Scene Prior Inspired Deep Underwater Image and Video Enhancement", Pattern Recognition, 2019
★ 182All-In-One-Underwater-Image-Enhancement-using-Domain-Adversarial-Learning. [CVPRW 2019] All-In-One Underwater Image Enhancement using Domain-Adversarial Learning
★ 71WCT2. Software that can perform photorealistic style transfer without the need of any post-processing steps.
★ 914Underwater-Image-Enhancement-via-Style-Transfer. This is the official repository for Exemplar Based Underwater Image Enhancement augmented by Wavelet Corrected Transforms.
★ 39FUnIE-GAN. Fast underwater image enhancement for Improved Visual Perception. #TensorFlow #PyTorch #RAL2020
★ 638UnderwaterImageRestoration. Underwater Image Restoration.
★ 64Single-Underwater-Image-Enhancement-and-Color-Restoration. Single Underwater Image Enhancement and Color Restoration, which is Python implementation for a comprehensive review paper "An Experimental-based Review of Image Enhancement and Image Restoration Methods for Underwater Imaging"
★ 763GANTransferLimitedData. This is a pytorch implementation of the paper "On Leveraging Pretrained GANs for Limited-Data Generation".
★ 59ReReVST-Code. Released code of Consistent Video Style Transfer via Relaxation and Regularization, TIP 2020
★ 65awesome-low-light-image-enhancement. This is a resouce list for low light image enhancement
★ 1.8kStableLLVE. Learning Temporal Consistency for Low Light Video Enhancement from Single Images (CVPR2021)
★ 164BIMEF. Code and data for the research paper "A Bio-Inspired Multi-Exposure Fusion Framework for Low-light Image Enhancement" (Submitted to IEEE Transactions on Cybernetics)
★ 275ActionDetection-AFSD. Code for CVPR2021 paper "Learning Salient Boundary Feature for Anchor-free Temporal Action Localization"
★ 185MUSES. [CVPR 2021] Multi-shot Temporal Event Localization: a Benchmark
★ 55TSP. TSP: Temporally-Sensitive Pretraining of Video Encoders for Localization Tasks (ICCVW 2021)
★ 119TDN. [CVPR 2021] TDN: Temporal Difference Networks for Efficient Action Recognition
★ 386RSPNet. Official Pytorch implementation for AAAI2021 paper (RSPNet: Relative Speed Perception for Unsupervised Video Representation Learning)
★ 37TimeSformer-pytorch. Implementation of TimeSformer from Facebook AI, a pure attention-based solution for video classification
★ 729A2Net. Revisiting Anchor Mechanisms for Temporal Action Localization (TIP 2020)
★ 36PyTorch-MFNet. Python
★ 251Video-Grounding-from-Text. Source code for "Weakly-Supervised Video Object Grounding from Text by Loss Weighting and Object Interaction"
★ 47CenterTrack. Simultaneous object detection and tracking using center points.
★ 2.5kTPN. [CVPR 2020] Temporal Pyramid Network for Action Recognition
★ 394DeepLearningInMedicalImagingAndMedicalImageAnalysis.
★ 532CharadesDet. Charades Object Detection Dataset (ICCV 2017)
★ 31charades-algorithms. Activity Recognition Algorithms for the Charades Dataset
★ 206something_else. Code repository for the paper: 'Something-Else: Compositional Action Recognition with Spatial-Temporal Interaction Networks'
★ 148NICE-GAN-pytorch. Official PyTorch implementation of NICE-GAN: Reusing Discriminators for Encoding: Towards Unsupervised Image-to-Image Translation
★ 242evar. Explainable Video Action Reasoning via Prior Knowledge and State Transitions
★ 21generative_inpainting. DeepFill v1/v2 with Contextual Attention and Gated Convolution, CVPR 2018, and ICCV 2019 Oral
★ 3.5kVideoX. VideoX: a collection of video cross-modal models
★ 1.1kActionDetection-DBG. Code for AAAI2020 paper "Fast Learning of Temporal Action Proposal via Dense Boundary Generator"
★ 351