This is your work, valued
two-stream-pytorch. PyTorch implementation of two-stream networks for video action recognition
579agentic-ai-system-course. Use agent to learn agent - A skeleton course on how to design, build, and operate production AI agents
539Hidden-Two-Stream. Caffe implementation for "Hidden Two-Stream Convolutional Networks for Action Recognition"
190Video-Tutorial-CVPR2020. A Comprehensive Tutorial on Video Modeling
66GuidedNet. Caffe implementation for "Guided Optical Flow Learning"
33deepOF. TensorFlow implementation for "Guided Optical Flow Learning"
25paper-reading. 深度学习经典、新论文逐段精读
24tiny-ucf101.
2autogluon. AutoGluon: AutoML for Image, Text, and Tabular Data
2higgs-audio. Text-audio foundation model from Boson AI
1bark. 🔊 Text-Prompted Generative Audio Model
1ViLT. Code for the ICML 2021 (long talk) paper: "ViLT: Vision-and-Language Transformer Without Convolution or Region Supervision"
1Video-Swin-Transformer. This is an official implementation for "Video Swin Transformers".
1semantic-segmentation. Improving Semantic Segmentation via Video Propagation and Label Relaxation
1SlowFast. PySlowFast: video understanding codebase from FAIR for reproducing state-of-the-art video models.
1