This is your work, valued
TSA-Net. [ACM MM 2021] TSA-Net: Tube Self-Attention Network for Action Quality Assessment
★ 40CA-SpaceNet. [IROS 2022] CA-SpaceNet: Counterfactual Analysis for 6D Pose Estimation in Space
★ 26Tiny-YOLO-LSQ. This is an implementation of YOLO using LSQ network quantization method.
★ 22CPR-Coach. Coach-Project
★ 19MyAlgorithms. Algorithms notes learning from ZuoShen.
★ 10CubeRobot. Code of 4-axis cube robot.
★ 9CPR-CLIP. [IEEE SPL 2023] CPR-CLIP: Multimodal Pre-training for Composite Error Recognition in CPR Training.
★ 8CNN_Numpy. We used Numpy as a tool to build a convolutional neural network from scratch to complete the task of handwritten digit recognition.
★ 1EasyOffer. 《EasyOffer》(<大模型面经合集>)是针对LLM宝宝们量身打造的大模型暑期实习Offer指南,主要记录大模型暑期实习和秋招准备的一些常见大厂手撕代码、大厂面经经验、常见大厂思考题等;小白一个,正在学习ing......有问题各位大佬随时指正,希望大家都能拿到心仪Offer!
★ 812Embodied-AI-Guide. [Lumina具身智能社区] 具身智能技术指南 Embodied-AI-Guide
★ 15kAwesome-Unified-Multimodal-Models. Awesome Unified Multimodal Models
★ 1.3kAI-Face-FairnessBench. We introduce AI-Face, the first million-scale AI-generated face dataset with demographic annotations, and conduct a comprehensive fairness benchmark. Our work has been accepted at CVPR 2025.
★ 98SOLO. Python
★ 9Effort-AIGI-Detection. Official implementation of ICML 2025 Oral 🏆 paper "Orthogonal Subspace Decomposition for Generalizable AI-Generated Image Detection".
★ 227DeepfakeBench. A comprehensive benchmark of deepfake detection
★ 1.1kAwesome-Deepfakes-Detection. A list of tools, papers and code related to Deepfake Detection.
★ 1.8kAwesome-AI-generated-Image-Detection. A list of awesome papers on AI-generated Image Detection.
★ 104DF40. Official repository for the next-generation deepfake detection dataset (DF40), comprising 40 distinct deepfake techniques, even the just released SoTAs. Our work has been accepted by NeurIPS 2024.
★ 347Mamba-in-CV. A paper list of some recent Mamba-based CV works.
★ 491langchain-examples. 🦜通过演示 LangChain 最具有代表性的应用范例,带你快速上手 LangChain 各个使用场景。(包含完整代码和数据集)
★ 553D-Occupancy-Perception. [Information Fusion 2025] A Survey on Occupancy Perception for Autonomous Driving: The Information Fusion Perspective
★ 625Awesome-occupancy-perception.
★ 310Awesome-BEV-Perception. Awesome-BEV-Perception
★ 33BEV-Perception. Bird's Eye View Perception
★ 712AISystem. AISystem 主要是指AI系统,包括AI芯片、AI编译器、AI推理和训练框架等AI全栈底层技术
★ 17kDiffusion-Models-Papers-Survey-Taxonomy. Diffusion model papers, survey, and taxonomy
★ 3.4kAwesome-Video-Diffusion-Models. [CSUR] A Survey on Video Diffusion Models
★ 2.3krouting-transformer. Fully featured implementation of Routing Transformer
★ 300denoising-diffusion-pytorch. Implementation of Denoising Diffusion Probabilistic Model in Pytorch
★ 11kvideo-action-recognition-datasets. This repository contains video datasets that can be used for training coarse to fine-grained (phase, step and action) temporal classification tasks.
★ 16TMRNet. [TMI'2021] Temporal Memory Relation Network for Workflow Recognition from Surgical Video
★ 67coding-for-great-offer. 大厂算法和数据结构刷题班
★ 419course. 高性能并行编程与优化 - 课件
★ 4.2kalgorithm-journey. 左程云的算法和数据结构通关课
★ 3.3kCPR-CLIP. [IEEE SPL 2023] CPR-CLIP: Multimodal Pre-training for Composite Error Recognition in CPR Training.
★ 8conv2d_direct. Cuda
★ 36DeepKE. [EMNLP 2022] An Open Toolkit for Knowledge Graph Extraction and Construction
★ 4.5kCA-SpaceNet. [IROS 2022] CA-SpaceNet: Counterfactual Analysis for 6D Pose Estimation in Space
★ 26DiffAct. Code for Diffusion Action Segmentation (ICCV 2023)
★ 78Video-Relation-Detection.
★ 4VideoRelationDetection2021.
★ 3LeetCode-Book. 《剑指 Offer》《图解算法数据结构》《Krahets 笔面试精选 88 题》Python, Java, C++ 解题代码
★ 8.4kCALM. Python
★ 610KnowLM. An Open-sourced Knowledgable Large Language Model Framework.
★ 1.4kCPR-Coach. Coach-Project
★ 19reformer-pytorch. Reformer, the efficient Transformer, in Pytorch
★ 2.2kCVPR2026-Papers-with-Code. CVPR 2026 论文和开源项目合集
★ 23kActionCLIP. This is the official implement of paper "ActionCLIP: A New Paradigm for Action Recognition"
★ 615pytorch-softdtw-cuda. Fast CUDA implementation of (differentiable) soft dynamic time warping for PyTorch
★ 735fast-transformers. Pytorch library for fast transformer implementations
★ 1.8kTriDet. [CVPR2023] Code for the paper, TriDet: Temporal Action Detection with Relative Boundary Modeling
★ 219ms-tcn. Python
★ 260pytorch-i3d. Python
★ 1.1kpytorch-i3d.
★ 1C2F-TCN. Official implementation of Coarse to Fine Multi-Resolution Temporal Convolutional Network for Temporal Action Segmentation
★ 21MS-TCN2. MS-TCN++: Multi-Stage Temporal Convolutional Network for Action Segmentation (TPAMI 2020)
★ 190ASFormer. Official repo for BMVC2021 paper ASFormer: Transformer for action segmentation
★ 139UVAST. Official Implementation of the paper "Unified Fully and Timestamp Supervised Temporal Action Segmentation via Sequence to Sequence Translation" (ECCV 2022)
★ 40Hands-on-GEMM. Cuda
★ 157cholect50. A repository for surgical action triplet dataset. Data are videos of laparoscopic cholecystectomy that have been annotated with <instrument, verb, target> labels for every surgical fine-grained activity.
★ 85Reading-for-Surgical-Vision. Reading list and publicly available datasets for surgical vision
★ 41CUDA_Freshman. Cuda
★ 2.8kMMNet. Python
★ 49cutlass. CUDA Templates and Python DSLs for High-Performance Linear Algebra
★ 10kcuda_sgemm. Cuda
★ 121Tutorials. Parallel programming tutorials
★ 639Chinese-medical-dialogue-data. Chinese medical dialogue data 中文医疗对话数据集
★ 1.8kMING. 明医 (MING):中文医疗问诊大模型
★ 1.2kalpaca-lora. Instruct-tune LLaMA on consumer hardware
★ 19kstanford_alpaca. Code and documentation to train Stanford's Alpaca models, and generate the data.
★ 30kcuda-samples. Samples for CUDA Developers which demonstrates features in CUDA Toolkit
★ 9.5kblock.bootstrap.pytorch. BLOCK (AAAI 2019), with a multimodal fusion library for deep learning models
★ 354DeepIE. DeepIE: Deep Learning for Information Extraction
★ 1.9kGlobalPointer_torch. CMeEE/CBLUE/NER实体识别
★ 135GPLinker_torch. CMeIE/CBLUE/CHIP/实体关系抽取/SPO抽取
★ 243cnSchema. 开放中文知识图谱的schema
★ 832flownet2-pytorch. Pytorch implementation of FlowNet 2.0: Evolution of Optical Flow Estimation with Deep Networks
★ 3.3kTrain_Custom_Dataset. 标注自己的数据集,训练、评估、测试、部署自己的人工智能算法
★ 4.1kCAM. Class Activation Mapping
★ 1.9kaction-modifiers. Code for the CVPR 2020 paper 'Action Modifiers: Learning from Adverbs in Instructional Videos'
★ 23denseflow. Extracting optical flow and frames
★ 319AggPose. [IJCAI 2022] Official PyTorch implementation of AggPose: Deep Aggregation Vision Transformer for Infant Pose Estimation
★ 32multiview_human_dataset. dataset for cvpr2020 "4D Association Graph for Realtime Multi-person Motion Capture Using Multiple Video Cameras"
★ 48YLearn. YLearn, a pun of "learn why", is a python package for causal inference
★ 433hmr-survey. [TPAMI 2023] Recovering 3D Human Mesh from Monocular Images: A Survey
★ 360video_features. Extract video features from raw videos using multiple GPUs. We support RAFT flow frames as well as S3D, I3D, R(2+1)D, VGGish, CLIP, and TIMM models.
★ 654upt. [CVPR'22] Official PyTorch implementation for paper "Efficient Two-Stage Detection of Human–Object Interactions with a Novel Unary–Pairwise Transformer"
★ 167mmdetection. OpenMMLab Detection Toolbox and Benchmark
★ 33kcuda_programming. Code from the "CUDA Crash Course" YouTube series by CoffeeBeforeArch
★ 966Hierarchical-Task-Modeling. Python
★ 8actseg. Action Segmentation Utilities: Evaluation and More ...
★ 5MuCon. Official Implementation for "Fast Weakly Supervised Action Segmentation Using Mutual Consistency" - TPAMI 2021
★ 21TimeSformer. The official pytorch implementation of our paper "Is Space-Time Attention All You Need for Video Understanding?"
★ 1.9kawesome-public-datasets. A topic-centric list of HQ open datasets.
★ 78kFineDiving. FineDiving: A Fine-grained Dataset for Procedure-aware Action Quality Assessment
★ 153pytorch-parallel. Optimize an example model with Python, CPP, and CUDA extensions and Ring-Allreduce.
★ 110matrix-operations. some common used matrix operations (for artificial neural network model)
★ 1matrix-cuda. matrix multiplication in CUDA
★ 126step-by-step-neural-network. step-by-step, implement a neural network with python and numpy to recognize handwritten number
★ 8My-First-CUDA-Code. The introduction to cuda, a simple and easy cuda project
★ 23vit-pytorch. Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch
★ 25kApollo-11. Original Apollo 11 Guidance Computer (AGC) source code for the command and lunar modules.
★ 72kAwesome-Temporal-Action-Localization. A curated list of temporal action localization/detection and related area (e.g. temporal action proposal) resources.
★ 587RTD-Action. [ICCV 2021] Relaxed Transformer Decoders for Direct Action Proposal Generation
★ 92ZeroShotVideoClassification. Zero-shot video classification by end-to-end training of 3D convolutional neural networks
★ 151BSN-boundary-sensitive-network. Codes of our paper: "BSN: Boundary Sensitive Network for Temporal Action Proposal Generation"
★ 411Bridge-Prompt. [CVPR2022] Bridge-Prompt: Towards Ordinal Action Understanding in Instructional Videos
★ 102TAdaConv. [ICLR 2022] TAda! Temporally-Adaptive Convolutions for Video Understanding. This codebase provides solutions for video classification, video representation learning and temporal detection.
★ 246