This is your work, valued
Training neural nets for a while
cnn-lstm. CNN LSTM architecture implemented in Pytorch for Video Classification
308head-detection-using-yolo. Detection of head using YOLO
165End-to-End-Trainable-Multi-Instance-Pose-Estimation-with-Transformers. pose-detection-transformer
19text-to-video. Experimental text-to-video generation
13scene-graph-vit. Implementation of the Paper Scene-Graph ViT
10large-scale-visual-relationship-understanding. Visual Relationship Understanding
10lstm-object-detection. Object Detection using LSTM-SSD
7ml-from-scratch. Python
6attention-models. Simplified Implementation of SOTA Deep Learning Papers in Pytorch
5dynamic-Pix2Pix. Dynamic-Pix2Pix: Noise Injected cGAN for Modeling Input and Target Domain Joint Distributions with Limited Training Data
4pix2seq. Python
3TexTok. Python
1D2iT. Python
1token-shuffle. Python
1open-sim. A Tool to convert human videos doing actions (eg : cooking) to realistic robotic videos + actions for training/fine tuning Foundation VLM/VAM models
1