This is your work, valued
multimodal-distillation. Codebase for "Multimodal Distillation for Egocentric Action Recognition" (ICCV 2023)
★ 32revisiting-spatial-temporal-layouts. Codebase for "Revisiting spatio-temporal layouts for compositional action recognition" (Oral at BMVC 2021).
★ 27sr-bert. Codebase for "Decoding language spatial relations to 2D spatial arrangements" (Findings of EMNLP 2020).
★ 11pruning_deep_nets. Layer-wise pruning of a neural network implemented in Tensorflow 1.x.
★ 7text2atlas. Codebase for "Learning to ground medical text in a 3D human atlas (CoNLL 2020)".
★ 6dave. Codebase for "DAVE: Diagnostic benchmark for Audio Visual Evaluation" (NeurIPS 2025 Datasets & Benchmarks)
★ 6vsepp_tensorflow. Implementation of "VSE++: Improving Visual-Semantic Embeddings with Hard Negatives" in Tensorflow.
★ 5SMHA. My master thesis: Siamese multi-hop attention for cross-modal retrieval.
★ 5cross_modal_full_transfer. PyTorch code for cross-modal-retrieval on Flickr8k/30k using Bert and EfficientNet
★ 4mini-bert. A simplified PyTorch implementation of BERT.
★ 1macchina. Codebase for "Self-supervised context-aware Covid-19 document exploration through atlas grounding" as well as links to the tools mentioned in the paper. Work done within ESAT-PSI at KU Leuven.
★ 1gpu-monitor. Python
★ 1SteeringTokens. Repository for the ACL 2026 paper "Compositional Steering of Large Language Models with Steering Tokens": https://arxiv.org/pdf/2601.05062
★ 4Synergy. Python
★ 1MLB. ◉ Improving Multimodal Learning with Multi-Loss Gradient Modulation (BMVC 2024) ◉ Self-Balancing Multimodal Models via Multi-Loss Gradient Modulation (IJCV 2025)
★ 3CoRe-Sleep. ◉ CoRe-Sleep: A Multimodal Fusion Framework for Time Series Robust to Imperfect Modalities (TNSRE 2024)
★ 2nanocode. Minimal Claude Code alternative. Single Python file, zero dependencies, ~250 lines.
★ 2.5kSpecGloss-GS. [WACV 2026 Oral] Spec-Gloss Surfels and Normal–Diffuse Priors for Relightable Glossy Objects
★ 12nanochat. The best ChatGPT that $100 can buy.
★ 57kMCR. ◉ Balancing Multimodal Training Through Game-Theoretic Regularization (Spotlight NeurIPS 2025)
★ 12ocebo. Object-Centric Pretraining via Target Encoder Bootstrapping (ICLR 2025)
★ 11ai-deadlines. ⏰ AI conference deadline countdowns
★ 346simzss. This repo contains the code for the ICLR 2025 paper: "A Simple Framework for Open-Vocabulary Zero-Shot Segmentation".
★ 7label-shift-calibration. Codebase for "LaSCal: Label-Shift Calibration without target labels", published at NeurIPS 2024.
★ 5aisuite. Simple, unified interface to multiple Generative AI providers
★ 16kBay-CAT. [ECCV’24] Official Implementation for CAT: Enhancing Multimodal Large Language Model to Answer Questions in Dynamic Audio-Visual Scenarios
★ 59proper-calibration-error. Codebase for "Consistent and Asymptotically Unbiased Estimation of Proper Calibration Errors", published at AISTATS 2024.
★ 5prismatic-vlms. A flexible and efficient codebase for training visually-conditioned language models (VLMs)
★ 1kchug. Minimal sharded dataset loaders, decoders, and utils for multi-modal document, image, and text datasets.
★ 163CrIBo. This repo contains the code for the ICLR 2024 paper: "CrIBo: Self-Supervised Learning via Cross-Image Object-Level Bootstrapping".
★ 12Babel-ImageNet. Jupyter Notebook
★ 23tensordict. TensorDict is a pytorch dedicated tensor container.
★ 1kcalibration-object-detection. Codebase for "Beyond Classification: Definition and Density-based Estimation of Calibration in Object Detection", published at WACV 2024.
★ 8WEHI-ResearchComputing.github.io. This is the website for the RCP at WEHI
★ 4abstention. Algorithms for abstention, calibration and domain adaptation to label shift.
★ 38SynthIE. The data and the PyTorch implementation for the models and experiments in the paper "Exploiting Asymmetry for Synthetic Training Data Generation: SynthIE and the Case of Information Extraction".
★ 65fact-linking. Codebase for "Linking Surface Facts to Large-Scale Knowledge Graphs" (EMNLP 2023)
★ 12AdaSim. This repo contains the code for the ICCV 2023 paper: "Adaptive Similarity Bootstrapping for Self-Distillation based Representation Learning".
★ 7global-local-self-distillation. This repo contains the code for the WACV 2023 paper Global-Local Self-Distillation for Visual Representation Learning.
★ 6motion-prediction-tim. This repo contains the code for the ACCV 2020 paper Motion Prediction using Temporal Inception Module.
★ 19CrOC. This repo contains the code for the CVPR 2023 paper: "CrOC : Cross-View Online Clustering for Dense Visual Representation Learning".
★ 19mBLIP. Python
★ 88indexed_gzip. Fast random access of gzip files in Python
★ 115wikimapper. Mapping Wikipedia pages to Wikidata IDs and vice versa.
★ 175faiss. A library for efficient similarity search and clustering of dense vectors.
★ 41kconfidence-intervals. Official code repository for "Distribution-Independent Confidence Intervals for the Eigendecomposition of Covariance Matrices via the Eigenvalue-Eigenvector Identity" (ICML 2021 Workshop on Distribution-Free Uncertainty Quantification).
★ 6explanatory-guided-learning. Codebase for “Machine Guides, Human Supervises: Interactive Learning with Global Explanations”.
★ 4ece-kde. Codebase for "A Consistent and Differentiable Lp Canonical Calibration Error Estimator", published at NeurIPS 2022.
★ 16revisiting-spatial-temporal-layouts. Codebase for "Revisiting spatio-temporal layouts for compositional action recognition" (Oral at BMVC 2021).
★ 27pixel. Research code for pixel-based encoders of language (PIXEL)
★ 348rezero. Official PyTorch Repo for "ReZero is All You Need: Fast Convergence at Large Depth"
★ 416refer. Referring Expression Datasets API
★ 573MTCN. Implementation of "With a Little Help from my Temporal Context: Multimodal Egocentric Action Recognition, BMVC, 2021" in PyTorch
★ 20scene_graph_benchmark. image scene graph generation benchmark
★ 402startjekyll. An example and guide to getting started with Jekyll and static site generators.
★ 119ConvNeXt. Code release for ConvNeXt model
★ 6.4kimg2dataset. Easily turn large sets of image urls to an image dataset. Can download, resize and package 100M urls in 20h on one machine.
★ 4.4kcalibration_and_bias. Codebase for "On the relationship between calibrated predictors and unbiased volume estimation" (MICCAI 2021).
★ 9clever. The Curious Layperson: Fine-Grained Image Recognition without Expert Labels (BMVC 2021 best student paper)
★ 25SimCSE. [EMNLP 2021] SimCSE: Simple Contrastive Learning of Sentence Embeddings https://arxiv.org/abs/2104.08821
★ 3.7kbertviz. BertViz: Visualize Attention in Transformer Models
★ 8.1kperceiver-pytorch. Implementation of Perceiver, General Perception with Iterative Attention, in Pytorch
★ 1.2kx-transformers. A concise but complete full-attention transformer with a set of promising experimental features from various papers
★ 5.9kMultimodal-Alignment-Framework. Implementation for MAF: Multimodal Alignment Framework
★ 45paperswithcode-client. API Client for paperswithcode.com
★ 189einops. Flexible and powerful tensor operations for readable and reliable code (for pytorch, jax, TF and others)
★ 9.6kdo-you-even-need-attention. Is the attention layer even necessary? (https://arxiv.org/abs/2105.02723)
★ 485SceneGraphParser. A python toolkit for parsing captions (in natural language) into scene graphs (as symbolic representations).
★ 595accelerate. 🚀 A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (including fp8), and easy-to-configure FSDP and DeepSpeed support
★ 9.8ktorchtyping. Type annotations and dynamic checking for a tensor's shape, dtype, names, etc.
★ 1.5knotation. TeX
★ 1083D-ResNets-PyTorch. 3D ResNets for Action Recognition (CVPR 2018)
★ 4kTRN-pytorch. Temporal Relation Networks
★ 789glances. Glances an Eye on your system. A top/htop alternative for GNU/Linux, BSD, Mac OS and Windows operating systems.
★ 33kMiDaS. Code for robust monocular depth estimation described in "Ranftl et. al., Towards Robust Monocular Depth Estimation: Mixing Datasets for Zero-shot Cross-dataset Transfer, TPAMI 2022"
★ 5.4kno_frills_hoi_det. A strong HOI Detection model without Frills!
★ 59detectron2. Detectron2 is a platform for object detection, segmentation and other visual recognition tasks.
★ 35kdetr. End-to-End Object Detection with Transformers
★ 15kffmpeg-python. Python bindings for FFmpeg - with complex filtering support
★ 11kvit-pytorch. Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch
★ 25kS3D. Release of the pretrained S3D Network in PyTorch (ECCV 2018)
★ 138something_else. Code repository for the paper: 'Something-Else: Compositional Action Recognition with Spatial-Temporal Interaction Networks'
★ 148awesome-shell. A curated list of awesome command-line frameworks, toolkits, guides and gizmos. Inspired by awesome-php.
★ 37kexa. A modern replacement for ‘ls’.
★ 24kminGPT. A minimal PyTorch re-implementation of the OpenAI GPT (Generative Pretrained Transformer) training
★ 25kPyTorch-GAN. PyTorch implementations of Generative Adversarial Networks.
★ 17klatexify_py. A library to generate LaTeX expression from Python code.
★ 7.6klocalized-narratives. Localized Narratives
★ 86dlbook_exercises. Exercises for the Deep Learning textbook at www.deeplearningbook.org
★ 1.4kpapers-with-annotations. Research papers with annotations, illustrations and explanations
★ 826pytorch-image-models. The largest collection of PyTorch image encoders / backbones. Including train, eval, inference, export scripts, and pretrained weights -- ResNet, ResNeXT, EfficientNet, NFNet, Vision Transformer (ViT), MobileNetV4, MobileNet-V3 & V2, RegNet, DPN, CSPNet, Swin Transformer, MaxViT, CoAtNet, ConvNeXt, and more
★ 37kdatasets. 🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools
★ 22kopen_lth. A repository in preparation for open-sourcing lottery ticket hypothesis code.
★ 640wttr.in. :partly_sunny: The right way to check the weather
★ 30kgit-lfs. Git extension for versioning large files
★ 14kawesome-papers. Papers & presentation materials from Hugging Face's internal science day
★ 2.1kgdown. Google Drive public file downloader when curl/wget fails.
★ 5.3kNYU-DLSP20. NYU Deep Learning Spring 2020
★ 6.8ktorchlayers. Shape and dimension inference (Keras-like) for PyTorch layers and neural networks
★ 570d2l-en. Interactive deep learning book with multi-framework code, math, and discussions. Adopted at 500 universities from 70 countries including Stanford, MIT, Harvard, and Cambridge.
★ 29kvoila. Voilà turns Jupyter notebooks into standalone web applications
★ 5.9kvirus-spreading. Simple virus spreading simulation tool made with plain/vanilla JavaScript
★ 83solt. Streaming over lightweight data transformations
★ 268Made-With-ML. Learn how to develop, deploy and iterate on production-grade ML applications.
★ 49kknockknock. 🚪✊Knock Knock: Get notified when your training ends with only two additional lines of code
★ 2.8kfastbook. The fastai book, published as Jupyter Notebooks
★ 25kbat. A cat(1) clone with wings.
★ 60kjax. Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and more
★ 36kpytorch-metric-learning. The easiest way to use deep metric learning in your application. Modular, flexible, and extensible. Written in PyTorch.
★ 6.3kconceptual-captions. Conceptual Captions is a dataset containing (image-URL, caption) pairs designed for the training and evaluation of machine learned image captioning systems.
★ 567brave-browser. Brave browser for Android, iOS, Linux, macOS, Windows.
★ 23kvscode. Visual Studio Code
★ 188k