This is your work, valued
ML @HuggingFace. Interested in deep learning, NLP. Contributed 40+ models to HuggingFace Transformers
Transformers-Tutorials. This repository contains demos I made with the Transformers library by HuggingFace.
12kVision-Transformer-papers. This repository contains an overview of important follow-up works based on the original Vision Transformer (ViT) by Google.
202tutorials. A repository containing general tutorials I'd like to share with the world.
79transformers. 🤗Transformers: State-of-the-art Natural Language Processing for Pytorch and TensorFlow 2.0.
50awesome-huggingface. Repository containing awesome resources regarding Hugging Face tooling.
49Description2Process. Transforming textual descriptions into process models using deep learning
15coco-eval. A tiny package supporting distributed computation of COCO metrics for PyTorch models.
15NielsRogge. Short README about myself.
13unilm. UniLM - Unified Language Model Pre-training / Pre-training for NLP and Beyond
11tapas_utils. A package containing utils for the PyTorch version of the Tapas algorithm.
11diffusion-notes. Some notes I took when learning about diffusion models.
8CogVLM. a state-of-the-art-level open visual language model
8yolov9. Implementation of paper - YOLOv9: Learning What You Want to Learn Using Programmable Gradient Information
7Open-Sora. Open-Sora: Democratizing Efficient Video Production for All
5Depth-Anything. Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data
5notebooks. Notebooks using the Hugging Face libraries 🤗
5LLaVA. Large Language-and-Vision Assistant built towards multimodal GPT-4 level capabilities.
5mistral-src. Reference implementation of Mistral AI 7B v0.1 model.
4rf-detr. RF-DETR is a real-time object detection model architecture developed by Roboflow, released under the Apache 2.0 license.
4Matcha-TTS. [ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching
4rasa-chatbot. JavaScript
4big_vision. Official codebase used to develop Vision Transformer, MLP-Mixer, LiT and more.
4MedSAM. The official repository for MedSAM: Segment Anything in Medical Images.
4ImageBind. ImageBind One Embedding Space to Bind Them All
4ml-aim. This repository provides the code and model checkpoints of the research paper: Scalable Pre-training of Large Autoregressive Image Models
3agents-poc. Python
3UniDepth. Universal Monocular Metric Depth Estimation
3VideoMamba. VideoMamba: State Space Model for Efficient Video Understanding
3blog. Public repo for HF blog posts
3MobileSAM. This is the official code for MobileSAM project that makes SAM lightweight for mobile applications and beyond!
3vstar. PyTorch Implementation of "V* : Guided Visual Search as a Core Mechanism in Multimodal LLMs"
3VoiceCraft. Zero-Shot Speech Editing and Text-to-Speech in the Wild
3Deep-Learning-for-Computer-Vision. This repository contains the assignments I made during the 2019 version of the Deep Learning for Computer Vision course taught at the University of Michigan.
3micro_diffusion. Official repository for our work on micro-budget training of large-scale diffusion models.
3yolov10. YOLOv10: Real-Time End-to-End Object Detection
3table-transformer. Model training and evaluation code for our dataset PubTables-1M, developed to support the task of table extraction from unstructured documents.
3textract. extract text from any document. no muss. no fuss.
3nougat. Implementation of Nougat Neural Optical Understanding for Academic Documents
3GaNDLF. A generalizable application framework for segmentation, regression, and classification using PyTorch
3alignment-handbook. Robust recipes for to align language models with human and AI preferences
3DarkIR. DarkIR: Robust Low-Light Image Restoration [Official PyTorch Implementation]
2axcell. Tools for extracting tables and results from Machine Learning papers
2VisualQuality-R1. VisualQuality-R1 is the first open-sourced NR-IQA model can accurately describe and rate the image quality.
2verl. veRL: Volcano Engine Reinforcement Learning for LLM
2whispering. Rust
2alltracker. Python
2MagicDriveDiT. Official implementation of the paper “MagicDriveDiT: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control”
2huggingface.js. Utilities to use the Hugging Face Hub API
2releasing-research-code. Tips for releasing research code in Machine Learning (with official NeurIPS 2020 recommendations)
2GST. Official implementation of "GST: Precise 3D Human Body from a Single Image with Gaussian Splatting Transformers"
2scikit-image. Image processing in Python
2trl. Train transformer language models with reinforcement learning.
2optimum. 🚀 Accelerate training and inference of 🤗 Transformers and 🤗 Diffusers with easy to use hardware optimization tools
2AudioSep. Official implementation of "Separate Anything You Describe"
2CrossMAE. Official Implementation of the CrossMAE paper: Rethinking Patch Dependence for Masked Autoencoders
2