Belgium

NielsRogge

Elite
@NielsRogge

ML @HuggingFace. Interested in deep learning, NLP. Contributed 40+ models to HuggingFace Transformers

Transformers-Tutorials. This repository contains demos I made with the Transformers library by HuggingFace.

12k

Vision-Transformer-papers. This repository contains an overview of important follow-up works based on the original Vision Transformer (ViT) by Google.

202

tutorials. A repository containing general tutorials I'd like to share with the world.

79

transformers. 🤗Transformers: State-of-the-art Natural Language Processing for Pytorch and TensorFlow 2.0.

50

awesome-huggingface. Repository containing awesome resources regarding Hugging Face tooling.

49

Description2Process. Transforming textual descriptions into process models using deep learning

15

coco-eval. A tiny package supporting distributed computation of COCO metrics for PyTorch models.

15

NielsRogge. Short README about myself.

13

unilm. UniLM - Unified Language Model Pre-training / Pre-training for NLP and Beyond

11

tapas_utils. A package containing utils for the PyTorch version of the Tapas algorithm.

11

diffusion-notes. Some notes I took when learning about diffusion models.

8

CogVLM. a state-of-the-art-level open visual language model

8

yolov9. Implementation of paper - YOLOv9: Learning What You Want to Learn Using Programmable Gradient Information

7

Open-Sora. Open-Sora: Democratizing Efficient Video Production for All

5

Depth-Anything. Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data

5

notebooks. Notebooks using the Hugging Face libraries 🤗

5

LLaVA. Large Language-and-Vision Assistant built towards multimodal GPT-4 level capabilities.

5

mistral-src. Reference implementation of Mistral AI 7B v0.1 model.

4

rf-detr. RF-DETR is a real-time object detection model architecture developed by Roboflow, released under the Apache 2.0 license.

4

Matcha-TTS. [ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching

4

rasa-chatbot. JavaScript

4

big_vision. Official codebase used to develop Vision Transformer, MLP-Mixer, LiT and more.

4

MedSAM. The official repository for MedSAM: Segment Anything in Medical Images.

4

ImageBind. ImageBind One Embedding Space to Bind Them All

4

ml-aim. This repository provides the code and model checkpoints of the research paper: Scalable Pre-training of Large Autoregressive Image Models

3

agents-poc. Python

3

UniDepth. Universal Monocular Metric Depth Estimation

3

VideoMamba. VideoMamba: State Space Model for Efficient Video Understanding

3

blog. Public repo for HF blog posts

3

MobileSAM. This is the official code for MobileSAM project that makes SAM lightweight for mobile applications and beyond!

3

vstar. PyTorch Implementation of "V* : Guided Visual Search as a Core Mechanism in Multimodal LLMs"

3

VoiceCraft. Zero-Shot Speech Editing and Text-to-Speech in the Wild

3

Deep-Learning-for-Computer-Vision. This repository contains the assignments I made during the 2019 version of the Deep Learning for Computer Vision course taught at the University of Michigan.

3

micro_diffusion. Official repository for our work on micro-budget training of large-scale diffusion models.

3

yolov10. YOLOv10: Real-Time End-to-End Object Detection

3

table-transformer. Model training and evaluation code for our dataset PubTables-1M, developed to support the task of table extraction from unstructured documents.

3

textract. extract text from any document. no muss. no fuss.

3

nougat. Implementation of Nougat Neural Optical Understanding for Academic Documents

3

GaNDLF. A generalizable application framework for segmentation, regression, and classification using PyTorch

3

alignment-handbook. Robust recipes for to align language models with human and AI preferences

3

DarkIR. DarkIR: Robust Low-Light Image Restoration [Official PyTorch Implementation]

2

axcell. Tools for extracting tables and results from Machine Learning papers

2

VisualQuality-R1. VisualQuality-R1 is the first open-sourced NR-IQA model can accurately describe and rate the image quality.

2

verl. veRL: Volcano Engine Reinforcement Learning for LLM

2

whispering. Rust

2

alltracker. Python

2

MagicDriveDiT. Official implementation of the paper “MagicDriveDiT: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control”

2

huggingface.js. Utilities to use the Hugging Face Hub API

2

releasing-research-code. Tips for releasing research code in Machine Learning (with official NeurIPS 2020 recommendations)

2

GST. Official implementation of "GST: Precise 3D Human Body from a Single Image with Gaussian Splatting Transformers"

2

scikit-image. Image processing in Python

2

trl. Train transformer language models with reinforcement learning.

2

optimum. 🚀 Accelerate training and inference of 🤗 Transformers and 🤗 Diffusers with easy to use hardware optimization tools

2

AudioSep. Official implementation of "Separate Anything You Describe"

2

CrossMAE. Official Implementation of the CrossMAE paper: Rethinking Patch Dependence for Masked Autoencoders

2
55
Apply