This is your work, valued
Llama2-Code-Interpreter. Make Llama2 use Code Execution, Debug, Save Code, Reuse it, Access to Internet
★ 681GroupFace. https://arxiv.org/abs/2005.10497
★ 72Adaptive-Wing-Loss-for-Robust-Face-Alignment-via-Heatmap-Regression. Adaptive-Wing-Loss-for-Robust-Face-Alignment-via-Heatmap-Regression
★ 56CAGFace. Component Attention Guided Face Super-Resolution Network: CAGFace
★ 28minimal-r1. Python
★ 26EXTD. EXTD :: Extremely Tiny Face Detector via Iterative Filter Reuse
★ 13SPIN. Unofficial implementation of Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models
★ 7Past-as-a-Guide. Project Page of "Past as a Guide: Leveraging retrospective learning for Python code completion"
★ 6RFBNet_for_head_detection. RFBNet for head detection (from https://github.com/ruinmessi/RFBNet)
★ 4tau-retail-rl. End-to-end reinforcement learning for retail domain tasks focused on exchange and cancel actions, inspired by the τ-bench framework.
★ 4LaneNet. Towards End-to-End Lane Detection: an Instance Segmentation Approach Variation
★ 3MultiThor. MultiAgent Cooperative instruction following Env.
★ 3Classify_Cars. simple CNN model classifying 4 classes of car (Keras)
★ 3LLM2Act. Make Large Language Model to Plan,Act,Self-improved
★ 2SSD_from_scratch_pytorch. pytorch version of : https://gluon.mxnet.io/chapter08_computer-vision/object-detection.html#SSD:-Single-Shot-MultiBox-Detector
★ 1minimal-mcts-llm. minimal implementation of mcts-llm
★ 1ChatGPTUtils. gpt4 cli interace
★ 1object_detection_for_retail. an autonomous solution for unmanned shops based on computer vision
★ 1Dog_Bread_Identification-Kaggle-. kaggle competition
★ 1XAI-501-Project. Python
★ 1SDPO. Reinforcement Learning via Self-Distillation (SDPO)
★ 1kself-distillation-analysis. Codebase for the work “Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs?”
★ 75oh-my-openclaw. oh-my-opencode patterns ported to OpenClaw
★ 184openclaw. Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
★ 385kQwen3-TTS. Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voice cloning.
★ 13knanoRLHF. nanoRLHF: from-scratch journey into how LLMs and RLHF really work.
★ 195AI-Scientist-v2. The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search
★ 6.9kToolOrchestra. ToolOrchestra is an end-to-end RL training framework for orchestrating tools and agentic workflows.
★ 753aideml. AIDE: AI-Driven Exploration in the Space of Code. The machine Learning engineering agent that automates AI R&D.
★ 1.5kaudio-ai-hub. The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
★ 948slime. slime is an LLM post-training framework for RL Scaling.
★ 7.7kmiles. Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.
★ 1.8kdeepagents. The batteries-included agent harness.
★ 27kverl. verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
★ 23klarge-scale-lm-tutorials. Large-scale language modeling tutorials with PyTorch
★ 6Qwen3-Omni. Qwen3-omni is a natively end-to-end, omni-modal LLM developed by the Qwen team at Alibaba Cloud, capable of understanding text, audio, images, and video, as well as generating speech in real time.
★ 3.9kDICE-Bench. [ACL 2025] DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues
★ 26DeepResearch. Tongyi Deep Research, the Leading Open-source Deep Research Agent
★ 20kmatcha_tts_e. Jupyter Notebook
★ 4Agent-R1. Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning
★ 1.6kminimal-r1. Python
★ 26gpt-oss. gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI
★ 20knano-vllm. Nano vLLM
★ 15kopenai-cs-agents-demo. Demo of a customer service use case implemented with the OpenAI Agents SDK
★ 6.5kdelayed-streams-modeling. Kyutai's Speech-To-Text and Text-To-Speech models based on the Delayed Streams Modeling framework.
★ 3klerobot-mujoco-tutorial. Jupyter Notebook
★ 589ch-tts-llasa-rl-grpo. Python
★ 51Absolute-Zero-Reasoner. Official Repository of Absolute Zero Reasoner
★ 1.9kcsm-hf. Implementation of Sesame's Conversational Speech Model for Hugging Face Transformers
★ 58ToRL. Python
★ 353LUCY. LUCY: Linguistic Understanding and Control Yielding Early Stage of Her
★ 60OpenManus-RL. A live stream development of RL tunning for LLM agents
★ 4.1kaudio_refiner. Python
★ 3MOSHI_forTPU. Python
★ 3Decision-Transformers-For-Trading. Jupyter Notebook
★ 34simpleRL-reason. Simple RL training for reasoning
★ 3.9kopen-r1. Fully open reproduction of DeepSeek-R1
★ 26ksearch-and-learn. Recipes to scale inference-time compute of open models
★ 1.1kNeurIPS-2024-LLM-Papers. Accepted LLM Papers in NeurIPS 2024
★ 38stable-codec. A family of state-of-the-art Transformer-based audio codecs for low-bitrate high-quality audio coding.
★ 438moshi. Moshi is a speech-text foundation model and full-duplex spoken dialogue framework. It uses Mimi, a state-of-the-art streaming neural audio codec.
★ 11kBayLing-Speech. LLaMA-Omni is a low-latency and high-quality end-to-end speech interaction model built upon Llama-3.1-8B-Instruct, aiming to achieve speech capabilities at the GPT-4o level.
★ 3.1kquiet-star. Code for Quiet-STaR
★ 739vllm. A high-throughput and memory-efficient inference and serving engine for LLMs
★ 88kexamples. A set of examples around pytorch in Vision, Text, Reinforcement Learning, etc.
★ 24kLLM101n. LLM101n: Let's build a Storyteller
★ 38kweave. Weave is a toolkit for developing AI-powered applications, built by Weights & Biases.
★ 1.1kllama3-tokenizer-js. JS tokenizer for LLaMA 3 and LLaMA 3.1
★ 118tau-bench. Code and Data for Tau-Bench
★ 1.4kVLM_manipulation. Manipulation in the simulation and realworld environment. GPT4 and G-DINO based reasoning is utilized for commonsense image reasoning.
★ 1lerobot. 🤗 LeRobot: Making AI for Robotics more accessible with end-to-end learning
★ 26kSolarLLMChatDemo. Full Stack SolarLLM Zero to All
★ 169gym-aloha. A gym environment for ALOHA
★ 215FishNet. Implementation code of the paper: FishNet: A Versatile Backbone for Image, Region, and Pixel Level Prediction, NeurIPS 2018
★ 543ri-motion. This repo contains useful implementations for handling motions maintained by rilab.ku
★ 1evalverse. The Universe of Evaluation. All about the evaluation for LLMs.
★ 235llama-cookbook. Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama model family and using them on various provider services
★ 19kInstructDet. Python
★ 37SPIN. The official implementation of Self-Play Fine-Tuning (SPIN)
★ 1.2kreflect. [CoRL 2023] REFLECT: Summarizing Robot Experiences for Failure Explanation and Correction
★ 108spots. Implementation about 'SPOTS: Stable Placement of Objects with Reasoning in Semi-Autonomous Teleoperation Systems'
★ 13yet-another-gpt-tutorial-v2. Jupyter Notebook
★ 14open_x_embodiment. Jupyter Notebook
★ 2kCLARA-SaGC-Code. Jupyter Notebook
★ 9Retrieval-QA-Benchmark. Benchmark baseline for retrieval qa applications
★ 121yet-another-pytorch-tutorial-v2. Jupyter Notebook
★ 49llm-search. Querying local documents, powered by LLM
★ 659chat-ui. JavaScript
★ 1inst-inpaint. A novel inpainting framework that can remove objects from images based on the instructions given as text prompts.
★ 385Llama2-Code-Interpreter. Make Llama2 use Code Execution, Debug, Save Code, Reuse it, Access to Internet
★ 681yet-another-gpt-tutorial. Jupyter Notebook
★ 42KoVicuna.
★ 122MultiThor. MultiAgent Cooperative instruction following Env.
★ 3skypilot. The AI Compute Platform for frontier teams. SkyPilot turns fragmented AI compute into one AI supercomputer, so frontier AI teams build custom intelligence faster.
★ 10kRILAB-SEG. Python
★ 3KoAlpaca. KoAlpaca: 한국어 명령어를 이해하는 오픈소스 언어모델 (KoAlpaca: An open-source language model to understand Korean instructions)
★ 1.6kthe-algorithm. Source code for the X Recommendation Algorithm
★ 74kflame. Official implementation of the paper "FLAME: Free-form Language-based Motion Synthesis & Editing"
★ 118FILM. Official repository of ICLR 2022 paper FILM: Following Instructions in Language with Modular Methods
★ 128TEACh_FILM. Official Repo of EMNLP 2022 "Don’t Copy the Teacher: Data and Model Challenges in Embodied Dialogue"
★ 8denoising-diffusion-pytorch. Implementation of Denoising Diffusion Probabilistic Model in Pytorch
★ 11kdiffuser. Code for the paper "Planning with Diffusion for Flexible Behavior Synthesis"
★ 1.3kAwesome-LLM-Robotics. A comprehensive list of papers using large language/multi-modal models for Robotics/RL, including papers, codes, and related websites
★ 4.4ksimple-mujoco-usage-v2. Jupyter Notebook
★ 13mdetr. Python
★ 1.1kyolov5. Ultralytics YOLOv5 in PyTorch for object detection, instance segmentation, classification, training, and export.
★ 58kFAI-Model. FriendliAI Model Hub
★ 892022-1-deep-learning-applications.
★ 185Conditional-Motion-In-Betweening. 🕹️ Official Implementation of Conditional Motion In-betweening (CMIB) 🏃
★ 135onnxjs. ONNX.js: run ONNX models using JavaScript
★ 1.8kMotionCLIP. Official Pytorch implementation of the paper "MotionCLIP: Exposing Human Motion Generation to CLIP Space"
★ 498Uncertainty-Aware-Robust-Learning. Jupyter Notebook
★ 6InteractivePickup. Interactive Text2Pickup Network for Natural Language based Human-Robot Collaboration
★ 11BABEL. Code accompanying the BABEL dataset (CVPR 2021).
★ 188mindall-e. PyTorch implementation of a 1.3B text-to-image generation model trained on 14 million image-text pairs
★ 631Mixquality_AL. Final term project for Bayesian Machine Learning Lecture (XAI-623)
★ 3CenterNet. Object detection, 3D detection, and pose estimation using center point detection:
★ 7.6kneat. [ICCV'21] NEAT: Neural Attention Fields for End-to-End Autonomous Driving
★ 329deep-text-recognition-benchmark. Text recognition (optical character recognition) with deep learning methods.
★ 3Transformer-in-Vision. Recent Transformer-based CV and related works.
★ 1.3ktransformers. 🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
★ 163kData_Structures_and_Algorithms_in_Python. :book: Worked Solutions of "Data Structures & Algorithms in Python", written by Michael T. Goodrich, Roberto Tamassia and Michael H. Goldwasser. ✏️
★ 400Human-In-The-Loop-MLN-Project. Python
★ 7peg-in-hole. Python
★ 26pytorch-mnist-VAE. Jupyter Notebook
★ 113Transformer-Explainability. [CVPR 2021] Official PyTorch implementation for Transformer Interpretability Beyond Attention Visualization, a novel method to visualize classifications by Transformer based networks.
★ 2kTransformer-MM-Explainability. [ICCV 2021- Oral] Official PyTorch implementation for Generic Attention-model Explainability for Interpreting Bi-Modal and Encoder-Decoder Transformers, a novel method to visualize any Transformer-based network. Including examples for DETR, VQA.
★ 911Cartoon-StyleGAN. Fine-tuning StyleGAN2 for Cartoon Face Generation
★ 656choicenet. Implementation of ChoiceNet
★ 128AlgorithmArchive. Algorithm / Data Structure archive
★ 3JoCoR. CVPR'20: Combating Noisy Labels by Agreement: A Joint Training Method with Co-Regularization
★ 131KLUE. 📖 Korean NLU Benchmark
★ 602bbopt. Black Box Optimization Methods
★ 14upstage-basic-deeplearning. PyTorch Tutorial for Boostcamp AI Tech
★ 52applied-ml. 📚 Papers & tech blogs by companies sharing their work on data science & machine learning in production.
★ 30kDeepjsccf_Autotune. https://arxiv.org/abs/1911.11174
★ 8irl_rocks. Cool Inverse Reinforcement Learning Papers
★ 124bayes-nn. Lecture notes on Bayesian deep learning
★ 487LossLandscape. Explores the ideas presented in Deep Ensembles: A Loss Landscape Perspective (https://arxiv.org/abs/1912.02757) by Stanislav Fort, Huiyi Hu, and Balaji Lakshminarayanan.
★ 66yet-another-pytorch-tutorial. Yet Another PyTorch Tutorial
★ 12mnist-c. Jupyter Notebook
★ 72Deep-learning-Uncertainty-papers. Uncertainty in Deep Learning Papers
★ 35confidence-aware-learning. Confidence-Aware Learning for Deep Neural Networks (ICML2020)
★ 74injae-kim.github.io. 지식과 경험을 공유하는 개발 블로그
★ 3enjoy-design-pattern. 🏡Java 언어로 배우는 디자인 패턴 입문
★ 14SellyDev. Autonomous-driving delivery robot project : Selly
★ 10gpt-3. GPT-3: Language Models are Few-Shot Learners
★ 16karcface-pytorch. Python
★ 1.9kkaggle-humpback. Code for 3rd place solution in Kaggle Humpback Whale Identification Challenge.
★ 164Adaptive-Wing-Loss-for-Robust-Face-Alignment-via-Heatmap-Regression. Adaptive-Wing-Loss-for-Robust-Face-Alignment-via-Heatmap-Regression
★ 56object_detection_for_retail. an autonomous solution for unmanned shops based on computer vision
★ 1EXTD_Pytorch. Official EXTD Pytorch code
★ 2PoseFix_RELEASE. Official TensorFlow implementation of "PoseFix: Model-agnostic General Human Pose Refinement Network", CVPR 2019
★ 22019-2-OSSP1-3355-3. 2019-2-OSSP1-3355-3
★ 4Realtime_Multi-Person_Pose_Estimation. Code repo for realtime multi-person pose estimation in CVPR'17 (Oral)
★ 5.1kEXTD_Pytorch. Official EXTD Pytorch code
★ 187people-counting-pose. Odin: Pose estimation-based tracking and counting of people in videos
★ 178Classify_Cars. simple CNN model classifying 4 classes of car (Keras)
★ 3iamleejihye.github.io. My portfolio site based on blog
★ 63BERT-pytorch. Google AI 2018 BERT pytorch implementation
★ 6.5k