This is your work, valued

Shen zhen

Robin

Elite
@jianzhnie

LLM, Infra, Reinforcement Learning, Transformers

awesome-instruction-datasets. A collection of awesome-prompt-datasets, awesome-instruction-dataset, to train ChatLLM such as chatgpt 收录各种各样的指令数据集, 用于训练 ChatLLM 模型。

738

awesome-text-to-video. A Survey on Text-to-Video Generation/Synthesis.

737

LLamaTuner. Easy and Efficient Finetuning LLMs. (Supported LLama, LLama2, LLama3, Qwen, Baichuan, GLM , Falcon) 大模型高效量化训练+部署.

620

Open-R1. The open source implementation of DeepSeek-R1. 开源复现 DeepSeek-R1

275

deep-marl-toolkit. MARLToolkit: The Multi-Agent Rainforcement Learning Toolkit. Include implementation of MAPPO, MADDPG, QMIX, VDN, COMA, IPPO, QTRAN, MAT...

170

AutoTabular. Automatic machine learning for tabular data. ⚡🔥⚡

69

GigaGAN. Implementation of GigaGAN in pytorch

58

GroupNorm-MXNet. This is the re-implementation of group normalization in MXNet Symbol,Module and Gluon

23

pyramidbox_pytorch. pytorch实现的Pyramidbox 人脸检测模型, 对原来代码的部分模块进行了修改,更简洁高效

22

RLToolkit. RLToolkit is a flexible and high-efficient reinforcement learning framework. Include implementation of DQN, AC, ACER, A2C, A3C, PG, DDPG, TRPO, PPO, SAC, TD3 and ....

20

RLZero. A clean and easy implementation of MuZero, AlphaZero and Self-Play reinforcement learning algorithms for any game.

17

RFBNet_Pytorch. RFBNet in Pytorch

15

MultimodalTookit. Incorporate Image, Text and Tabular Data with HuggingFace Transformers

13

llmtech. LLMTechSite, 专注于通用人工智能领域的技术生态。

12

TsFormer. TsFormer is a toolbox that implement transformer models on Time series model

11

deep-rl-toolkit. RLToolkit is a flexible and high-efficient reinforcement learning framework. Include implementation of DQN, AC,A2C, A3C, PG, DDPG, TRPO, PPO, SAC, TD3 and ....

9

awesome-open-chatgpt. Open efforts to implement ChatGPT-like models and beyond.

9

AutoTimm. Auto torch image models: train and evaluation

8

S3FD_pytorch. pytorch 实现的S3FD,对原来的代码进行了优化,更简洁高效

8

ScaleRL. ScaleRL is a simple and scalable distributed reinforcement learning framework based on Python and PyTorch

8

DSFD_pytorch. pytorch 实现的DSFD, 更高效更简洁

6

MXNet-im2rec_tutorial. 如何使用mxnet的im2rec函数制作自己的物体检测数据集

6

mini-vllm. A compact implementation of vLLM, designed to demystify the complexities of modern LLM serving systems.

6

machine_learning_notes. 工作学习笔记

5

LLMToolkit. LLMToolkit is a toolkit for NLP(Natural Language Processing) and LLM(Large Language Models) using Pytorch.

5

age_gender_estimation. age_gender_estimation

4

yolov3_pytorch. pytorch实现的yolov3, 对原来代码的数据读取模块进行了修改,更简洁高效, 修复了原来代码的bugs,支持Pytorch-1.1 更高的版本

4

ssd_pytorch. pytorch 版本的SSD实现

4

deep_head_pose. deep_head_pose in pytorch

3

ProteinTransformer. ProteinTransformer is a toolkit using deep learning for protein function annotation

3

AutoCAD. A Reinforcement Learning (RL) simulation environment built for training and evaluating autonomous cyber attack & defense models on simulated networks.

2

diffusion-toolkit. Diffuser-toolkit: All kinds of diffusion models for image and audio generation in PyTorch

2

robin-academic-blog. 🎓 Hugo Academic Theme 创建一个学术网站. https://jianzhnie.github.io/

2

models. Models and examples built with TensorFlow

2

learnc. 用来学习 C ++ 编程项目

2

MultimodalTransformers. lmmtoolkit is a toolkit for Multi-Modal Learning

2

LLMPractice. 主要记录大语言模型、强化学习等前沿AI技术的工程实践和技术积累

2

self_supervised. self-supervised learning

1

RefineDet_Pytorch. RefineDet in Pytorch

1

jianzhnie.

1

jianzhnie.github.io. Robin's homepage. Visit https://jianzhnie.github.io/

1

CropAndResize_pytorch. CropAndResize in pytorch

1

LMMRobot. LMMRobot is a professional end-to-end development framework that uses multimodal large models to enable embodied intelligent robot development.

1

LLMEval. LLMEval is your all-in-one toolkit for evaluating LLMs,suport vllm, sglang as backend

1

RetinaNet_Pytorch. RetinaNet in Pytorch

1

DPSNet. Python

1

FaceBoxes. FaceBoxes in Pytorch

1

retinaface_pytorch. A PyTorch implementation of RetinaFace

1

AutoML-Tools. AutoML-Tools

1

ScaleTorch. A PyTorch toolkit for large model training

1

CyberAttackSimulator. CyberAttackSimulator

1

ascend-llm-ops. Shell

1

oh-my-claude-code. This a template repository

1