This is your work, valued
ChatGLM-Tuning. 基于ChatGLM-6B + LoRA的Fintune方案
★ 3.7kgpt2-quickly. Python
★ 141Trading-Gym. A Trading environment base on Gym
★ 83vscode-writer-assistant. 基于生成模型的vscode中文AI写作助手插件
★ 48Baidu-Hot. 记录每天百度搜索热点
★ 24CPM-TF2Transformer. CPM的Transformer版
★ 17AI-XIUXIAN. Python
★ 12Baidu-Translation-SDK. Baidu Translate SDK in python.
★ 6django-view-limiter. view limiter for django
★ 6peft. 🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
★ 5django-cms-bootstrap. A easy Django CMS project using bootstrap css
★ 4Auto-CLUD. 一键运行 中文语言理解测评基准 CLUE 任务
★ 3Wechat-Oauth-SDK. 基于Wechat Oauth 的Python实现 | Wechat Oauth API And SDK With Python
★ 3gpt2-ml. GPT2 for Multiple Languages, including pretrained models. GPT2 多语言支持, 15亿参数中文预训练模型
★ 2weibo_analyze. weibo spider & machine learning
★ 2Picture-Wall. This is a project that generate a picture by many small picture
★ 2sentence-transformers-tf. sentence-transformers with tensorflow
★ 1trlx. A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
★ 1google-research. Google Research
★ 1Django-Rest-API-Template. A Django restful api web server template
★ 1ComfyUI_IPAdapter_plus. Python
★ 1Ants. A human spider framework base on Django
★ 1the-art-of-debugging. The Art of Debugging Open Book
★ 1.7kdailypaper-skills. 用Agent skills打造我的论文流水线
★ 1.1kungguf. GGUF -> Safetensor converter
★ 26ASI-Evolve. Python
★ 822frequencylaw. Official repository for textual frequency law
★ 1.5ksmolvlm-realtime-webcam. Real-time webcam demo with SmolVLM and llama.cpp server
★ 5.6klobehub. 🤯 LobeHub is your Chief Agent Operator, organizing your agents into 7×24 operations by hiring, scheduling, and reporting on your entire AI team.
★ 81kLangBot. Production-grade platform for building agentic IM bots - 生产级多平台智能机器人开发平台/ Agent、知识库编排、插件系统 / Bots for Discord / Slack / LINE / Telegram / WeChat(企业微信, 企微智能机器人, 公众号) / 飞书 / 钉钉 / QQ / Matrix e.g. Integrated with ChatGPT(GPT), DeepSeek, Dify, n8n, Langflow, Coze, Claude, Gemini, GLM, Ollama, SiliconFlow, Moonshot, openclaw / hermes agent, deerflow
★ 17kagent-lightning. The absolute trainer to light up AI agents.
★ 17klerobot. 🤗 LeRobot: Making AI for Robotics more accessible with end-to-end learning
★ 26kSSDD. Official implementation for SSDD Single-Step Diffusion Decoder for Efficient Image Tokenization.
★ 65so-novel. 小说下载|网文下载 | 网络小说
★ 7.5kxiaohongshu-mcp. MCP for xiaohongshu.com
★ 15kUNO. [ICCV 2025] 🔥🔥 UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioning
★ 1.4kdiffusion-4k. [CVPR 2025] Diffusion-4K: Ultra-High-Resolution Image Synthesis with Latent Diffusion Models
★ 364FlexControl.
★ 6UniVRM. UniVRM is a gltf-based VRM format implementation for Unity. English is here https://vrm.dev/en/ . 日本語 はこちら https://vrm.dev/
★ 3.3kWan2.1. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kThinkPO. Python
★ 17DeepEP. DeepEP: an efficient expert-parallel communication library
★ 9.9kyolov12. [NeurIPS 2025] YOLOv12: Attention-Centric Real-Time Object Detectors
★ 2.9kAwesome-LLM-Judges. ⚖️ Awesome LLM Judges ⚖️
★ 202MetaGPT. 🌟 The Multi-Agent Framework: First AI Software Company, Towards Natural Language Programming
★ 70kJudgeBench. Python
★ 128dify. Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self-hosted, so teams move from prototype to production without rebuilding the stack.
★ 151klm-evaluation-harness. A framework for few-shot evaluation of language models.
★ 13kopik. Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.
★ 21kdanbooru-tag-tree. A tag tree for getting a group of tags or checking what group a tag in
★ 12diffusion-pipe. A pipeline parallel training script for diffusion models.
★ 2kStableAnimator. [CVPR2025] We present StableAnimator, the first end-to-end ID-preserving video diffusion framework, which synthesizes high-quality videos without any post-processing, conditioned on a reference image and a sequence of poses.
★ 1.4kComfyUI-RMBG. A ComfyUI custom node designed for advanced image background removal and object, face, clothes, and fashion segmentation, utilizing multiple models including RMBG-2.0, INSPYRENET, BEN, BEN2, BiRefNet, SDMatte, SAM, SAM2, SAM3 and GroundingDINO.
★ 2kSystemAnimatorOnline. XR Animator, AI-based Full Body Motion Capture and Extended Reality (XR) solution, powered by System Animator Online
★ 1.8kRF-Inversion. Rectified Flow Inversion (RF-Inversion) - ICLR 2025
★ 478Weylus. Use your tablet as graphic tablet/touch screen on your computer.
★ 9.4kGraphite. Community-built comprehensive 2D content creation appplication for graphic design, digital art, and interactive real-time motion graphics powered by a node-based procedural graphics engine
★ 27kctrlora. [ICLR 2025] Codebase for "CtrLoRA: An Extensible and Efficient Framework for Controllable Image Generation"
★ 268krita-ai-diffusion. Streamlined interface for generating images with AI in Krita. Inpaint and outpaint with optional text prompt, no tweaking required.
★ 10kAnimate-X. [ICLR 2025] Animate-X - PyTorch Implementation
★ 305ComfyUI-PyramidFlowWrapper. Python
★ 364comfyui-flux-accelerator. Accelerates Flux.1 image generation, just by using this node.
★ 141models. All my self trained & released AI upscaling models. After gathering and applying over 600 different upscaling models, I learned how to train my own models, and these are the results.
★ 618pose-depot. A collection of ControlNet poses
★ 324lora-scripts. SD-Trainer. LoRA & Dreambooth training scripts & GUI use kohya-ss's trainer, for diffusion model.
★ 6.1kNeural_Gaffer. [NeurIPS 2024] Official code for "Neural Gaffer: Relighting Any Object via Diffusion"
★ 346PhotoPoster. To support and further the research in the field of portrait animation , we are excited to launch PhotoPoster, an open project for pose-driven image generation.
★ 211speech-to-speech. Build local voice agents with open-source models
★ 7.8kControlNeXt. Controllable video and image Generation, SVD, Animate Anyone, ControlNet, ControlNeXt, LoRA
★ 1.6kargilla. Argilla is a collaboration tool for AI engineers and domain experts to build high-quality datasets
★ 5.1kstyle-transfer-comfyui-workflow. A style transfer testing workflow for ComfyUI. Made with 💚 by the CozyMantis squad.
★ 55StabilityMatrix. Multi-Platform Package Manager for Stable Diffusion
★ 8.6kPhotoMaker. PhotoMaker [CVPR 2024]
★ 10kpublic-apis. A collective list of free APIs
★ 453kKolors. Kolors Team
★ 4.6kpyramidinfer. Python
★ 47PowerInfer. High-speed Large Language Model Serving for Local Deployment
★ 9.7kAsyncDiff. [NeurIPS 2024] AsyncDiff: Parallelizing Diffusion Models by Asynchronous Denoising
★ 215ComfyUI_densediffusion. DenseDiffusion custom node for ComfyUI
★ 161InstanceDiffusion. [CVPR 2024] Code release for "InstanceDiffusion: Instance-level Control for Image Generation"
★ 614redka. Redis re-implemented with SQL
★ 4.6kseed-tts-eval. Python
★ 1.6kComfyUI_omost. ComfyUI implementation of Omost
★ 445Omost. Your image is almost there!
★ 7.6kComfyUI-3D-Pack. An extensive node suite that enables ComfyUI to process 3D inputs (Mesh & UV Texture, etc) using cutting edge algorithms (3DGS, NeRF, etc.)
★ 3.8kPhased-Consistency-Model. [NeurIPS 2024] Boosting the performance of consistency models with PCM!
★ 522ComfyUI-Easy-Use. In order to make it easier to use the ComfyUI, I have made some optimizations and integrations to some commonly used nodes.
★ 2.6kMetaCLIP. NeurIPS 2025 Spotlight; ICLR2024 Spotlight; CVPR 2024; EMNLP 2024
★ 1.9kICLR24. Official code for ICLR 2024 paper, "A Hard-to-Beat Baseline for Training-free CLIP-based Adaptation"
★ 86FT-CLIP. CLIP Itself is a Strong Fine-tuner: Achieving 85.7% and 88.0% Top-1 Accuracy with ViT-B and ViT-L on ImageNet
★ 222IC-Light. More relighting!
★ 8.5kChatTTS. A generative speech model for daily dialogue.
★ 40kcomfyui-inpaint-nodes. Nodes for better inpainting with ComfyUI: Fooocus inpaint model for SDXL, LaMa, MAT, and various other tools for pre-filling inpaint & outpaint areas.
★ 1.2kstreaming. A Data Streaming Library for Efficient Neural Network Training
★ 1.5kopen-tts-tracker.
★ 1.1kComfyUI-PuLID-ZHO. Unofficial implementation of PuLID(diffusers) for ComfyUI
★ 239Awesome-Chinese-LLM. 整理开源的中文大语言模型,以规模较小、可私有化部署、训练成本较低的模型为主,包括底座模型,垂直领域微调及应用,数据集与教程等。
★ 23kComfyUI_InstantID. Python
★ 1.8kComfyui-Yolov8. Python
★ 26APISR. APISR: Anime Production Inspired Real-World Anime Super-Resolution (CVPR 2024)
★ 1.1kFace-Upscalers-ONNX. ONNX-Powered Inference for State-of-the-Art Face Upscalers
★ 112ComfyUI-AnimateAnyone-Evolved. Improved AnimateAnyone implementation that allows you to use the opse image sequence and reference image to generate stylized video
★ 561AnimateLCM. [SIGGRAPH ASIA 2024 TCS] AnimateLCM: Computation-Efficient Personalized Style Video Generation without Personalized Video Data
★ 659fastdup. fastdup is a powerful, free tool designed to rapidly generate valuable insights from image and video datasets. It helps enhance the quality of both images and labels, while significantly reducing data operation costs, all with unmatched scalability.
★ 1.9kOpenMoE. A family of open-sourced Mixture-of-Experts (MoE) Large Language Models
★ 1.7kVIRL. (ECCV 2024) Code for V-IRL: Grounding Virtual Intelligence in Real Life
★ 367CapHuman. [CVPR2024] CapHuman: Capture Your Moments in Parallel Universes
★ 99LLM-Uncertainty-Bench. Benchmarking LLMs via Uncertainty Quantification
★ 263X-Pose. [ECCV 2024] Official implementation of the paper "X-Pose: Detecting Any Keypoints"
★ 815HierSpeechpp. The official implementation of HierSpeech++
★ 1.2kaxolotl. Go ahead and axolotl questions
★ 12kMoore-AnimateAnyone. Character Animation (AnimateAnyone, Face Reenactment)
★ 3.5kFocus-on-Your-Instruction. [CVPR 2024] Focus on Your Instruction: Fine-grained and Multi-instruction Image Editing by Attention Modulation
★ 116TinyLlama. The TinyLlama project is an open endeavor to pretrain a 1.1B Llama model on 3 trillion tokens.
★ 9kVCoder. [CVPR 2024] VCoder: Versatile Vision Encoders for Multimodal Large Language Models
★ 280CartoonSegmentation. Instance segmentation for cartoon/anime characters and some visual techniques building around it.
★ 199TheChosenOne. Unofficial implementation of the paper "The Chosen One: Consistent Characters in Text-to-Image Diffusion Models"
★ 270Awesome-Video-Datasets. Video datasets
★ 1.7klitgpt. 20+ high-performance LLMs with recipes to pretrain, finetune and deploy at scale.
★ 14kOpen-AnimateAnyone. Unofficial Implementation of Animate Anyone
★ 2.9kCritiqueLLM. Python
★ 147Qwen-Audio. The official repo of Qwen-Audio (通义千问-Audio) chat & pretrained large audio language model proposed by Alibaba Cloud.
★ 1.9kRobustVideoMatting. Robust Video Matting in PyTorch, TensorFlow, TensorFlow.js, ONNX, CoreML!
★ 9.5kOpenVoice. Instant voice cloning by MIT and MyShell. Audio foundation model.
★ 37kefficientspeech. PyTorch code implementation of EfficientSpeech - to be presented at ICASSP2023.
★ 182Matcha-TTS. [ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching
★ 1.3kwhisper.cpp. Port of OpenAI's Whisper model in C/C++
★ 52kdistilabel. Distilabel is a framework for synthetic data and AI feedback for engineers who need fast, reliable and scalable pipelines based on verified research papers.
★ 3.3kswift-chat. Mac app to demonstrate swift-transformers
★ 596llm-viz. 3D Visualization of an GPT-style LLM
★ 5.5kFIREBALL. Home of FIREBALL: A Dataset of Dungeons and Dragons Actual-Play with Structured Game State Information (ACL 2023)
★ 53LLM-Agent-Paper-List. The paper list of the 86-page SCIS cover paper "The Rise and Potential of Large Language Model Based Agents: A Survey" by Zhiheng Xi et al.
★ 8.2kPixivUtil2. Download images from Pixiv and more!
★ 2.7kGenAI_LLM_timeline. ChatGPT, GenerativeAI and LLMs Timeline
★ 953VALL-E-X. An open source implementation of Microsoft's VALL-E X zero-shot TTS model. Demo is available in https://plachtaa.github.io/vallex/
★ 7.9kVITS-fast-fine-tuning. This repo is a pipeline of VITS finetuning for fast speaker adaptation TTS, and many-to-many voice conversion
★ 5kPokemonRedExperiments. Playing Pokemon Red with Reinforcement Learning
★ 7.9kLanguageAgentTreeSearch. [ICML 2024] Official repository for "Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models"
★ 848sd-webui-EasyPhoto. 📷 EasyPhoto | Your Smart AI Photo Generator.
★ 5.2kagents. An Open-source Framework for Data-centric, Self-evolving Autonomous Language Agents
★ 6kexllamav2. A fast inference library for running LLMs locally on modern consumer-class GPUs
★ 4.6kLLaSM. 第一个支持中英文双语语音-文本多模态对话的开源可商用对话模型。便捷的语音输入将大幅改善以文本为输入的大模型的使用体验,同时避免了基于 ASR 解决方案的繁琐流程以及可能引入的错误。
★ 561SoTaNa. Jupyter Notebook
★ 127MultiCapCLIP. (ACL'2023) MultiCapCLIP: Auto-Encoding Prompts for Zero-Shot Multilingual Visual Captioning
★ 36MockingBird. 🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
★ 37kaudiocraft_plus. Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor / tokenizer, along with MusicGen, a simple and controllable music generation LM with textual and melodic conditioning.
★ 640Platypus. Code for fine-tuning Platypus fam LLMs using LoRA
★ 625AgentBench. A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
★ 3.6kgenerative_agents. Generative Agents: Interactive Simulacra of Human Behavior
★ 22kgpt-researcher. An autonomous agent that conducts deep research on any data using any LLM providers
★ 29kawesome-talking-head-generation.
★ 1.9kLMOps. General technology for enabling AI capabilities w/ LLMs and MLLMs
★ 4.5kSadTalker. [CVPR 2023] SadTalker:Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face Animation
★ 14kvllm. A high-throughput and memory-efficient inference and serving engine for LLMs
★ 88kRecurrentGPT. Official Code for Paper: RecurrentGPT: Interactive Generation of (Arbitrarily) Long Text
★ 998gchar. Crawler and cleaner of data for novelai embedding's training
★ 21llama.onnx. LLaMa/RWKV onnx models, quantization and testcase
★ 368MOSS. An open-source tool-augmented conversational language model from Fudan University
★ 12kFlagInstruct.
★ 173textgen. Open-source desktop app for local LLMs. Text, vision, tool-calling, OpenAI/Anthropic-compatible API. 100% private.
★ 48kInpaint-Anything. Inpaint anything using Segment Anything and inpainting models.
★ 7.7kRRHF. [NIPS2023] RRHF & Wombat
★ 805JARVIS. JARVIS, a system to connect LLMs with ML community. Paper: https://arxiv.org/pdf/2303.17580.pdf
★ 25kmusiclm-pytorch. Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch
★ 3.3kChatGLM-Tuning. 基于ChatGLM-6B + LoRA的Fintune方案
★ 3.7kvisual-openllm. something like visual-chatgpt, 文心一言的开源版
★ 1.2kChatPaper. Use ChatGPT to summarize the arXiv papers. 全流程加速科研,利用chatgpt进行论文全文总结+专业翻译+润色+审稿+审稿回复
★ 20kChatGLM-6B. ChatGLM-6B: An Open Bilingual Dialogue Language Model | 开源双语对话语言模型
★ 41kminimal-llama. Python
★ 456llama_infer. Inference script for Meta's LLaMA models using Hugging Face wrapper
★ 109pix2pix. Image-to-image translation with conditional adversarial nets
★ 11kstable-diffusion-webui. Stable Diffusion web UI
★ 164kControlNet. Let us control diffusion models!
★ 34kattrs. Python Classes Without Boilerplate
★ 5.8kcattrs. Composable custom class converters for attrs, dataclasses and friends.
★ 1kRobyn. Robyn is a Super Fast Async Python Web Framework with a Rust runtime.
★ 7.3kPPHC. 📙《高并发的哲学原理》开源图书(CC BY-NC-ND)https://pphc.lvwenhan.com
★ 4ktrlx. A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
★ 4.8kbitsandbytes. Accessible large language models via k-bit quantization for PyTorch.
★ 8.4kPaLM-rlhf-pytorch. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
★ 7.9ktransformers_tasks. ⭐️ NLP Algorithms with transformers lib. Supporting Text-Classification, Text-Generation, Information-Extraction, Text-Matching, RLHF, SFT etc.
★ 2.4ksafetensors. Simple, safe way to store and distribute tensors
★ 3.8krust-bert. Rust native ready-to-use NLP pipelines and transformer-based models (BERT, DistilBERT, GPT2,...)
★ 3.1kpromptsource. Toolkit for creating, sharing and using natural language prompts.
★ 3kflash-linux0.11-talk. 你管这破玩意叫操作系统源码 — 像小说一样品读 Linux 0.11 核心代码
★ 22kpinyin-ime. Pinyin IME with rust - 拼音输入法rust版
★ 47rime-ice. Rime 配置:雾凇拼音 | 长期维护的简体词库
★ 19kprompt-to-prompt. Jupyter Notebook
★ 3.5kstable-diffusion-streamlit. Quantized stable-diffusion cutting down memory 75%, testing in streamlit, deploying in container
★ 52sentence-transformers-tf. sentence-transformers with tensorflow
★ 1awesome-sanic. A curated list of awesome Sanic resources and extensions
★ 765citybound. A work-in-progress, open-source, multi-player city simulation game.
★ 8.1kwezterm. A GPU-accelerated cross-platform terminal emulator and multiplexer written by @wez and implemented in Rust
★ 28kfastgpt. ⚡ boost inference speed of GPT models in transformers by onnxruntime
★ 51evaluate. 🤗 Evaluate: A library for easily evaluating machine learning models and datasets.
★ 2.5kmotan-go. The golang implementation of Motan
★ 474Fengshenbang-LM. Fengshenbang-LM(封神榜大模型)是IDEA研究院认知计算与自然语言研究中心主导的大模型开源体系,成为中文AIGC和认知智能的基础设施。
★ 4.1kChineseSTS. 中文文本语义相似度(Chinese Semantic Text Similarity)语料库建设
★ 478sanic. Accelerate your web app development | Build fast. Run fast.
★ 19krustpad. Efficient and minimal collaborative code editor, self-hosted, no database required
★ 4.1kColossalAI. Making large AI models cheaper, faster and more accessible
★ 41klightseq. LightSeq: A High Performance Library for Sequence Processing and Generation
★ 3.3kavvvatars. Beautifully crafted unique avatar placeholder for your next react project
★ 2kElectronBot. C
★ 9.4kpoetry. Python packaging and dependency management made easy
★ 34kgotop. A terminal based graphical activity monitor inspired by gtop and vtop
★ 3.1kmagic. CSS3 Animations with special effects
★ 8.6kawesome-robotics. A list of awesome Robotics resources
★ 6.8kwukong-robot. 🤖 wukong-robot 是一个简单、灵活、优雅的中文语音对话机器人/智能音箱项目,支持ChatGPT多轮对话能力,还可能是首个支持脑机交互的开源智能音箱项目。
★ 7.1kopenrave. Open Robotics Automation Virtual Environment: An environment for testing, developing, and deploying robotics motion planning algorithms.
★ 813Auto-CLUD. 一键运行 中文语言理解测评基准 CLUE 任务
★ 3CoSENT. 比Sentence-BERT更有效的句向量方案
★ 371best_AI_papers_2021. A curated list of the latest breakthroughs in AI (in 2021) by release date with a clear video explanation, link to a more in-depth article, and code.
★ 2.9kspark-sklearn. (Deprecated) Scikit-learn integration package for Apache Spark
★ 1.1kmujoco. Multi-Joint dynamics with Contact. A general purpose physics simulator.
★ 14kml-agents. The Unity Machine Learning Agents Toolkit (ML-Agents) is an open-source project that enables games and simulations to serve as environments for training intelligent agents using deep reinforcement learning and imitation learning.
★ 20kminerl. MineRL Competition for Sample Efficient Reinforcement Learning - Python Package
★ 969rocket-recycling. Rocket-recycling with Reinforcement Learning
★ 302automa. A browser extension for automating your browser by connecting blocks
★ 22kextendedsamplecharacter-dontstarvetogether. A template character mod for Don't Starve Together
★ 71SinglepassTextCluster. SinglepassTextCluster, an TextCluster tools based on Singlepass cluster algorithm that use tfidf vector and doc2vec,which can be used for individual real-time corpus cluster task。基于single-pass算法思想的自动文本聚类小组件,内置tfidf和doc2vec两种文本向量方法,可自动输出聚类数目、类簇文档集合和簇类大小,用于自有实时数据的聚类任务。
★ 65secguide. 面向开发人员梳理的代码安全指南
★ 13k