This is your work, valued
VisualBilibili. 爬取B站up视频详细信息,并进行可视化
★ 105ImageProcess. 数字图像处理,vue+Django restframework
★ 26SAM-webui. segment anything webui
★ 21RL-snack. 强化学习贪吃蛇
★ 17SearchEngine. 新闻搜索引擎,实现倒排索引等基本功能
★ 6SentimentClassify. 情感分类LSTM模型
★ 1of-lyre. 21键游戏乐器演奏程序, 支持合奏, 适用于开放空间, 原神等, 具有可共享曲谱的在线曲库, 使用midi作为输入.
★ 10Understand-Anything. Graphs that teach > graphs that impress. Turn any code into an interactive knowledge graph you can explore, search, and ask questions about. Works with Claude Code, Codex, Cursor, Copilot, Gemini CLI, and more.
★ 77kDoubaoFreeApi. 一个轻量级豆包API代理服务
★ 87xiaozhi-esp32-server. 本项目为xiaozhi-esp32提供后端服务,帮助您快速搭建ESP32设备控制服务器。Backend service for xiaozhi-esp32, helps you quickly build an ESP32 device control server.
★ 10kmusicdl. Musicdl: A lightweight music downloader written in pure python. (轻量级无损音乐下载器,支持数十个音乐/有声读物平台,例如网易云音乐,QQ音乐,酷狗音乐,酷我音乐,咪咕音乐,千千静听,汽水音乐,Bilibili,街声,喜马拉雅,懒人听书,荔枝FM,蜻蜓FM,JOOX,TIDAL,YouTube,Apple Music,Spotify,Qobuz,SoundCloud等主流音乐平台)
★ 5.5kMusic-Xiaozhi-ESP32-Server. 音乐小智AI服务端,仅供参考学习,尚不完善
★ 18of-ps. overfield emu server
★ 11LoloResource. 箱庭游戏资源文件
★ 4OverField-Music-Player. 🎹 Master the stage in OverField (开放空间)! Automated Piano & Guitar player with MIDI support, built with AutoHotkey v2.
★ 2Lolo. 开放空间/OverField 服务端部分实现
★ 17pyimport2pkg. 🐍 Reverse mapping tool: from Python import statements to pip package names. Perfect for AI-generated code!
★ 8LocalAIVoiceChat. Local AI talk with a custom voice based on Zephyr 7B model. Uses RealtimeSTT with faster_whisper for transcription and RealtimeTTS with Coqui XTTS for synthesis.
★ 726BiliTools. 本项目已停止维护。
★ 5.2kMusicFreeDesktop. 插件化、定制化、无广告的免费音乐播放器
★ 8.6kMicroverse. A god-simulation sandbox game built on Godot 4 as a multi-agent AI social simulation system. In this virtual world, AI characters possess independent thinking and memory, capable of autonomous social interactions, task completion, and developing complex social relationships through continuous communication.
★ 2.4kimage-to-slideshow-video-maker. The program is created to make slideshow videos from the images. you can add speech and background music with this project. thank you!
★ 16lrcget. Utility for mass-downloading LRC synced lyrics for your offline music library.
★ 3klrclib. LRCLIB server written in Rust with Axum and SQLite3 database
★ 1.9kDanceba. [ICCV 2025] Official code for "Align Your Rhythm: Generating Highly Aligned Dance Poses with Gating-Enhanced Rhythm-Aware Feature Representation"
★ 97piano_transcription_inference. Python
★ 474MeowField_AutoPiano. 自动弹琴软件,可自动转换音频文件为MIDI谱,适配开发空间PC端乐器自动演奏
★ 82OverField_Auto_Piano. OverField 开放空间 自动弹琴工具
★ 6tifffile. Read and write TIFF files.
★ 659Kontext-Lora-Trainer. Python
★ 67ComfyUI-TeaCache. Python
★ 1.1kfacefusion. Industry leading face manipulation platform
★ 29kclash-for-linux. 🐧 一个更完整、更优雅的 Linux Clash / Mihomo 代理运行平台
★ 5.7kComfyUI-nunchaku. ComfyUI Plugin of Nunchaku
★ 2.9kLaplacian-Pyramids. Image blending by using Gaussian and Laplacian pyramids
★ 28IOPaint. Image inpainting tool powered by SOTA AI Model. Remove any unwanted object, defect, people from your pictures or erase and replace(powered by stable diffusion) any thing on your pictures.
★ 23kMVANet. Multi-view Aggregation Network for Dichotomous Image Segmentation (CVPR24, Highlight)
★ 194BEN2. Python
★ 280CogVideo. text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
★ 13kComfyui-In-Context-Lora-Utils. Python
★ 245cos-python-sdk-v5. Python
★ 216gpu-monitor. JavaScript
★ 3ComfyUI-WD14-Tagger. A ComfyUI extension allowing for the interrogation of booru tags from images.
★ 1.2kChineseREADME. 📖中文标准README
★ 3Awesome-Try-On-Models. A repository for organizing papers, codes and other resources related to Virtual Try-on Models
★ 439ComfyUI_SLK_joy_caption_two. ComfyUI Node
★ 722ComfyUI-FluxTrainer. Python
★ 1.2kai-research. Settings for AI Training
★ 66ComfyUI-PuLID-Flux. PuLID-Flux ComfyUI implementation
★ 129guide-to-write-comfyui-custom-node. 一个创建comfyui自定义节点的指南(guide to write comfyui custom node,tutorial)
★ 66ComfyUI-IPAdapter-Flux. Python
★ 474ComfyUI_IPAdapter_plus. Python
★ 6.1kComfyUI_CatVTON_Wrapper. CatVTON warpper for ComfyUI
★ 373sd-webui-api-payload-display. Display the corresponding API payload after each generation on WebUI
★ 203InstantStyle. InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation 🔥
★ 2kComfyUI. The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
★ 123kComfyUI-Manager. ComfyUI-Manager is an extension designed to enhance the usability of ComfyUI. It offers management functions to install, remove, disable, and enable various custom nodes of ComfyUI. Furthermore, this extension provides a hub feature and convenience functions to access a wide range of information within ComfyUI.
★ 16kai-toolkit. The ultimate training toolkit for finetuning diffusion models
★ 11kgenerative-ai-for-beginners. 21 Lessons, Get Started Building with Generative AI
★ 114kself-llm. 《开源大模型食用指南》针对中国宝宝量身打造的基于Linux环境快速微调(全参数/Lora)、部署国内外开源大模型(LLM)/多模态大模型(MLLM)教程
★ 31ksd-scripts. Python
★ 7.2kllama_index. LlamaIndex is the leading document agent and OCR platform
★ 51kcgft-llm. Practice to LLM.
★ 2.5kFollowYourPose. [AAAI 2024] Follow-Your-Pose: This repo is the official implementation of "Follow-Your-Pose : Pose-Guided Text-to-Video Generation using Pose-Free Videos"
★ 1.4kMusePose. MusePose: a Pose-Driven Image-to-Video Framework for Virtual Human Generation
★ 2.7kAIAssistOnnx. 这是AI游戏助手的最新版代码库,其他的代码库已停止更新。AI游戏助手截取游戏屏幕进行对象识别,模拟自动瞄准/自动开枪等功能,提升玩家的游戏体验。这个版本使用onnx+yoyov6进行AI推理,微软的onnx在个人电脑上运行效率很高,配合美团的yoyov6对象检测模型,游戏图像检测速度快到飞起。
★ 139minimind. 🧠「大模型」2小时完全从0训练64M的小参数LLM!Train a 64M-parameter LLM from scratch in just 2h!
★ 54knvitop. An interactive NVIDIA-GPU process viewer and beyond, the one-stop solution for GPU process management.
★ 7.1kInternVL. [CVPR 2024 Oral] InternVL Family: A Pioneering Open-Source Alternative to GPT-4o. 接近GPT-4o表现的开源多模态对话模型
★ 10kCatVTON. [ICLR 2025] CatVTON is a simple and efficient virtual try-on diffusion model with 1) Lightweight Network (899.06M parameters totally), 2) Parameter-Efficient Training (49.57M parameters trainable) and 3) Simplified Inference (< 8G VRAM for 1024X768 resolution).
★ 1.8kawesome-diffusion-categorized. collection of diffusion model papers categorized by their subareas
★ 2.2kIDM-VTON-train. Python
★ 66Image-to-Image-Search. A reverse image search engine powered by elastic search and tensorflow
★ 326search-by-image. Browser extension for reverse image search, available for Chrome, Edge and Safari
★ 3.6ktransfiner. Mask Transfiner for High-Quality Instance Segmentation, CVPR 2022
★ 547pix2gestalt. Code for the paper "pix2gestalt: Amodal Segmentation by Synthesizing Wholes" (CVPR 2024)
★ 208weiboSpider. 新浪微博爬虫,用python爬取新浪微博数据
★ 9.7kdeocclusion. Code for our CVPR 2020 work.
★ 811awesome-virtual-try-on. A curated list of awesome research papers, projects, code, dataset, workshops etc. related to virtual try-on.
★ 3.1kOOTDiffusion-train. Python
★ 139IDM-VTON. [ECCV2024] IDM-VTON : Improving Diffusion Models for Authentic Virtual Try-on in the Wild
★ 5.1kbilibili-API-collect.
★ 20kAwesome-Chinese-LLM. 整理开源的中文大语言模型,以规模较小、可私有化部署、训练成本较低的模型为主,包括底座模型,垂直领域微调及应用,数据集与教程等。
★ 23kSam_LoRA. Low rank adaptation for segmentation anything model (SAM)
★ 251stable-diffusion-webui-forge. Python
★ 13kDownGit. Create GitHub Resource Download Link
★ 2.2kPaddleSeg. Easy-to-use image segmentation library with awesome pre-trained model zoo, supporting wide-range of practical tasks in Semantic Segmentation, Interactive Segmentation, Panoptic Segmentation, Image Matting, 3D Segmentation, etc.
★ 9.4kControlNet. Let us control diffusion models!
★ 34kSam_LoRA. Segment Your Ring (SYR) - Segment Anything model adapted with LoRA to segment rings.
★ 145U-2-Net. The code for our newly accepted paper in Pattern Recognition 2020: "U^2-Net: Going Deeper with Nested U-Structure for Salient Object Detection."
★ 9.8kdeep-learning-for-image-processing. deep learning for image processing including classification and object-detection etc.
★ 26kcloth-segmentation. This repo contains code and a pre-trained model for clothes segmentation.
★ 673winmerge. WinMerge is an Open Source differencing and merging tool for Windows. WinMerge can compare both folders and files, presenting differences in a visual text format that is easy to understand and handle.
★ 9kiodraw. ioDraw is a free online drawing software, which is used to make flow chart, mind map, Gantt chart, whiteboard, mermaid, poster, and more—no registration required. It also features AI-generated charts.
★ 289diffusers. 🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
★ 34kstable_diffusion_dreambooth_inpainting. Stable Diffusion Dreambooth Inpainting Finetuning
★ 67stable-diffusion-webui-dataset-tag-editor. Extension to edit dataset captions for SD web UI by AUTOMATIC1111
★ 735GPT-SoVITS. 1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
★ 60kkohya_ss. Python
★ 13kpng_info_editor. Copy and edit Stable Diffusion image generation data between images
★ 11GPT_API_free. Free ChatGPT&DeepSeek API Key,免费ChatGPT&DeepSeek API。免费接入DeepSeek API和GPT4 API,支持 gpt | deepseek | claude | gemini | grok 等排名靠前的常用大模型。
★ 39kstable-diffusion-webui. Stable Diffusion web UI
★ 164kopenpose-editor. Openpose Editor for AUTOMATIC1111's stable-diffusion-webui
★ 1.8kChatGLM-6B. ChatGLM-6B: An Open Bilingual Dialogue Language Model | 开源双语对话语言模型
★ 41kSadTalker. [CVPR 2023] SadTalker:Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face Animation
★ 14kSAM-webui. Python
★ 69huggingface-cloth-segmentation. Huggingface cloth segmentation using U2NET
★ 197Real-ESRGAN. Real-ESRGAN aims at developing Practical Algorithms for General Image/Video Restoration.
★ 36kfacechain. FaceChain is a deep-learning toolchain for generating your Digital-Twin.
★ 9.5kFOPA-Fast-Object-Placement-Assessment. A discriminative object placement approach
★ 38keypoint_rcnn_training_pytorch. How to Train a Custom Keypoint Detection Model with PyTorch (Article on Medium)
★ 93DragDiffusion. [CVPR2024, Highlight] Official code for DragDiffusion
★ 1.3kProgrammers-Overseas-Job-Interview-Handbook. 🏂🏻 程序员海外工作/英文面试手册
★ 4.8kxhs. 基于小红书 Web 端进行的请求封装。https://reajason.github.io/xhs/
★ 2.2kxhs_spider. 使用爬虫抓取小红书信息,并通过企业微信发送给自己
★ 47Deep-Reinforcement-Learning-Algorithms. 32 projects in the framework of Deep Reinforcement Learning algorithms: Q-learning, DQN, PPO, DDPG, TD3, SAC, A2C and others. Each project is provided with a detailed training log.
★ 1kproxy_pool. Python ProxyPool for web spider
★ 24khexo-theme-Chic. An elegant, powerful, easy-to-read Hexo theme.
★ 919