This is your work, valued
Courage is the magic that turns dreams into reality
pixivpy. Pixiv API for Python
★ 2kclone-fastcoll. clone fastcoll_v1.0.0.5_source.zip from http://www.win.tue.nl/hashclash/
★ 80lldb-capstone-arm. Capstone disassemble scripts for lldb
★ 72IOTrackerOnWeb. An HTTP/File I/O tracker tweak for iOS. Just inject the dylib to target App and view log on http://iPhone.local:8080
★ 56My-iDevice-Tools. A set of console tools for iOS devices
★ 56python-imobiledevice_demo. libimobiledevice demo for Python
★ 41PixivAPI_iOS. Pixiv API for IOS
★ 27ue5-ffmpeg. Record game screen and push RTMP in UE5.0
★ 22pyalgotrade-step-by-step. Simples for PyAlgoTrade(https://github.com/gbeced/pyalgotrade)
★ 22GPT-SoVITS. 1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
★ 17iOS_3rdTrackingBlocker. 第三方SDK反跟踪插件,可以查看或屏蔽第三方SDK的上报
★ 16HiddenMarkovModel. Python implementation of Hidden Markov Model, with demo of Chinese Part-of-Speech tagging
★ 16tweibo-pysdk. 腾讯微博Python SDK (Tencent Weibo SDK for Python)
★ 12bloomfilter. This is a stand-alone bloomfilter implementation written in C. Simple but powerful
★ 11PixivBot. Pixiv机器人 for GAE/腾讯微博 (Pixiv bot for GAE/Tencent Weibo)
★ 10Pixiv-RankingLog. P站过去排行扫图专用App - Pixiv RankingLog for iOS
★ 7FakeSMS. createFakeSms() with GSM_UCS2 supported
★ 7cn_segment. Chinese word segmentation based on statistical methods (for Python)
★ 7upbit.github.io. 个人博客
★ 6zwbot. Twitter中文单词机器人 for Google App Engine
★ 6go-echarts. ECharts server written by Golang
★ 5react-native-layout-playground. React Native layout playground
★ 4libimobiledevice. A cross-platform protocol library to communicate with iOS devices
★ 3groupcache. groupcache (with expiration functionality) is a caching and cache-filling library, intended as a replacement for memcached in many cases.
★ 2ExStatsD. An Elixir ports client for Statsd
★ 2Padavan-build-RM2100. Padavan 自动编译
★ 2Anime4K. A High-Quality Real Time Upscaler for Anime Video
★ 2StreamChatPlayground. Yet Another Chat Playground for Large Language Models.
★ 2crash-exporter. Go
★ 2Keychain-Dumper. A tool to check which keychain items are available to an attacker once an iOS device has been jailbroken
★ 1vimrc. My vimrc
★ 1flask_whiteboard. A jieba whiteboard write by Flask
★ 1zserver. Personal Use Only
★ 1ssl-kill-switch2. Blackbox tool to disable SSL certificate validation - including certificate pinning - within iOS and OS X Apps
★ 1goalbatch. A simple way to execute functions asynchronously and waits for results
★ 1opencl-helper. A simple OpenCL example written in C
★ 1Unbiased_LambdaMart. Implementation for "Unbiased LambdaMART: An Unbiased PairwiseLearning-to-Rank Algorithm", merge acbull/Unbiased_LambdaMart changes into microsoft/LightGBM
★ 1HunyuanOCR. HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better
★ 1.9ksystem_prompts_leaks. Extracted system prompts from Anthropic - Claude Fable 5, Opus 5, Claude Design, Claude Code. OpenAI - ChatGPT GPT-5.6-Sol, Codex. Google - Gemini 3.5 Flash, 3.1 Pro, Antigravity. xAI - Grok, Cursor, Copilot, VS Code, Perplexity, and more. Updated regularly.
★ 62kCL4R1T4S. LEAKED SYSTEM PROMPTS FOR CHATGPT, CLAUDE, GEMINI, GROK, PERPLEXITY, CURSOR, LOVABLE, REPLIT, AND MORE! - AI SYSTEMS TRANSPARENCY FOR ALL! 👐
★ 47kMooncake. Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.
★ 6.1kxpu_benchmark. Python
★ 1rtk. CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies
★ 74kVLMEvalKit. Python
★ 23lmms-eval. One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks
★ 4.3kevalscope. A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
★ 3.2kimmich. High performance self-hosted photo and video management solution.
★ 109kclaw-code. An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.
★ 195kOminiX-MLX. MLX implementation of OminiX for LLM, image generataion, ASR and TTS
★ 57CUDA-Agent. CUDA Agent: Large-Scale Agentic RL for High-Performance CUDA Kernel Generation
★ 1.1kzeroclaw. Fast, small, and fully autonomous AI personal assistant infrastructure, any OS, any platform — deploy anywhere, swap anything 🦀
★ 32kautoresearch-macos. AI agents running research on single-GPU nanochat training automatically adopted for MacOS
★ 2.3kautoresearch. AI agents running research on single-GPU nanochat training automatically
★ 93kverl. verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
★ 23kopenclaw. Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
★ 385khpc-ops. High Performance LLM Inference Operator Library
★ 1.1kvllm-omni. A framework for efficient model inference with omni-modality models
★ 5.8ktokenizers. Go, Wasm bindings for HF Tokenizers and Tiktoken
★ 222Qwen3-Omni. Qwen3-omni is a natively end-to-end, omni-modal LLM developed by the Qwen team at Alibaba Cloud, capable of understanding text, audio, images, and video, as well as generating speech in real time.
★ 3.9kRSSHUB-MCP. RSSHUB-MCP
★ 14servers. Model Context Protocol Servers
★ 89khass-xiaomi-miot. Automatic integrate all Xiaomi devices to HomeAssistant via miot-spec, support Wi-Fi, BLE, ZigBee devices. 小米米家智能家居设备接入Hass集成
★ 6kimmich-go. An alternative to the immich-CLI command that doesn't depend on nodejs installation. It tries its best for importing google photos takeout archives.
★ 6.6kComfyUI-GGUF. GGUF Quantization support for native ComfyUI models
★ 3.9kWan2.1. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17k3FS. A high-performance distributed file system designed to address the challenges of AI training and inference workloads.
★ 10kFlashMLA. FlashMLA: Efficient Multi-head Latent Attention Kernels
★ 13kDeepGEMM. DeepGEMM: clean and efficient BLAS kernel library on GPU
★ 7.6ktriton. Development repository for the Triton language and compiler
★ 20kComfy-WaveSpeed. https://wavespeed.ai/ [WIP] The all in one inference optimization solution for ComfyUI, universal, flexible, and fast.
★ 1.2kget-pixivpy-token. Get your Pixiv token easily (for running upbit/pixivpy)
★ 191silero-vad. Silero VAD: pre-trained enterprise-grade Voice Activity Detector
★ 9.8kwaveterm. An open-source, AI-integrated, cross-platform terminal for seamless workflows
★ 22kPDFMathTranslate. [EMNLP 2025 Demo] PDF scientific paper translation with preserved formats - 基于 AI 完整保留排版的 PDF 文档全文双语翻译,支持 Google/DeepL/Ollama/OpenAI 等服务,提供 CLI/GUI/MCP/Docker/Zotero
★ 36kpurego. A library for calling C functions from Go without Cgo
★ 3.8ksherpa-onnx. Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages
★ 14krust2go. Call Between Golang and Rust Asynchronously
★ 419lokirus. A Loki hook for Logrus
★ 20walk. A Windows GUI toolkit for the Go Programming Language
★ 7.1kwebrtc. Pure Go implementation of the WebRTC API
★ 17kmoshi. Moshi is a speech-text foundation model and full-duplex spoken dialogue framework. It uses Mimi, a state-of-the-art streaming neural audio codec.
★ 11kGOT-OCR2.0. Official code implementation of General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model
★ 8.2kPlayCover. Community fork of PlayCover
★ 12kgvisor. Application Kernel for Containers
★ 19kideviceinstaller. Manage apps of iOS devices
★ 1.4kMiniCPM-V. A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone
★ 26kDeep-Live-Cam. real time face swap and one-click video deepfake with only a single image
★ 95kmlx. MLX: An array framework for Apple silicon
★ 28kexo. Run frontier AI locally.
★ 47kblog. my blog
★ 25lancedb. Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less.
★ 11kzincsearch. ZincSearch . A lightweight alternative to elasticsearch that requires minimal resources, written in Go.
★ 18kdejavu. A Web UI for Elasticsearch and OpenSearch: Import, browse and edit data with rich filters and query views, create reference search UIs.
★ 8.5konyx. Open Source AI Platform - AI Chat with advanced features that works with every LLM
★ 31kgo-redis. Redis Go client
★ 22kstructure-vision. Viewer for the structure extracted by Grobid on PDF documents
★ 57pdfplumber. Plumb a PDF for detailed information about each char, rectangle, line, et cetera — and easily extract text and tables.
★ 11kfish-speech. SOTA Open Source TTS
★ 32kdspy. DSPy: The framework for programming—not prompting—language models
★ 36kclink. Bash's powerful command line editing in cmd.exe
★ 5.4kpyannote-audio. Neural building blocks for speaker diarization: speech activity detection, speaker change detection, overlapped speech detection, speaker embedding
★ 10kpdfminer.six. Community maintained fork of pdfminer - we fathom PDF
★ 7kragflow. RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create a superior context layer for LLMs
★ 87kwhisper.cpp. Port of OpenAI's Whisper model in C/C++
★ 52kSWE-agent. SWE-agent takes a GitHub issue and tries to automatically fix it, using your LM of choice. It can also be employed for offensive cybersecurity or competitive coding challenges. [NeurIPS 2024]
★ 20kray. Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
★ 43kAniPortrait. AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation
★ 5ksimulated-trial-and-error. Python
★ 124SmsForwarder. 短信转发器——监控Android手机短信、来电、APP通知,并根据指定规则转发到其他手机:钉钉群自定义机器人、钉钉企业内机器人、企业微信群机器人、飞书机器人、企业微信应用消息、邮箱、bark、webhook、Telegram机器人、Server酱、PushPlus、手机短信等。包括主动控制服务端与客户端,让你轻松远程发短信、查短信、查通话、查话簿、查电量等。(V3.0 新增)PS.这个APK主要是学习与自用,如有BUG请提ISSUE,同时欢迎大家提PR指正
★ 27kgpt-pilot. The first real AI developer
★ 34kUE4-AES-Key-Extracting-Guide. A simple guide on how to extract AES-256 keys and use them to decrypt .pak file(s) in most steam-games.
★ 464continue. open-source coding agent
★ 35kGaLore. GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection
★ 1.7kChatGLM3. ChatGLM3 series: Open Bilingual Chat LLMs | 开源双语对话语言模型
★ 14kComfyUI_InstantID. Python
★ 1.8kfastllm. fastllm是后端无依赖的高性能大模型推理库。同时支持张量并行推理稠密模型和混合模式推理MOE模型,任意10G以上显卡即可推理满血DeepSeek。双路9004/9005服务器+单显卡部署DeepSeek满血满精度原版模型,单并发20tps;INT4量化模型单并发30tps,多并发可达60+。
★ 4.9ksd-forge-layerdiffuse. [WIP] Layer Diffusion for WebUI (via Forge)
★ 4.1kUnrealPakViewer. 查看 UE4 Pak 文件的图形化工具,支持 UE4 pak/ucas 文件
★ 1.4kgpt4free. The official gpt4free repository | various collection of powerful language models | opus 4.6 gpt 5.3 kimi 2.5 deepseek v3.2 gemini 3
★ 67kGPT-SoVITS. 1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
★ 60kcomfy_sd_krita_plugin. Make AI art between canvas and nodes with Krita.
★ 156magic-animate. [CVPR 2024] Official repository for "MagicAnimate: Temporally Consistent Human Image Animation using Diffusion Model"
★ 11kAwesome-Video-Diffusion-Models. [CSUR] A Survey on Video Diffusion Models
★ 2.3kDiffSynth-Studio. Enjoy the magic of Diffusion models!
★ 13kopeninterpreter. A coding agent for open models like Kimi K3
★ 67kjupyterlab-kernelspy. A Jupyter Lab extension for inspecting messages to/from a kernel
★ 85HuggingFace-Download-Accelerator. 利用HuggingFace的官方下载工具从镜像网站进行高速下载。
★ 1.4kStreamDiffusion. StreamDiffusion: A Pipeline-Level Solution for Real-Time Interactive Generation
★ 11kCogVLM. a state-of-the-art-level open visual language model | 多模态预训练模型
★ 6.7kdlssg-to-fsr3. Adds AMD FSR 3 Frame Generation to games by replacing Nvidia DLSS Frame Generation (nvngx_dlssg).
★ 5kkrita-ai-diffusion. Streamlined interface for generating images with AI in Krita. Inpaint and outpaint with optional text prompt, no tweaking required.
★ 10kgaussian-splatting. Original reference implementation of "3D Gaussian Splatting for Real-Time Radiance Field Rendering"
★ 23klatent-consistency-model. Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference
★ 4.6k4DGaussians. [CVPR 2024] 4D Gaussian Splatting for Real-Time Dynamic Scene Rendering
★ 3.9kGPTs. leaked prompts of GPTs
★ 32kCodeFormer. [NeurIPS 2022] Towards Robust Blind Face Restoration with Codebook Lookup Transformer
★ 18k018_UEGaussianSplatting_Public. C#
★ 45ChatGPT-AutoExpert. 🚀🧠💬 Supercharged Custom Instructions for ChatGPT (non-coding) and ChatGPT Advanced Data Analysis (coding).
★ 6.6kstable-diffusion-prompt-reader. A simple standalone viewer for reading prompts from Stable Diffusion generated image outside the webui.
★ 1.3ksd-webui-animatediff. AnimateDiff for AUTOMATIC1111 Stable Diffusion WebUI
★ 3.4krife-ncnn-vulkan. RIFE, Real-Time Intermediate Flow Estimation for Video Frame Interpolation implemented with ncnn library
★ 1.1ksd-webui-fabric. Python
★ 408ProPainter. [ICCV 2023] ProPainter: Improving Propagation and Transformer for Video Inpainting
★ 6.8kChatDev. ChatDev 2.0: Dev All through LLM-powered Multi-Agent Collaboration
★ 34kLLM-Agent-Paper-List. The paper list of the 86-page SCIS cover paper "The Rise and Potential of Large Language Model Based Agents: A Survey" by Zhiheng Xi et al.
★ 8.2kwonderful-prompts. 🔥中文 prompt 精选🔥,ChatGPT 使用指南,提升 ChatGPT 可玩性和可用性!🚀
★ 6.2kImageReward. [NeurIPS 2023] ImageReward: Learning and Evaluating Human Preferences for Text-to-image Generation
★ 1.7kpromptflow. Build high-quality LLM apps - from prototyping, testing to production deployment and monitoring.
★ 11kAgentBench. A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
★ 3.6kgaussian-splatting-colab. Jupyter Notebook
★ 512airoboros. Customizable implementation of the self-instruct paper.
★ 1.1kTaskMatrix. Python
★ 34kco-tracker. CoTracker is a model for tracking any point (pixel) on a video.
★ 5ksemantic-kernel. Integrate cutting-edge LLM technology quickly and easily into your apps
★ 28kllama.cpp. LLM inference in C/C++
★ 122kseamless_communication. Foundational Models for State-of-the-Art Speech and Text Translation
★ 12kGrounded-Segment-Anything. Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything
★ 18kProfessor-Synapse. Python
★ 3.4kSegment-and-Track-Anything. An open-source project dedicated to tracking and segmenting any objects in videos, either automatically or interactively. The primary algorithms utilized include the Segment Anything Model (SAM) for key-frame segmentation and Associating Objects with Transformers (AOT) for efficient tracking and propagation purposes.
★ 3.1kStableVideo. [ICCV 2023] StableVideo: Text-driven Consistency-aware Diffusion Video Editing
★ 1.4kCoDeF. [CVPR'24 Highlight] Official PyTorch implementation of CoDeF: Content Deformation Fields for Temporally Consistent Video Processing
★ 4.8kFastChat. An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
★ 40ktext-generation-inference. Large Language Model Text Generation Inference
★ 11kLLMAgentPapers. Must-read Papers on LLM Agents.
★ 3.1kMetaHuman-DNA-Calibration. Mathematica
★ 603vllm. A high-throughput and memory-efficient inference and serving engine for LLMs
★ 88kToolBench. [ICLR'24 spotlight] An open platform for training, serving, and evaluating large language model for tool learning.
★ 5.7kgenerative_agents. Generative Agents: Interactive Simulacra of Human Behavior
★ 22kvideo2bvh. Extracts human motion in video and save it as bvh mocap file.
★ 657FlagEmbedding. Retrieval and Retrieval-augmented LLMs
★ 12kgpt-researcher. An autonomous agent that conducts deep research on any data using any LLM providers
★ 29kWebGLM. WebGLM: An Efficient Web-enhanced Question Answering System (KDD 2023)
★ 1.6kGLM-130B. GLM-130B: An Open Bilingual Pre-Trained Model (ICLR 2023)
★ 7.7kdemucs. Code for the paper Hybrid Spectrogram and Waveform Source Separation
★ 10kVITS-fast-fine-tuning. This repo is a pipeline of VITS finetuning for fast speaker adaptation TTS, and many-to-many voice conversion
★ 5kChatPaper2Xmind. 论文XMind笔记生成工具,将论文pdf通过ChatGPT转换为带有图片和公式的简要XMind笔记,提高论文阅读效率。
★ 588AgentVerse. 🤖 AgentVerse 🪐 is designed to facilitate the deployment of multiple LLM-based agents in various applications, which primarily provides two frameworks: task-solving and simulation
★ 5.1kGenshin_Datasets. Genshin Datasets For SVC/SVS/TTS
★ 736StarRail_Datasets. StarRail Datasets For SVC/SVS/TTS
★ 347loguru. Python logging made (stupidly) simple
★ 24kprompt-patterns. Prompt 编写模式:如何将思维框架赋予机器,以设计模式的形式来思考 prompt
★ 3.1kfactool. FacTool: Factuality Detection in Generative AI
★ 936Stable-Diffusion. FLUX, Stable Diffusion, SDXL, SD3, LoRA, Fine Tuning, DreamBooth, Training, Automatic1111, Forge WebUI, SwarmUI, DeepFake, TTS, Animation, Text To Video, Tutorials, Guides, Lectures, Courses, ComfyUI, Google Colab, RunPod, Kaggle, NoteBooks, ControlNet, TTS, Voice Cloning, AI, AI News, ML, ML News, News, Tech, Tech News, Kohya, Midjourney, RunPod
★ 2.8kopen_clip. An open source implementation of CLIP.
★ 14kReAct. [ICLR 2023] ReAct: Synergizing Reasoning and Acting in Language Models
★ 4.1kchain-of-hindsight. Simple next-token-prediction for RLHF
★ 228reflexion. [NeurIPS 2023] Reflexion: Language Agents with Verbal Reinforcement Learning
★ 3.2kimage-search. Simple Image Search powered by Multimodal Foundation Models (OpenAI Clip and Microsoft GLIP)
★ 18ShortGPT. 🚀🎬 ShortGPT - Experimental AI framework for youtube shorts / tiktok channel automation
★ 7.7klang-segment-anything. SAM with text prompt
★ 2.6kCLIP_Surgery. [Pattern Recognition 25] CLIP Surgery for Better Explainability with Enhancement in Open-Vocabulary Tasks
★ 483Personalize-SAM. Personalize Segment Anything Model (SAM) with 1 shot in 10 seconds
★ 1.7kCLIP. CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
★ 34kawesome-segment-anything. Tracking and collecting papers/projects/others related to Segment Anything.
★ 1.7kTrack-Anything. Track-Anything is a flexible and interactive tool for video object tracking and segmentation, based on Segment Anything, XMem, and E2FGVI.
★ 7kMotionBERT. [ICCV 2023] PyTorch Implementation of "MotionBERT: A Unified Perspective on Learning Human Motion Representations"
★ 1.4kstable-diffusion-webui-Layer-Divider. Layer-Divider, an extension for stable-diffusion-webui using the segment-anything model (SAM)
★ 184Painter. Painter & SegGPT Series: Vision Foundation Models from BAAI
★ 2.6kcodeinterpreter-api. 👾 Open source implementation of the ChatGPT Code Interpreter
★ 3.8kMetaGPT. 🌟 The Multi-Agent Framework: First AI Software Company, Towards Natural Language Programming
★ 70kAwesome-LLM. Awesome-LLM: a curated list of Large Language Model
★ 27kannotated_deep_learning_paper_implementations. 🧑🏫 60+ Implementations/tutorials of deep learning papers with side-by-side notes 📝; including transformers (original, xl, switch, feedback, vit, ...), optimizers (adam, adabelief, sophia, ...), gans(cyclegan, stylegan2, ...), 🎮 reinforcement learning (ppo, dqn), capsnet, distillation, ... 🧠
★ 67kawesome-open-gpt. Collection of Open Source Projects Related to GPT,GPT相关开源项目合集🚀、精选🔥🔥
★ 6kLocalAI. LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.
★ 48kFlowise. Build AI Agents, Visually
★ 55kMotionGPT. [NeurIPS 2023] MotionGPT: Human Motion as a Foreign Language, a unified motion-language generation model using LLMs
★ 1.9ksdnext. SD.Next: All-in-one WebUI for AI generative image and video creation, captioning and processing
★ 7.2kComfyUI-Manager. ComfyUI-Manager is an extension designed to enhance the usability of ComfyUI. It offers management functions to install, remove, disable, and enable various custom nodes of ComfyUI. Furthermore, this extension provides a hub feature and convenience functions to access a wide range of information within ComfyUI.
★ 16kMegatron-DeepSpeed. Ongoing research training transformer language models at scale, including: BERT & GPT-2
★ 2.3kpeft. 🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
★ 21kChatGLM2-6B. ChatGLM2-6B: An Open Bilingual Chat LLM | 开源双语对话语言模型
★ 16kOpenLLM. Run any open-source LLMs, such as DeepSeek and Llama, as OpenAI compatible API endpoint in the cloud.
★ 12kgpt-engineer. CLI platform to experiment with codegen. Precursor to: https://lovable.dev
★ 55kunreal-pcg-examples. Accompanying project to the Unreal PCG Tutorial series: https://www.youtube.com/watch?v=BqPhdQOweqU&list=PLA03OHAaHgYpo0enf8p-2oEpja3grLOKZ
★ 132openai-cookbook. Examples and guides for using the OpenAI API
★ 75kVoyager. An Open-Ended Embodied Agent with Large Language Models
★ 7.1kLLM-ToolMaker. Jupyter Notebook
★ 1.1kStyleAvatar3D. Official repo for StyleAvatar3D
★ 505guidance. A guidance language for controlling large language models.
★ 22kRetrieval-based-Voice-Conversion-WebUI. Easily train a good VC model with voice data <= 10 mins!
★ 37kthreestudio. A unified framework for 3D content generation.
★ 7kprolificdreamer. ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score Distillation (NeurIPS 2023 Spotlight)
★ 1.6kWizardVicunaLM. LLM that combines the principles of wizardLM and vicunaLM
★ 716tree-of-thoughts. Plug in and Play Implementation of Tree of Thoughts: Deliberate Problem Solving with Large Language Models that Elevates Model Reasoning by atleast 70%
★ 4.6ktree-of-thought-llm. [NeurIPS 2023] Tree of Thoughts: Deliberate Problem Solving with Large Language Models
★ 6kglTFRuntime. Unreal Engine Plugin for loading glTF files at runtime
★ 549safe-rlhf. Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback
★ 1.6kDragGAN. Official Code for DragGAN (SIGGRAPH 2023)
★ 36kJARVIS. JARVIS, a system to connect LLMs with ML community. Paper: https://arxiv.org/pdf/2303.17580.pdf
★ 25kadetailer. Auto detecting, masking and inpainting with detection model.
★ 4.8kAutoGPT. AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters.
★ 186kAuto-Photoshop-StableDiffusion-Plugin. A user-friendly plug-in that makes it easy to generate stable diffusion images inside Photoshop using either Automatic or ComfyUI as a backend.
★ 7.3kVideoRenderer. RTX HDR modded into MPC-VideoRenderer.
★ 1.5kSD-CN-Animation. This script allows to automate video stylization task using StableDiffusion and ControlNet.
★ 819sd-webui-aspect-ratio-helper. Simple extension to easily maintain aspect ratio while changing dimensions. Install via the extensions tab on the AUTOMATIC1111 webui.
★ 438Exporters. Exporters for Babylon.js and gltf file formats
★ 701was-node-suite-comfyui. An extensive node suite for ComfyUI with over 210 new nodes
★ 1.8k