This is your work, valued
disambiguation. Python
★ 92AutoPrescribe. Learning to Prescribe
★ 18Coursera-Machine-Learning. source from exercises in Coursera.
★ 8aminer-spider. Python
★ 8aminer-core-1. The NEXT Generation of ArnetMiner.org
★ 6Vogue. Python
★ 2windows-phone-7-workshop. Application connection to Twitter and Flickr
★ 2octopress. The source of my octopress blog
★ 1experimental-code-archive. C++
★ 1Archi. Computer Architecture
★ 1GameMaker-HTML5-Player. A tool to convert and play GameMaker games in the browser
★ 1GWAP. JavaScript
★ 1MIPS-CPU-Simulator. This is a mips simulator I wrote once to help my understanding of pipelines, branch prediction, assembly language, and more.
★ 1NeMo. Neural Modules: a toolkit for conversational AI
★ 1AWS-SDK-for-Windows-Phone-7. This is source code for an asynchronous SDK for interacting with Amazon S3 webservices for Windows Phone 7.x (WP7)
★ 1html5-game-book. Code Samples and Demos for "Learning HTML5 Game Programming"
★ 1knowledge-evolution. TeX
★ 1AutoVocal. An Experimental Automatical Singing Voice Syntheziser
★ 1cron. Small R-Type shooter based on craftyjs.com
★ 1CERMINE. Content ExtRactor and MINEr
★ 1PCV. Open source Python module for computer vision
★ 1AlsoView-Net. Large scale network analysis(Linkedin also-view network).
★ 1aminer-reader. PDF Reader in JavaScript
★ 1Kinect-HCI.
★ 1Coding4fun.
★ 1valuecell. ValueCell is a community-driven, multi-agent platform for financial applications.
★ 11kkimi-cli. Kimi Code CLI is your next CLI agent.
★ 11knautilus_trader. Production-grade Rust-native trading engine with deterministic event-driven architecture
★ 25kagents-for-openbb. Custom agents for OpenBB Workspace
★ 361OpenBB. Open Data Platform for analysts, quants and AI agents.
★ 71kccxt. A unified trading API with more than 100 crypto exchanges and prediction markets in JavaScript / TypeScript / Python / C# / PHP / Go / Java
★ 43kBiomni. Biomni: a general-purpose biomedical AI agent
★ 3.6khummingbot. Open source software that helps you create and deploy high-frequency crypto trading bots
★ 19kthetagang. ThetaGang is an IBKR bot for collecting money
★ 2.7kibkr-ai-agent. Interactive Brokers MCP-based AI Agent
★ 4ibkr_agent. Python
★ 2ibkr-mcp-server. MCP Server for IBKR Client
★ 64a-list-of-claude-code-agents. A list of Claude Code Sub-Agents submitted by the community.
★ 1.3kKimi-K2. Kimi K2 is the large language model series developed by Moonshot AI team
★ 11kFeishu-MCP. Feishu / Lark 飞书文档与任务管理工具,支持 MCP 服务器和 CLI + Skill 两种使用方式,可无缝集成 Cursor、Claude Code、Cline 等 AI 编码工具
★ 715opcode. A powerful GUI app and Toolkit for Claude Code - Create custom agents, manage interactive Claude Code sessions, run secure background agents, and more.
★ 22kBrowserOS. 🌐 The open-source Agentic browser; alternative to ChatGPT Atlas, Perplexity Comet, Dia.
★ 13kquantstats. Portfolio analytics for quants, written in Python
★ 7.5ksuperagent. Superagent protects your AI applications against prompt injections, data leaks, and harmful outputs. Embed safety directly into your app and prove compliance to your customers.
★ 6.7kmoonpalace. MoonPalace(月宫)是由 Moonshot AI 月之暗面提供的 API 调试工具。
★ 255stagehand. The SDK For Browser Agents
★ 24kTradingAgents. TradingAgents: Multi-Agents LLM Financial Trading Framework
★ 95kquantitative_analysis. 量化分析
★ 376mcp-yahoo-finance. A Model Context Protocol (MCP) server for Yahoo Finance.
★ 26autogen. A programming framework for agentic AI
★ 60kxhs-mcp-server. 小红书MCP服务器 | 秒级工具调用执行,笔记用户搜索、通知消息监控等;打造2026最简单、最快速、最精准、最好用的小红书MCPServer!
★ 178deepwiki-open. Open Source DeepWiki: AI-Powered Wiki Generator for GitHub/Gitlab/Bitbucket Repositories. Join the discord: https://discord.gg/gMwThUMeme
★ 17kdeer-flow. An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, subagents and message gateway, it handles different levels of tasks that could take minutes to hours.
★ 78kFramePack. Lets make video diffusion practical!
★ 17kai-hedge-fund. An AI Hedge Fund Team
★ 62kOpenHands. 🙌 OpenHands: AI-Driven Development
★ 83kRD-Agent. Research and development (R&D) is crucial for the enhancement of industrial productivity, especially in the AI era, where the core aspects of R&D are mainly focused on data and models. We are committed to automating these high-value generic R&D processes through R&D-Agent, which lets AI drive data-driven AI. 🔗https://aka.ms/RD-Agent-Tech-Report
★ 14kqlib. Qlib is an AI-oriented Quant investment platform that aims to use AI tech to empower Quant Research, from exploring ideas to implementing productions. Qlib supports diverse ML modeling paradigms, including supervised learning, market dynamics modeling, and RL, and is now equipped with https://github.com/microsoft/RD-Agent to automate R&D process.
★ 47kKunQuant. A compiler, optimizer and executor for financial expressions and factors
★ 297CHRONOS. Repo for NAACL 2025 Paper "Unfolding the Headline: Iterative Self-Questioning for News Retrieval and Timeline Summarization"
★ 297MoBA. MoBA: Mixture of Block Attention for Long-Context LLMs
★ 2.2kFunClip. FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.
★ 6.1kai2srt. 利用 GeminiAI 一键为长视频创建解说短视频,并支持音视频转录字幕
★ 69browser-use. 🌐 Make websites accessible for AI agents. Automate tasks online with ease.
★ 107koctogen. Octogen is an Open-Source Code Interpreter Agent Framework
★ 282electron-browser-shell. A minimal, tabbed web browser with support for Chrome extensions—built on Electron.
★ 722music. a music player forked from YesPlayMusic。高颜值的第三方网易云播放器,支持 Windows / macOS / Linux :electron/Docker:
★ 693CogVLM. a state-of-the-art-level open visual language model | 多模态预训练模型
★ 6.7kbark-voice-cloning-HuBERT-quantizer. The code for the bark-voicecloning model. Training and inference.
★ 710fastsdcpu. Fast stable diffusion on CPU and AI PC
★ 2.1kworktool. 一款安全稳定的Android无障碍服务工具,支持控制企微/微信来运行的无人值守群管理企业微信机器人
★ 2.7kopeninterpreter. A coding agent for open models like Kimi K3
★ 67kGrounding_LLMs_with_online_RL. We perform functional grounding of LLMs' knowledge in BabyAI-Text
★ 276FastGPT. FastGPT is a knowledge-based platform built on the LLMs, offers a comprehensive suite of out-of-the-box capabilities such as data processing, RAG retrieval, and visual AI workflow orchestration, letting you easily develop and deploy complex question-answering systems without the need for extensive setup or configuration.
★ 29kToolBench. [ICLR'24 spotlight] An open platform for training, serving, and evaluating large language model for tool learning.
★ 5.7kSuperAGI. <⚡️> SuperAGI - A dev-first open source autonomous AI agent framework. Enabling developers to build, manage & run useful autonomous agents quickly and reliably.
★ 18kcodeinterpreter-api. 👾 Open source implementation of the ChatGPT Code Interpreter
★ 3.8kgpt-code-ui. An open source implementation of OpenAI's ChatGPT Code interpreter
★ 3.5kDB-GPT. open-source agentic AI data assistant for the next generation of AI + Data products.
★ 20kso-vits-svc. SoftVC VITS Singing Voice Conversion
★ 28kbark. 🔊 Text-Prompted Generative Audio Model
★ 39kdocprompting. Data and code for "DocPrompting: Generating Code by Retrieving the Docs" @ICLR 2023
★ 253JARVIS. JARVIS, a system to connect LLMs with ML community. Paper: https://arxiv.org/pdf/2303.17580.pdf
★ 25kgerev. 🧠 AI-powered enterprise search engine 🔎
★ 2.8kfastertransformer_backend. Python
★ 413plantuml. Generate diagrams from textual description
★ 13kapitable. 🚀🎉📚 APITable, an API-oriented low-code platform for building collaborative apps and better than all other Airtable open-source alternatives.
★ 15kchat-gpt-ppt. Use ChatGPT (or other backends) to generate PPT automatically, all in one single file.
★ 919Co-Speech-Motion-Generation. Freeform Body Motion Generation from Speech
★ 210ml-neuman. Official repository of NeuMan: Neural Human Radiance Field from a Single Video (ECCV 2022)
★ 1.3kRPA. Ui.Vision Open-Source RPA Software with Computer Vision, OCR, Anthropic Computer Use/LLM. Selenium IDE import/export.
★ 2klowdefy. Build apps that AI can generate, humans can review, and teams can maintain. Config that works between code and natural language.
★ 3kVideoPose3D. Efficient 3D human pose estimation in video using 2D keypoint trajectories
★ 4.1kPaddleBoBo. 基于飞桨开发的虚拟主播
★ 1.1kdocusaurus. Easy to maintain open source documentation websites.
★ 66kvuepress-theme-antdocs. 🔥🎨 An Ant Design style theme for VuePress. (QQ Group: 867711329) [NOTE: The AntDocs-next is WIP.]
★ 216FastDeploy. High-performance Inference and Deployment Toolkit for LLMs and VLMs based on PaddlePaddle
★ 3.7kfastDeploy. Deploy DL/ ML inference pipelines with minimal extra code.
★ 105aframe. :a: Web framework for building virtual reality experiences.
★ 18kNVFlare. NVIDIA Federated Learning Application Runtime Environment
★ 952mellotron. Mellotron: a multispeaker voice synthesis model based on Tacotron 2 GST that can make a voice emote and sing without emotive or singing training data
★ 869nnsvs. Neural network-based singing voice synthesis library for research
★ 747TikTok-Compilation-Video-Generator. A system of bots that collects clips automatically via custom made filters, lets you easily browse these clips, and puts them together into a compilation video ready to be uploaded straight to any social media platform. Full VPS support is provided, along with an accounts system so multiple users can use the bot at once. This bot is split up into three separate programs. The server. The client. The video generator. These programs perform different functions that when combined creates a very powerful system for auto generating compilation videos.
★ 947screenity. The free and privacy-friendly screen recorder with no limits 🎥
★ 18kDiffSinger. DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code
★ 4.8kWGANSing. Multi-voice singing voice synthesis
★ 237deckdeckgo. The web open source editor for presentations
★ 1.7kmoving-letters. Text animated with anime.js
★ 571taichi. Productive, portable, and performant GPU programming in Python.
★ 28kluci. LuCI - OpenWrt Configuration Interface
★ 7.8konline_speaker_change_detector. Online streaming speaker change detection model in Pytorch
★ 44n8n. Fair-code workflow automation platform with native AI capabilities. Combine visual building with custom code, self-host or cloud, 400+ integrations.
★ 199ktalking-head-anime-demo. Demo for the "Talking Head Anime from a Single Image."
★ 2kLipSync. LipSync for Unity3D 根据语音生成口型动画 支持fmod
★ 493talking-head-anime-2-demo. Demo programs for the Talking Head Anime from a Single Image 2: More Expressive project.
★ 1.2kWav2Lip. This repository contains the codes of "A Lip Sync Expert Is All You Need for Speech to Lip Generation In the Wild", published at ACM Multimedia 2020. For HD commercial model, please try out Sync Labs
★ 13kopenpose. OpenPose: Real-time multi-person keypoint detection library for body, face, hands, and foot estimation
★ 34kOpenVtuber. 虚拟爱抖露(アイドル)共享计划, 是基于单目RGB摄像头的人眼与人脸特征点检测算法, 在实时3D面部捕捉以及模型驱动领域的应用.
★ 967vignette. The open source VTuber software. ❤
★ 525VTuber_Unity. Use Unity 3D character and Python deep learning algorithms to stream as a VTuber!
★ 810OpenVTuberProject. Open Vtuber project containing all sub projects
★ 245EEND_PyTorch. A PyTorch implementation of End-to-End Neural Diarization
★ 110cyclone. Powerful workflow engine and end-to-end pipeline solutions implemented with native Kubernetes resources. https://cyclone.dev
★ 1.1kormb. Docker for Your ML/DL Models Based on OCI Artifacts
★ 473TTS. 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
★ 46kuhubctl. uhubctl - USB hub per-port power control
★ 2.8kAutoHotkey. AutoHotkey - macro-creation and automation-oriented scripting utility for Windows.
★ 13kpywinauto. Windows GUI Automation with Python (based on text properties)
★ 6.1kfind3-cli-scanner. The command-line scanner that supports Bluetooth and WiFi
★ 151indoor-location. balabala
★ 45find3. High-precision indoor positioning framework, version 3.
★ 4.8kIndoorPos. 这是一个采用蓝牙4.0--iBeacon技术的室内定位服务端程序。
★ 640docker-wxwork. DoWork is a Dockerized WeChat Work (盒装企业微信) PC Windows Client for Linux
★ 134docker-wechat. DoChat is a Dockerized WeChat (盒装微信) PC Windows Client for Linux
★ 2.3kPython-UIAutomation-for-Windows. 🐍Python 3 wrapper of Microsoft UIAutomation. Support UIAutomation for MFC, WindowsForm, WPF, Modern UI(Metro UI), Qt, IE, Firefox, Chrome ...
★ 3.5kTagUI. Free RPA tool by AI Singapore
★ 6.3kspeexdsp. Speex audio processing library - THIS IS A MIRROR, DEVELOPMENT HAPPENS AT https://gitlab.xiph.org/xiph/speexdsp
★ 718ec. Echo Canceller, part of Voice Engine project
★ 291EarTrumpet. EarTrumpet - Volume Control for Windows
★ 11kBackgroundMusic. Background Music, a macOS audio utility: automatically pause your music, set individual apps' volumes and record system audio.
★ 19kgo-webrtcvad. cgo interface to WebRTC Voice Activity Dectection
★ 70gotapestry. Go on Tapestry
★ 3ice. 🚀 ice.js: The Progressive App Framework Based On React(基于 React 的渐进式应用框架)
★ 19krpi-audio-receiver. Raspberry Pi Audio Receiver with Bluetooth A2DP, AirPlay 2, and Spotify Connect
★ 1.6kLibreASR. :speech_balloon: An On-Premises, Streaming Speech Recognition System
★ 679CPWechatXposed. 使用Xposed Hook微信等APP
★ 628bluez-alsa. Bluetooth Audio ALSA Backend
★ 975pybluez. Bluetooth Python extension module
★ 2.4kBluetoothGattMitm. Revealing unknown GATT/BLE protocol by mediation in communication between device and client
★ 21Bluetooth_LE_MITM. Man-in-the-Middle Relay program between a Bluetooth Low-Energy (BTLE) Peripheral and Central
★ 14phasen. A unofficial Pytorch implementation of Microsoft's PHASEN
★ 235webcamoid. Webcamoid is a full featured and multiplatform camera suite.
★ 2.5kvoice-engine. building blocks to create voice interface applications
★ 204odas. ODAS: Open embeddeD Audition System
★ 1kelectron-transparency-demo. Demo for transparent windows with click-through in Electron
★ 8SpeechSplit. Unsupervised Speech Decomposition Via Triple Information Bottleneck
★ 697Realtime-audio2audio-alignment. temp_files
★ 7QASystemOnMedicalKG. A tutorial and implement of disease centered Medical knowledge graph and qa system based on it。知识图谱构建,自动问答,基于kg的自动问答。以疾病为中心的一定规模医药领域知识图谱,并以该知识图谱完成自动问答与分析服务。
★ 7.3kVGGVox. VGGVox models for Speaker Identification and Verification trained on the VoxCeleb (1 & 2) datasets
★ 402WebRTCVAD_Wrapper. A simple Python wrapper to simplify working with WebRTC VAD and its rougher analogue based on RMS and ZCR (useful for processing audio recordings before using them with neural networks).
★ 9openrpa. Free Open Source Enterprise Grade RPA
★ 3kTNN. TNN: developed by Tencent Youtu Lab and Guangying Lab, a uniform deep learning inference framework for mobile、desktop and server. TNN is distinguished by several outstanding features, including its cross-platform capability, high performance, model compression and code pruning. Based on ncnn and Rapidnet, TNN further strengthens the support and performance optimization for mobile devices, and also draws on the advantages of good extensibility and high performance from existed open source efforts. TNN has been deployed in multiple Apps from Tencent, such as Mobile QQ, Weishi, Pitu, etc. Contributions are welcome to work in collaborative with us and make TNN a better framework.
★ 4.6kzoomChatBot. A chatbot for zoom, in python.
★ 12Zoogle. Zoom Google Assistant Bot
★ 10MediaStreamRecorder. Cross browser audio/video/screen recording. It supports Chrome, Firefox, Opera and Microsoft Edge. It even works on Android browsers. It follows latest MediaRecorder API standards and provides similar APIs.
★ 2.7kGuyu. Chinese GPT2: pre-training and fine-tuning framework for text generation
★ 187SongNet. Code for ACL 2020 paper "Rigid Formats Controlled Text Generation":https://www.aclweb.org/anthology/2020.acl-main.68/
★ 235Resemblyzer. A python package to analyze and compare voices with deep learning
★ 3.3kopenvino. OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
★ 11kwechat.hook.sdk. 商用版微信开发SDK,非微信ipad/mac/android协议,已封装好全部API接口,可实现微信99%功能; 可开发微信群控、云控、微信机器人、个人号SCRM客服系统等, 微信二次开发SDK,微信个人号开发API接口协议。(提供源代码)
★ 433DALI. A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference applications.
★ 5.7kcheetah. On-device streaming speech-to-text engine powered by deep learning
★ 669CnC_Remastered_Collection. Command & Conquer: Remastered Collection
★ 21kCTC_pytorch. CTC end -to-end ASR for timit and 863 corpus.
★ 219espresso. Espresso: A Fast End-to-End Neural Speech Recognition Toolkit
★ 939StreamingTransformer. Python
★ 277Tacotron-2. DeepMind's Tacotron-2 Tensorflow implementation
★ 2.3kdeep-voice-conversion. Deep neural networks for voice conversion (voice style transfer) in Tensorflow
★ 3.9kasteroid. The PyTorch-based audio source separation toolkit for researchers
★ 2.6krnnt-speech-recognition. End-to-end speech recognition using RNN Transducers in Tensorflow 2.0
★ 250photo2cartoon. 人像卡通化探索项目 (photo-to-cartoon translation project)
★ 4kespnet. End-to-End Speech Processing Toolkit
★ 9.9kprefix-beam-search. Code for prefix beam search tutorial by @labodk
★ 187aeneas. aeneas is a Python/C library and a set of tools to automagically synchronize audio and text (aka forced alignment)
★ 2.9klip-reading-deeplearning. :unlock: Lip Reading - Cross Audio-Visual Recognition using 3D Architectures
★ 1.9kopen_stt_e2e. PyTorch end-to-end speech recognition
★ 50python-wechaty. Python Wechaty is a Conversational RPA SDK for Chatbot Makers written in Python
★ 1.8kwarp-rnnt. CUDA-Warp RNN-Transducer
★ 215warp-transducer. A fast parallel implementation of RNN Transducer.
★ 313server. The Triton Inference Server provides an optimized cloud and edge inferencing solution.
★ 11klabel-studio. Label Studio is a multi-type data labeling and annotation tool with standardized output format
★ 28krasa-webchat. A feature-rich chat widget for Rasa and Botfront
★ 1kMegEngine. MegEngine 是一个快速、可拓展、易于使用且支持自动求导的深度学习框架
★ 4.8knemo_examples. Python
★ 6dl_inference. 通用深度学习推理工具,可在生产环境中快速上线由TensorFlow、PyTorch、Caffe框架训练出的深度学习模型。
★ 417RE-VERB. speaker diarization system using an LSTM
★ 50TensorRT. PyTorch/TorchScript/FX compiler for NVIDIA GPUs using TensorRT
★ 3kffsubsync. Automagically synchronize subtitles with video.
★ 7.8kvadnet. Real-time Voice Activity Detection in Noisy Eniviroments using Deep Neural Networks
★ 465WechatEnhancement. 仅供学习交流,禁止用于其他用途,请及时删除,禁止任何公司或个人发布与传播,不接受任何捐赠
★ 1.2kjittor. Jittor is a high-performance deep learning framework based on JIT compiling and meta-operators.
★ 3.2kPyTorch_ONNX_TensorRT. A tutorial about how to build a TensorRT Engine from a PyTorch Model with the help of ONNX
★ 248Speech. A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
★ 18kDSAlign. DeepSpeech based forced alignment tool
★ 239DSAlign. DeepSpeech based forced alignment tool
★ 2gorgonia. Gorgonia is a library that helps facilitate machine learning in Go.
★ 5.9kdeepspeech2-online-decoder. Online (real-time) decoder to be used with DeepSpeech2 model
★ 25Speech-Transformer. A PyTorch implementation of Speech Transformer, an End-to-End ASR with Transformer network on Mandarin Chinese.
★ 810audiomentations. A Python library for audio data augmentation. Useful for making audio ML models work well in the real world, not just in the lab.
★ 2.3kwukong-robot. 🤖 wukong-robot 是一个简单、灵活、优雅的中文语音对话机器人/智能音箱项目,支持ChatGPT多轮对话能力,还可能是首个支持脑机交互的开源智能音箱项目。
★ 7.1k_rasa_chatbot. A Chinese task oriented chatbot in IVR(Interactive Voice Response) domain, implement by rasa. This is a demo with toy dataset, more data should be added for performance.
★ 495CTCWordBeamSearch. Connectionist Temporal Classification (CTC) decoder with dictionary and language model.
★ 578CTCDecoder. Connectionist Temporal Classification (CTC) decoding algorithms: best path, beam search, lexicon search, prefix search, and token passing. Implemented in Python.
★ 837mtl-text-recognition. multi-task learning for text recognition with joint CTC-attention
★ 117ASRFrame. An Automatic Speech Recognition Frame ,一个中文语音识别的完整框架, 提供了多个模型
★ 252mypy. Optional static typing for Python
★ 21kcode-server. VS Code in the browser
★ 79kev3play. 买了个lego ev3玩
★ 3transformers. 🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
★ 163kdeep-dregs. A streaming Speech to Text server using DeepSpeech
★ 16Gather-Deployment. Gathers Python deployment, infrastructure and practices.
★ 348service-streamer. Boosting your Web Services of Deep Learning Applications.
★ 1.2kargo-workflows. Workflow Engine for Kubernetes
★ 17kspleeter. Deezer source separation library including pretrained models.
★ 28kpie. 百度云流式语音识别客户端 SDK
★ 80