This is your work, valued
ArtFormer. [CVPR 2025] ArtFormer: Controllable Generation of Diverse 3D Articulated Objects
★ 43Homework-of-Algorithm-Analysis-in-2404. Homework of Algorithm Analysis in 2404
★ 4OOP-Java-homework-PVZ. homework of `OOP(java)`: PVZ in JavaFx.
★ 3Assignment-of-Compile-Principle. Assignment of Compile Principle
★ 2Generative-Articulated-Object-s-. 满纸荒唐言
★ 2memo-with-backend. 帮新闻系同学小组作业写的后端
★ 2TransArticulate. [CVPR 2025] ArtFormer: Controllable Generation of Diverse 3D Articulated Objects, contains all commits version.
★ 1quant-for-beginners. Jupyter Notebook
★ 401rclone. "rsync for cloud storage" - Google Drive, S3, Dropbox, Backblaze B2, One Drive, Swift, Hubic, Wasabi, Google Cloud Storage, Azure Blob, Azure Files, Yandex Files
★ 59kgen2act. Python
★ 13x-ui. Xray panel supporting multi-protocol multi-user expire day & traffic & IP limit (Vmess, Vless, Trojan, ShadowSocks, Wireguard, Hysteria, Tunnel, Mixed, HTTP, Tun, MTProto)
★ 44kcc-switch-cli. ⭐️ A cross-platform CLI All-in-One assistant tool for Claude Code, Codex & Gemini CLI.
★ 4.5kRoboDojo. RoboDojo Official Repo
★ 320cc-switch. A cross-platform desktop All-in-One assistant for Claude Code, Codex, OpenCode, OpenClaw, Grok Build & Hermes Agent. Only official website: ccswitch.io
★ 123kskills. Public repository for Agent Skills
★ 165kLandPPT. 一个基于LLM的演示文稿生成平台,能够自动将文档内容转换为专业的PPT演示文稿。平台支持多种AI模型,提供丰富的模板和样式选择,让用户能够创建高质量的演示文稿。
★ 3.5kSSR-V2ray-Trojan. 2026机场推荐与机场评测
★ 17kbiliup. 自动直播录制、投稿、twitch、ytb频道搬运工具。命令行投稿(B站)和视频下载工具,提供多种登录方式,支持多p。
★ 5.3kbilibili-cli. A CLI for Bilibili — browse videos, users, search, and feeds from the terminal
★ 951galbot_one_golf_description. Python
★ 22flux2. Official inference repo for FLUX.2 models
★ 2.6kflux. Official inference repo for FLUX.1 models
★ 26kImageWAM. ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing?
★ 143academic-figure-generator. AI 驱动的学术论文配图生成平台。上传论文 → AI 分析内容生成 Prompt → 一键生成高质量科研配图,还有配套的skill可在主流agent中使用
★ 1.9kcosmos. NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.
★ 11kpdfrip. A multi-threaded PDF password cracking utility equipped with commonly encountered password format builders and dictionary attacks.
★ 1.4kmihomo. A simple Python Pydantic model for Honkai: Star Rail parsed data from the Mihomo API.
★ 33kcli. The official Lark/Feishu CLI tool, maintained by the larksuite team — built for humans and AI Agents. Covers core business domains including Messenger, Docs, Base, Sheets, Calendar, Mail, Tasks, Meetings, and more, with 200+ commands and 20+ AI Agent Skills.
★ 16ktaehv. Tiny AutoEncoder for Hunyuan Video (and other video models)
★ 450RLinf. RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI
★ 4.3kBEHAVIOR-1K. BEHAVIOR-1K: a platform for accelerating Embodied AI research. Join our Discord for support: https://discord.gg/bccR5vGFEx
★ 1.6kInternDataEngine. InternDataEngine: Pioneering High-Fidelity Synthetic Data Generator for Robotic Manipulation
★ 120DeepSeek-OCR. Contexts Optical Compression
★ 24kclashfree. clash节点、免费clash节点、免费节点、免费梯子、clash科学上网、clash翻墙、clash订阅链接、clash for Windows、clash教程、免费公益节点、最新clash免费节点订阅地址、clash免费节点每日更新
★ 16kRoboCOIN. RoboCoin + LeRobot integration
★ 207ProcVLM. ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation https://procvlm.github.io/
★ 18FastVideo. A unified inference and post-training framework for accelerated video generation.
★ 3.9kopen-webui. User-friendly AI Interface (Supports Ollama, OpenAI API, ...)
★ 147kAudioVisual. 仓库已停止维护
★ 3.2kcodex-loop-plugin. Claude Code-style recurring loop scheduling for OpenAI Codex.
★ 9codex. Lightweight coding agent that runs in your terminal
★ 103kRoboLab. Python
★ 401voice-input-src.
★ 2.4ksub2api. Sub2API 一站式开源中转服务,让 Claude、Openai 、Gemini、Grok订阅统一接入,支持拼车共享,更高效分摊成本,原生工具无缝使用。
★ 35kAlphaBrain. The Comprehensive Toolkit for Embodied AI Models
★ 231DeepSpeed. DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
★ 43kpolaris. A real2sim evaluation framework for generalist policies
★ 222best-igl-wiki. HTML
★ 24codex-cli-best-practice. from vibe coding to agentic engineering - practice makes codex perfect
★ 948andrej-karpathy-skills. A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.
★ 198kclaude-code-best-practice. from vibe coding to agentic engineering - practice makes claude perfect
★ 64kApollo-11. Original Apollo 11 Guidance Computer (AGC) source code for the command and lunar modules.
★ 72ktailscale. The easiest, most secure way to use WireGuard and 2FA.
★ 35kopencv. Open Source Computer Vision Library
★ 90kmink. Python inverse kinematics based on MuJoCo.
★ 1.5kany4lerobot. 🎁 A collection of utilities for LeRobot.
★ 1.1ktarsier. Tarsier -- a family of large-scale video-language models, which is designed to generate high-quality video descriptions , together with good capability of general video understanding.
★ 548Self-Forcing. Official codebase for "Self Forcing: Bridging Training and Inference in Autoregressive Video Diffusion" (NeurIPS 2025 Spotlight)
★ 3.5kdify. Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self-hosted, so teams move from prototype to production without rebuilding the stack.
★ 151kjobs. A research tool for visually exploring Bureau of Labor Statistics Occupational Outlook Handbook data. This is not a report, a paper, or a serious economic publication — it is a development tool for exploring BLS data visually.
★ 1.9kpixi. Powerful system-level package manager for Linux, macOS and Windows written in Rust – building on top of the Conda ecosystem.
★ 7.5kCausal-Forcing. [ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation" & Causal Forcing++
★ 891CLIP. CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
★ 34kleonidk.github.io. a fork of https://jonbarron.info/ for use in jekyll builds with markdown page updates
★ 361lingbot-world. Advancing Open-source World Models
★ 4.3kcursor-deepseek. A high-performance HTTP/2-enabled proxy server designed specifically to enable Cursor IDE's Composer to use DeepSeek's and OpenRouter's language models. This proxy translates OpenAI-compatible API requests to DeepSeek/OpenRouter API format, allowing Cursor's Composer and other OpenAI API-compatible tools to seamlessly work with these models.
★ 640new-api. A unified AI model hub for aggregation & distribution. It supports cross-converting various LLMs into OpenAI-compatible, Claude-compatible, or Gemini-compatible formats. A centralized gateway for personal and enterprise model management. 🍥
★ 44kdas-datakit. Python
★ 54FinRL. FinRL®: Financial Reinforcement Learning. 🔥
★ 16kvjepa2. PyTorch code and models for VJEPA2 self-supervised learning from video.
★ 4.4kjepa. PyTorch code and models for V-JEPA self-supervised learning from video.
★ 4.1kItChat. A complete and graceful API for Wechat. 微信个人号接口、微信机器人及命令行微信,三十行即可自定义个人号机器人。
★ 26kwechaty. Conversational RPA SDK for Chatbot Makers. Join our Discord: https://discord.gg/7q8NBZbQzt
★ 23kLDA-1B. [RSS 2026] LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion
★ 289Action100M. A Large-scale Video Action Dataset
★ 483rcm. rCM & Causal-rCM: Leading and Unified Algorithms/Infrastructures for Bidirectional/Autoregressive Video Diffusion Distillation at Scale
★ 776lingbot-va. [RSS 2026] Causal video-action world model for generalist robot control
★ 1.7kredis. For developers, who are building real-time data-driven applications, Redis is the preferred, fastest, and most feature-rich cache, data structure server, and document and vector query engine.
★ 76kGaVS. [SIGGRAPH 2025] Official Repo for Paper - GaVS: 3D-Grounded Video Stabilization via Temporally-Consistent Local Reconstruction and Rendering
★ 294dreamzero. Code to pretrain, fine-tune, and evaluate DreamZero and run sim & real-world evals
★ 2.5kdacite. Simple creation of data classes from dictionaries.
★ 2kfourier-feature-networks. Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional Domains
★ 1.4kopenclaw. Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
★ 385kPyAV. Pythonic bindings for FFmpeg's libraries.
★ 3.3kfrankx. High-Level Motion Library for Collaborative Robots
★ 338lingbot-vla. A Pragmatic VLA Foundation Model
★ 1.7kdecord. An efficient video loader for deep learning with smart shuffling that's super easy to digest
★ 2.5klerobot. 🤗 LeRobot: Making AI for Robotics more accessible with end-to-end learning
★ 26kLTX-2. Official Python inference and LoRA trainer package for the LTX-2 audio–video generative model.
★ 8.5kdocker-install. Docker installation script
★ 3.2kDiffSynth-Studio. Enjoy the magic of Diffusion models!
★ 13kVideo-As-Prompt. [ICLR 2026] Official repo for paper "Video-As-Prompt: Unified Semantic Control for Video Generation"
★ 443LightX2V. Lightweight Image Video Action Generation Inference Framework
★ 2.5kmobile_manipulation_papers. Papers in Mobile Manipulation (Personal Collection)
★ 57sam-3d-objects. SAM 3D Objects
★ 7.2kAwesome-Video-Robotic-Papers. This repository compiles a list of papers related to the application of video technology in the field of robotics! Star⭐ the repo and follow me if you like what you see🤩.
★ 193openai-cookbook. Examples and guides for using the OpenAI API
★ 75kpolydown. Batch downloader for polyhaven.com. Download 'HDRIs, Textures and Models' in any sizes with preview images from Poly Haven.
★ 104cursor-memory-bank. A modular, documentation-driven framework using Cursor custom modes (VAN, PLAN, CREATIVE, IMPLEMENT) to provide persistent memory and guide AI through a structured development workflow with visual process maps.
★ 3.1kgenie_sim. Simulation Platform from AgiBot
★ 1.3khexo. A fast, simple & powerful blog framework, powered by Node.js.
★ 42kcurobo. CUDA Accelerated Robot Library
★ 1.7kmega-sam. Code for the project "MegaSaM: Accurate, Fast and Robust Structure and Motion from Casual Dynamic Videos"
★ 1.3kmolmo. Code for the Molmo Vision-Language Model
★ 923molmo2. Code for the Molmo2 Vision-Language Model
★ 697starVLA. StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing
★ 3.3krerun. Visualize, query, and stream to train on multimodal robotics data.
★ 11kdiffusion_policy. [RSS 2023] Diffusion Policy Visuomotor Policy Learning via Action Diffusion
★ 4.4kdex-retargeting. Various retargeting optimizers to translate human hand motion to robot hand motion.
★ 1.2klarge-video-planner. Python
★ 256acad-homepage.github.io. AcadHomepage: A Modern and Responsive Academic Personal Homepage
★ 2.9kVMwareWorkstation. 手动上传官网的VMwareWorkstation安装包
★ 3.8kTurboDiffusion. TurboDiffusion: 100–200× Acceleration for Video Diffusion Models
★ 3.6kninja. a small build system with a focus on speed
★ 13kPyFleX. Customized Python APIs for NVIDIA FleX
★ 144datasets. TFDS is a collection of datasets ready to use with TensorFlow, Jax, ...
★ 4.6kInternUtopia. A simulation platform for versatile Embodied AI research and developments.
★ 1.3kManiSkill-WidowX250S.
★ 3URDF. URDF file of A1 ,R1.
★ 24sim-evals. A simulation evaluation platform for DROID
★ 235Qwen-VL. The official repo of Qwen-VL (通义千问-VL) chat & pretrained large vision language model proposed by Alibaba Cloud.
★ 6.7kminiball. Fast computation of the smallest enclosing ball of a point set, in low or moderately high dimensions.
★ 152DanceGRPO. An official implementation of DanceGRPO: Unleashing GRPO on Visual Generation
★ 1.6kEmbodiedGen. Towards a Generative 3D World Engine for Embodied Intelligence
★ 5893dgrut. Ray tracing and hybrid rasterization of Gaussian particles
★ 2.3kTRELLIS. Official repo for paper "Structured 3D Latents for Scalable and Versatile 3D Generation" (CVPR'25 Spotlight).
★ 13kgiga-world-0. GigaWorld-0: World Models as Data Engine to Empower Embodied AI
★ 1.6kPointcept. Pointcept: Perceive the world with sparse points, a codebase for point cloud perception research. Latest works: Utonia (ICML'26), Concerto (NeurIPS'25), Sonata (CVPR'25 Highlight), PTv3 (CVPR'24 Oral)
★ 3.1kCtrl-World. ICLR 2026 Paper: Ctrl-World
★ 538giga-train. GigaTrain: An Efficient and Scalable Training Framework for AI Models
★ 1.1kCoACD. [SIGGRAPH2022] Approximate Convex Decomposition for 3D Meshes with Collision-Aware Concavity and Tree Search
★ 1.1kawesome-video-generation. A collection of awesome video generation studies.
★ 778co-tracker. CoTracker is a model for tracking any point (pixel) on a video.
★ 5kdexbotic. Dexbotic: Open-Source Vision-Language-Action Toolbox
★ 1.3kdaam. Diffusion attentive attribution maps for interpreting Stable Diffusion.
★ 803CausVid. (CVPR 2025) From Slow Bidirectional to Fast Autoregressive Video Diffusion Models
★ 1.4kgoku. [CVPR2025 Highlight] Video Generation Foundation Models: https://saiyan-world.github.io/goku/
★ 2.9kInfinity. [CVPR 2025 Oral]Infinity ∞ : Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
★ 1.6kdeskflow. Share a single keyboard and mouse between multiple computers.
★ 28kTrendRadar. ⭐AI-driven public opinion & trend monitor with multi-platform aggregation, RSS, and smart alerts.🎯 告别信息过载,你的 AI 舆情监控助手与热点筛选工具!聚合多平台热点 + RSS 订阅,支持关键词精准筛选。AI 智能筛选新闻 + AI 翻译 + AI 分析简报直推手机,也支持接入 MCP 架构,赋能 AI 自然语言对话分析、情感洞察与趋势预测等。支持 Docker ,数据本地/云端自持。集成微信/飞书/钉钉/Telegram/邮件/ntfy/bark/slack 等渠道智能推送。
★ 61kEmu3.5. Native Multimodal Models are World Learners
★ 1.5krobotics_arXiv_daily. Fetching Embodied AI Paper from ArXiv automatically
★ 331embodied-ai-start. [PKU EPIC Lab] 面向小白的具身智能入门指南
★ 1.1kmy_arXiv_daily. Python
★ 206debugpy. An implementation of the Debug Adapter Protocol for Python
★ 2.4kframe-guidance. [ICLR 2026] Frame Guidance: Training-Free Guidance for Frame-Level Control in Video Diffusion Models
★ 64pytorch3d. PyTorch3D is FAIR's library of reusable components for deep learning with 3D data
★ 9.9knope. [CVPR 2024] PyTorch implementation of NOPE: Novel Object Pose Estimation from a Single Image
★ 218Any6D. [CVPR 2025] Any6D: Model-free 6D Pose Estimation of Novel Objects
★ 478Awesome6DPoseEstimation. Python
★ 318FoundationPose. [CVPR 2024 Highlight] FoundationPose: Unified 6D Pose Estimation and Tracking of Novel Objects
★ 3.5kpanda3d. Powerful, mature open-source cross-platform game engine for Python and C++, developed by Disney and CMU
★ 5.2kbokeh. Interactive Data Visualization in the browser, from Python
★ 20kmegapose6d. Code for "MegaPose: 6D Pose Estimation of Novel Objects via Render & Compare", CoRL 2022.
★ 3633DObjectTracking. Algorithms and Publications on 3D Object Tracking
★ 1kSyncthingWindowsSetup. Syncthing Windows Setup
★ 4kReCamMaster. [ICCV'25 Best Paper Finalist] ReCamMaster: Camera-Controlled Generative Rendering from A Single Video
★ 1.8kdexmimicgen. This code corresponds to simulation environments used as part of the DexMimicGen project.
★ 266open_x_embodiment. Jupyter Notebook
★ 2kgdown. Google Drive public file downloader when curl/wget fails.
★ 5.3khamer. HaMeR: Reconstructing Hands in 3D with Transformers
★ 1.1kcosmos-transfer1-diffusion-renderer. Cosmos-Transfer1-DiffusionRenderer: High-quality video de-lighting and re-lighting based on Cosmos video diffusion framework
★ 833segment-anything. The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 55kfranka_description. Official models of Franka Robotics GmbH robots
★ 134Holistic-Robot-Pose-Estimation. [ECCV 2024] PyTorch implementation of "Real-time Holistic Robot Pose Estimation with Unknown States"
★ 40RoboPEPP. Python
★ 26DREAM. DREAM: Deep Robot-to-Camera Extrinsics for Articulated Manipulators (ICRA 2020)
★ 189mujoco_menagerie. A collection of high-quality models for the MuJoCo physics engine, curated by Google DeepMind.
★ 3.8kGR00T-Dreams. DreamGen: Nvidia GEAR Lab's initiative to solve the robotics data problem using world models
★ 592Cache4Diffusion. Aiming to integrate most existing feature caching-based diffusion acceleration schemes into a unified framework.
★ 110reloadium. Hot Reloading and Profiling for Python
★ 3kopenvla. OpenVLA: An open-source vision-language-action model for robotic manipulation.
★ 6.7kComfyUI. The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
★ 123kVideoX-Fun. 📹 A more flexible framework that can generate videos at any resolution and creates videos from images.
★ 2.2kVACE. [ICCV 2025] Official implementations for paper: VACE: All-in-One Video Creation and Editing
★ 3.9kWan2.2. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kWan2.1. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kHOGformer. Gradient as Conditions: Rethinking HOG for All-in-one Image Restoration
★ 43pytorch-differentiable-histogram. PyTorch implementation for differentiable histogram
★ 23HOG-PyTorch. Batched Histogram of oriented gradients using PyTorch
★ 8PhysTwin. [ICCV 2025] PhysTwin: Physics-Informed Reconstruction and Simulation of Deformable Objects from Videos
★ 435CogVideo. text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
★ 13kLuisaCompute. High-Performance Rendering Framework on Stream Architectures
★ 1kLuisaRender. High-Performance Cross-Platform Monte Carlo Renderer Based on LuisaCompute
★ 612MDL-SDK. NVIDIA Material Definition Language SDK
★ 532FlashGS. C++
★ 226random-fourier-features-pytorch. Implementation of "Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional Domains" by Tancik et al.
★ 116drrobot. Code for "Differentiable Robot Rendering" (CoRL 2024)
★ 187moviepy. Video editing with Python
★ 15kkornia. 🐍 Geometric Computer Vision Library for Spatial AI
★ 11kHunyuanImage-3.0. HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generation
★ 3.2kdiff-gaussian-rasterization. Cuda
★ 1.5kcosmos-transfer1. Cosmos-Transfer1 is a world-to-world transfer model designed to bridge the perceptual divide between simulated and real-world environments.
★ 813wgpu-3dgs-viewer. A 3D Gaussian Splatting Viewer written in Rust using wgpu.
★ 58flash-attention. Fast and memory-efficient exact attention
★ 25krobogs. Python
★ 555Hunyuan3D-2. High-Resolution 3D Assets Generation with Large Scale Hunyuan3D Diffusion Models.
★ 14kInstantMesh. InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models
★ 4.5k3dgs-render-blender-addon. 3DGS Render by KIRI Engine
★ 1.1kMani-GS. [CVPR' 2025'] Mani-GS: Gaussian Splatting Manipulation with Triangular Mesh
★ 197mesh2splat. Fast mesh to 3D gaussian splat conversion
★ 892Bagel. Open-source unified multimodal model
★ 6.1kGenSim2. Python
★ 85ReKep. ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation
★ 976RynnVLA-001. [ICRA 2026] RynnVLA-001: Using Human Demonstrations to Improve Robot Manipulation
★ 303astro-theme-pure. ⭐ A simple, fast and powerful blog & document theme built by Astro
★ 1kInternHumanoid. A versatile, all-in-one toolbox for whole-body humanoid robot control.
★ 185