This is your work, valued
HA3D_simulator. Official implementation of Human-Aware Vision-and-Language Navigation: Bridging Simulation to Reality with Dynamic Human Interactions (NeurIPS DB Track'24 Spotlight).
★ 58simple-TikTok.
★ 2OpenAI-learning. openAI chatgdp 开始
★ 1OmniNavBench. [RSS 2026] Official code & data for "OmniNavBench: Beyond Isolation — A Unified Benchmark for General-Purpose Navigation"
★ 87UrbanVerse. Scaling Urban Simulation - Infinite Physically-Plausible Urban Simulation = IsaacSim(Physically-Accurate Assets × Real-World City-Tour Layouts)
★ 61tongsim. A high-fidelity, general-purpose platform for embodied agent training and testing.
★ 188VLNTube. Python
★ 28InternUtopia. A simulation platform for versatile Embodied AI research and developments.
★ 1.3kSAGE-3D_Official. This is the official repository of the paper "Towards Physically Executable 3D Gaussian for Embodied Navigation".
★ 194timesfm. TimesFM (Time Series Foundation Model) is a pretrained time-series foundation model developed by Google Research for time-series forecasting.
★ 27kTradingView-API. 📈 Get real-time stocks from TradingView
★ 4.2kVibe-Trading. "Vibe-Trading: Your Personal Trading Agent"
★ 29kHorizon. 📡 Your own AI-powered news radar. Generates daily briefings in English & Chinese. | 用 AI 构建你专属的新闻雷达
★ 8.6kclash. Clash官网各版本Clash下载地址及备份下载地址
★ 7kAI-Trader. "AI-Trader: 100% Fully-Automated Agent-Native Trading"
★ 21kserenity-skill. Serenity-inspired Agent Skill for supply-chain bottleneck stock research
★ 3.6kserenity-aleabitoreddit. Installable Serenity tweet archive + AI/semi supply-chain skill. Install: npx skills add yan-labs/serenity-aleabitoreddit
★ 459TradingAgents. TradingAgents: Multi-Agents LLM Financial Trading Framework
★ 95kclaw-code. An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.
★ 195kECC. The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
★ 237kGalbotSDK. C++
★ 59OpenCLI. Make Any Website into CLI & Use your logged-in browser by AI agent.
★ 28kdimos. Dimensional is the agentic operating system for physical space. Command humanoids, quadrupeds, drones, and other hardware platforms in natural language and build multi-agent systems that work seamlessly with physical input (cameras, lidar, actuators).
★ 3.8kstreaming-llm. [ICLR 2024] Efficient Streaming Language Models with Attention Sinks
★ 7.3kisaac-go2-ros2. Unitree Go2 simulation platform for testing navigation, decision-making and autonomous tasks. (NVIDIA Isaac/ROS2)
★ 565IsaacSim. NVIDIA Isaac Sim™ is an open-source application on NVIDIA Omniverse for developing, simulating, and testing AI-driven robots in realistic virtual environments.
★ 3.8kIsaacLab. Unified framework for robot learning built on NVIDIA Isaac Sim
★ 7.8kHunyuan3D-2. High-Resolution 3D Assets Generation with Large Scale Hunyuan3D Diffusion Models.
★ 14kNavFoM-Web. JavaScript
★ 3deep_sort. Simple Online Realtime Tracking with a Deep Association Metric
★ 6.2kboxmot. BoxMOT: Pluggable Python and C++ SOTA multi-object tracking modules with support for axis-aligned and oriented bounding boxes
★ 8.3kYolov5-Deepsort-Fastreid. Python
★ 288LIO-SAM. LIO-SAM: Tightly-coupled Lidar Inertial Odometry via Smoothing and Mapping
★ 4.9kLiteReality. [NeurIPS 2025] LiteReality: Graphics-Ready 3D Scene Reconstruction from RGB-D Scans
★ 389StreamVLN. [ICRA 2026] Official implementation of the paper: "StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling"
★ 565unitree_ros. C++
★ 1.5krealsense-ros. ROS Wrapper for RealSense™ Cameras
★ 3.4ksnowboy. Future versions with model training module will be maintained through a forked version here: https://github.com/seasalt-ai/snowboy
★ 3.4kOmniPerception. Python
★ 519TrackVLA. [CoRL 2025] Repository relating to "TrackVLA: Embodied Visual Tracking in the Wild"
★ 420Segment-Any-Anomaly. Official implementation of "Segment Any Anomaly without Training via Hybrid Prompt Regularization (SAA+)".
★ 843RealNet. Offical implementation of "RealNet: A Feature Selection Network with Realistic Synthetic Anomaly for Anomaly Detection (CVPR 2024)"
★ 427AdaCLIP. [ECCV2024] The Official Implementation for ''AdaCLIP: Adapting CLIP with Hybrid Learnable Prompts for Zero-Shot Anomaly Detection''
★ 308piper. A fast, local neural text to speech system
★ 11kYOLO-World. [CVPR 2024] Real-Time Open-Vocabulary Object Detection
★ 6.5kwhisper. Robust Speech Recognition via Large-Scale Weak Supervision
★ 106kteb_local_planner. An optimal trajectory planner considering distinctive topologies for mobile robots based on Timed-Elastic-Bands (ROS Package)
★ 1.3kunitree_ros2. C++
★ 780unitree_sdk2. Unitree robot sdk version 2. https://support.unitree.com/home/zh/developer
★ 1.3kunitree_sdk2_python. Python interface for unitree sdk2
★ 751genesis-world. Simulation platform for general-purpose robotics & embodied AI learning.
★ 30kCM2. Python
★ 59waypoint-predictor. Training code of waypoint predictor in Discrete-to-Continuous VLN.
★ 32PQDiff. [ICLR 2024] Continuous-Multiple Image Outpainting in One-Step via Positional Query and A Diffusion-based Approach Link: https://arxiv.org/abs/2401.15652
★ 92LongVU. [ICML 2025] Official PyTorch implementation of LongVU
★ 431habitat-lab. A modular high-level library to train embodied AI agents across a variety of tasks and environments.
★ 3.1kjetson-containers. Machine Learning Containers for NVIDIA Jetson and JetPack-L4T
★ 4.8kjetson_dla_tutorial. A tutorial for getting started with the Deep Learning Accelerator (DLA) on NVIDIA Jetson
★ 371openvla. OpenVLA: An open-source vision-language-action model for robotic manipulation.
★ 6.7kml-slowfast-llava. SlowFast-LLaVA: A Strong Training-Free Baseline for Video Large Language Models
★ 291Embodied_AI_Paper_List. [Embodied-AI-Survey-2025] Paper List and Resource Repository for Embodied AI
★ 2.1kSim2Real-VLN-3DFF. Official implementation of Sim-to-Real Transfer via 3D Feature Fields for Vision-and-Language Navigation (CoRL'24).
★ 80afford-motion. Official implementation of CVPR24 highlight paper "Move as You Say, Interact as You Can: Language-guided Human Motion Generation with Scene Affordance"
★ 177home-robot. Mobile manipulation research tools for roboticists
★ 1.2kVLN-BEVBert. [ICCV 2023] Official repo of "BEVBert: Multimodal Map Pre-training for Language-guided Navigation"
★ 260LLaMA-VID. LLaMA-VID: An Image is Worth 2 Tokens in Large Language Models (ECCV 2024)
★ 861Awesome-Multimodal-Large-Language-Models. :sparkles::sparkles:Latest Advances on Multimodal Large Language Models
★ 18kAcFormer. Python
★ 29freeCodeCamp. freeCodeCamp.org's open-source codebase and curriculum. Learn math, programming, and computer science for free.
★ 453kOpenFace. OpenFace – a state-of-the art tool intended for facial landmark detection, head pose estimation, facial action unit recognition, and eye-gaze estimation.
★ 7.7kFastChat. An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
★ 40kLAVIS. LAVIS - A One-stop Library for Language-Vision Intelligence
★ 11kDreamPose. Official implementation of "DreamPose: Fashion Image-to-Video Synthesis via Stable Diffusion"
★ 1kMERTools. Toolkits for Multimodal Emotion Recognition
★ 326MM-Diffusion. [CVPR'23] MM-Diffusion: Learning Multi-Modal Diffusion Models for Joint Audio and Video Generation
★ 453Video-Motion-Customization. VMC: Video Motion Customization using Temporal Attention Adaption for Text-to-Video Diffusion Models (CVPR 2024)
★ 199generative-models. Generative Models by Stability AI
★ 27kHA3D_simulator. Official implementation of Human-Aware Vision-and-Language Navigation: Bridging Simulation to Reality with Dynamic Human Interactions (NeurIPS DB Track'24 Spotlight).
★ 58speaker_follower. Code release for Fried et al., Speaker-Follower Models for Vision-and-Language Navigation. in NeurIPS, 2018.
★ 138vilbert-multi-task. Multi Task Vision and Language
★ 824gpt_academic. 为GPT/GLM等LLM大语言模型提供实用化交互接口,特别优化论文阅读/润色/写作体验,模块化设计,支持自定义快捷按钮&函数插件,支持Python和C++等项目剖析&自译解功能,PDF/LaTex论文翻译&总结功能,支持并行问询多种LLM模型,支持chatglm3等本地模型。接入通义千问, deepseekcoder, 讯飞星火, 文心一言, llama2, rwkv, claude2, moss等。
★ 71kvln-bert. Code for the paper "Improving Vision-and-Language Navigation with Image-Text Pairs from the Web" (ECCV 2020)
★ 59ScaleVLN. [ICCV 2023 Oral]: Scaling Data Generation in Vision-and-Language Navigation
★ 225CLIP. CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image
★ 34kunitree_legged_sdk. SDK tools for control robots.
★ 425UnitreecameraSDK. Unitree GO1 camera SDK
★ 120habitat-sim. A flexible, high-performance 3D simulator for Embodied AI research.
★ 3.8kVLN-CE. Vision-and-Language Navigation in Continuous Environments using Habitat
★ 844FollowYourPose. [AAAI 2024] Follow-Your-Pose: This repo is the official implementation of "Follow-Your-Pose : Pose-Guided Text-to-Video Generation using Pose-Free Videos"
★ 1.4kstable-diffusion. A latent text-to-image diffusion model
★ 73kdiffusers. 🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
★ 34kpaper-template. ECCV 2024 paper template
★ 68YouTube-VLN. [ICCV'23] Learning Vision-and-Language Navigation from YouTube Videos
★ 71airbert. Codebase for the Airbert paper
★ 46airbert-recurrentvln. Python
★ 7jetson-inference. Hello AI World guide to deploying deep-learning inference networks and deep vision primitives with TensorRT and NVIDIA Jetson.
★ 8.9kvideo_features. Extract video features from raw videos using multiple GPUs. We support RAFT flow frames as well as S3D, I3D, R(2+1)D, VGGish, CLIP, and TIMM models.
★ 654anomalydiffusion. [AAAI 2024] AnomalyDiffusion: Few-Shot Anomaly Image Generation with Diffusion Model
★ 325DiffusionAD. [TPAMI] DiffusionAD: Norm-guided One-step Denoising Diffusion for Anomaly Detection
★ 221DDAD. Python
★ 202awesome-industrial-anomaly-detection. Paper list and datasets for industrial image anomaly/defect detection (updating). 工业异常/瑕疵检测论文及数据集检索库(持续更新)。
★ 3.7kanomalib. An anomaly detection library comprising state-of-the-art algorithms and features such as experiment management, hyper-parameter optimization, and edge inference.
★ 6kDistDepth. Repository for "Toward Practical Monocular Indoor Depth Estimation" (CVPR 2022)
★ 237Matterport3DSimulator. AI Research Platform for Reinforcement Learning from Real Panoramic Images.
★ 709Matterport. Matterport3D is a pretty awesome dataset for RGB-D machine learning tasks :)
★ 1.2kawesome-vision-language-navigation. A curated list for vision-and-language navigation. ACL 2022 paper "Vision-and-Language Navigation: A Survey of Tasks, Methods, and Future Directions"
★ 601motion-diffusion-model. The official PyTorch implementation of the paper "Human Motion Diffusion Model"
★ 4.1kT2M-GPT. (CVPR 2023) Pytorch implementation of “T2M-GPT: Generating Human Motion from Textual Descriptions with Discrete Representations”
★ 772awesome-anomaly-detection. A curated list of awesome anomaly detection resources
★ 2.9kmixed-segdec-net-comind2021. Official PyTorch implementation for "Mixed supervision for surface-defect detection: from weakly to fully supervised learning"
★ 319Surface-Defect-Detection. 📈 目前最大的工业缺陷检测数据库及论文集 Constantly summarizing open source dataset and critical papers in the field of surface defect research which are of great importance.
★ 4.1kmodelscope. ModelScope: bring the notion of Model-as-a-Service to life.
★ 9.1kmake-a-video-pytorch. Implementation of Make-A-Video, new SOTA text to video generator from Meta AI, in Pytorch
★ 2kkairos-pub.
★ 1llama. Inference code for Llama models
★ 60kAwesome-Crowd-Localization. Awesome Crowd Localization
★ 47Awesome-Crowd-Counting. Awesome Crowd Counting
★ 2.6kCVPR2024-Paper-Code-Interpretation. cvpr2024/cvpr2023/cvpr2022/cvpr2021/cvpr2020/cvpr2019/cvpr2018/cvpr2017 论文/代码/解读/直播合集,极市团队整理
★ 12kAwesome-Video-Generation. Python
★ 41video-generation-survey. A reading list of video generation
★ 723Awesome-Video-Diffusion. A curated list of recent diffusion models for video generation, editing, and various other applications.
★ 5.7kgpt4free. The official gpt4free repository | various collection of powerful language models | opus 4.6 gpt 5.3 kimi 2.5 deepseek v3.2 gemini 3
★ 67kORB_SLAM3_detailed_comments. Detailed comments for ORB-SLAM3
★ 1.5kORB_SLAM2_detailed_comments. Detailed comments for ORB-SLAM2 with trouble-shooting, key formula derivation, and diagrammatic drawing
★ 1.7ksegment-anything. The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 55kORB_SLAM3. ORB-SLAM3: An Accurate Open-Source Library for Visual, Visual-Inertial and Multi-Map SLAM
★ 8.9kVINS-Mono. A Robust and Versatile Monocular Visual-Inertial State Estimator
★ 6kopencv_zoo. Model Zoo For OpenCV DNN and Benchmarks.
★ 1kopen_model_zoo. Pre-trained Deep Learning models and demos (high quality and extremely fast)
★ 4.4kdarknet. Convolutional Neural Networks
★ 26kunitree_ros2_to_real. A ROS2 package you can use to control the real Go1 robot
★ 84STSGCN. AAAI 2020. Spatial-Temporal Synchronous Graph Convolutional Networks: A New Framework for Spatial-Temporal Network Data Forecasting
★ 47openai-cookbook. Examples and guides for using the OpenAI API
★ 75kopenai-python. The official Python library for the OpenAI API
★ 31kgo-openai. OpenAI ChatGPT, GPT-5, GPT-Image-1, Whisper API clients for Go
★ 11kCppTemplateTutorial. 中文的C++ Template的教学指南。与知名书籍C++ Templates不同,该系列教程将C++ Templates作为一门图灵完备的语言来讲授,以求帮助读者对Meta-Programming融会贯通。(正在施工中)
★ 11kmodern-cpp-tutorial. 📚 Modern C++ Tutorial: C++11 to C++26 On the Fly | https://changkun.de/modern-cpp/
★ 26kC-Plus-Plus. Collection of various algorithms in mathematics, machine learning, computer science and physics implemented in C++ for educational purposes.
★ 35kCPlusPlusThings. C++那些事
★ 43ksimple-TikTok.
★ 2