This is your work, valued
(*^▽^*)
TinyVLA. Python
★ 86Worldeval. Python
★ 29daily_management. 日常管理系统
★ 3vera-ui. TypeScript
★ 2VeraBlog. HTML
★ 1Milimili. music PWA
★ 1giga-world-1. A Roadmap to Build World Models for Robot Policy Evaluation
★ 782Pathwise_TTC.
★ 18FramePack. Lets make video diffusion practical!
★ 17kSelf-Forcing. Official codebase for "Self Forcing: Bridging Training and Inference in Autoregressive Video Diffusion" (NeurIPS 2025 Spotlight)
★ 3.5kStable-Video-Infinity. [ICLR 26 Oral] Stable Video Infinity: Infinite-Length Video Generation with Error Recycling
★ 2.5kvggt-omega. [CVPR 2026 Oral] VGGT Omega
★ 3.7kLance. A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
★ 1.3kRollingForcing. [ICLR 2026] Official Repo for Rolling Forcing: Autoregressive Long Video Diffusion in Real Time
★ 449PointWorld. PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation
★ 468mmpose. OpenMMLab Pose Estimation Toolbox and Benchmark.
★ 7.8kdas-datakit. Python
★ 54lingbot-world. Advancing Open-source World Models
★ 4.3kcc9s. A k9s-inspired CLI and TUI for managing Claude Code sessions — browse, search, inspect, and clean up your AI coding sessions
★ 69interactive_world_sim. [RSS 2026] Interactive World Simulator for Robot Policy Training and Evaluation
★ 281DreamDojo. Official Codebase for "DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos" (ICML 2026)
★ 1kWMPO. Official Implementation of Paper: WMPO: World Model-based Policy Optimization for Vision-Language-Action Models
★ 227world-model-eval. Code for "Evaluating Robot Policies in a World Model".
★ 102Ctrl-World. ICLR 2026 Paper: Ctrl-World
★ 538InfinityStar. [NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation
★ 774LLaMA-VID. LLaMA-VID: An Image is Worth 2 Tokens in Large Language Models (ECCV 2024)
★ 861MA-LMM. (2024CVPR) MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
★ 350ICEdit. [NeurIPS 2025] Image editing is worth a single LoRA! 0.1% training data for fantastic image editing! Surpasses GPT-4o in ID persistence~ MoE ckpt released! Only 4GB VRAM is enough to run!
★ 2.1kOpenTrajBooster. Official implementation of TrajBooster
★ 189video-prediction-policy. Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations https://video-prediction-policy.github.io
★ 407Worldeval. Python
★ 29Bagel. Open-source unified multimodal model
★ 6.1kcosmos-predict2. Cosmos-Predict2 is a collection of general-purpose world foundation models for Physical AI that can be fine-tuned into customized world models for downstream applications.
★ 794Awesome-World-Models. A comprehensive list of papers for the definition of World Models and using World Models for General Video Generation, Embodied AI, and Autonomous Driving, including papers, codes, and related websites.
★ 1.9kHunyuanWorld-1.0. Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels with Hunyuan3D World Model
★ 2.9kVideoWorld. [CVPR 2025] VideoWorld is a simple generative model that learns purely from unlabeled videos—much like how babies learn by observing their environment.
★ 793OmniGen2. OmniGen2: Exploration to Advanced Multimodal Generation. https://arxiv.org/abs/2506.18871
★ 4.1klearn-python3. Jupyter notebooks for teaching/learning Python 3
★ 6.8kUniVLA. [RSS 2025] Learning to Act Anywhere with Task-centric Latent Actions
★ 1.1kAwesome-From-Video-Generation-to-World-Model. A list of works on video generation towards world model
★ 504SimpleVLA-RL. [ICLR 2026] SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
★ 1.8kATI. Official implementation of ATI: Any Trajectory Instruction for Controllable Video Generation. https://arxiv.org/pdf/2505.22944
★ 355EX-4D. The implementation of Extreme Viewpoint 4D Video Generation
★ 267vjepa2. PyTorch code and models for VJEPA2 self-supervised learning from video.
★ 4.4kjepa. PyTorch code and models for V-JEPA self-supervised learning from video.
★ 4.1kDepth-Anything-V2. [NeurIPS 2024] Depth Anything V2. A More Capable Foundation Model for Monocular Depth Estimation
★ 8.6kEWMBench. Official code for EWMBench: Evaluating Scene, Motion, and Semantic Quality in Embodied World Models
★ 129Open-Sora-Plan. This project aim to reproduce Sora (Open AI T2V model), we wish the open source community contribute to this project.
★ 12kEasyControl. Implementation of "EasyControl: Adding Efficient and Flexible Control for Diffusion Transformer"(ICCV2025)
★ 1.7kOpenHelix. OpenHelix: An Open-source Dual-System VLA Model for Robotic Manipulation
★ 389MAGI-1. MAGI-1: Autoregressive Video Generation at Scale
★ 3.7kvq_bet_official. Official code for "Behavior Generation with Latent Actions" (ICML 2024 Spotlight)
★ 211Step-Video-TI2V. Python
★ 374open-genie. Pytorch implementation of "Genie: Generative Interactive Environments", Bruce et al. (2024).
★ 2921xgpt. world modeling challenge for humanoid robots
★ 564EmbodiedEval. Evaluate Multimodal LLMs as Embodied Agents
★ 59AMBER. An LLM-free Multi-dimensional Benchmark for Multi-modal Hallucination Evaluation
★ 173Mirror. Python
★ 14QueryAgent. Code and data for QueryAgent(ACL 2024)
★ 21g1. g1: Using Llama-3.1 70b on Groq to create o1-like reasoning chains
★ 4.2klumos. Code and data for "Lumos: Learning Agents with Unified Data, Modular Design, and Open-Source LLMs"
★ 477avatar. (NeurIPS 2024) AvaTaR: Optimizing LLM Agents for Tool Usage via Contrastive Reasoning
★ 241ToRA. ToRA is a series of Tool-integrated Reasoning LLM Agents designed to solve challenging mathematical reasoning problems by interacting with tools [ICLR'24].
★ 1.1krefiner. About The corresponding code from our paper " REFINER: Reasoning Feedback on Intermediate Representations" (EACL 2024). Do not hesitate to open an issue if you run into any trouble!
★ 76openr. OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models
★ 1.8ksearch-and-learn. Recipes to scale inference-time compute of open models
★ 1.1kLLMo1Wrapper. A Python wrapper that enables large language models (LLMs) to simulate the step-by-step thinking process of OpenAI’s o1 model, providing users with detailed reasoning and comprehensive answers.
★ 3DPO-ST. [ACL 2024] Self-Training with Direct Preference Optimization Improves Chain-of-Thought Reasoning
★ 54genesis-world. Simulation platform for general-purpose robotics & embodied AI learning.
★ 30kDoLa. Official implementation for the paper "DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models"
★ 557PromptGating4MCTG. This is the repo for our work “An Extensible Plug-and-Play Method for Multi-Aspect Controllable Text Generation” (ACL 2023).
★ 14alfworld. ALFWorld: Aligning Text and Embodied Environments for Interactive Learning
★ 813clip-image-search. Search images with a text or image query, using Open AI's pretrained CLIP model.
★ 268Style-LLM. Python
★ 10sNeuron-TST. EMNLP 2024 | Style-Specific Neurons for Steering LLMs in Text Style Transfer
★ 14annotated-transformer. An annotated implementation of the Transformer paper.
★ 7.4kTransformer. Transformer模型复现
★ 9Reproduce_the_Transformer_model. 通过阅读论文Attention is all you need来复现Transformer模型
★ 12nanoCLIP. A lightweight Text-to-Image Retrieval model [Web App]
★ 29image-retrieval. Content-Based Image Retrieval (CBIR) using Faiss (Facebook) and many different feature extraction methods ( VGG16, ResNet50, Local Binary Pattern, RGBHistogram)
★ 46image-retrieval. 图片向量检索服务,包含Numpy、Faiss、ES、Milvus多种计算引擎
★ 138TinyStyler. Code for TinyStyler: Efficient Few-Shot Text Style Transfer with Authorship Embeddings, EMNLP 2024 Findings
★ 32stylellm_models. StyleLLM文风大模型:基于大语言模型的文本风格迁移项目。Text style transfer base on Large Language Model. #文字修饰 # 润色 #风格模仿
★ 359CDR_Meets_LLMs. Python
★ 4Ask-before-Plan. [EMNLP 2024] Ask-before-Plan: Proactive Language Agents for Real-World Planning
★ 24XAgent. An Autonomous LLM Agent for Complex Task Solving
★ 8.5kTell_Me_More. Repo for paper "Tell Me More! Towards Implicit User Intention Understanding of Language Model Driven Agents"
★ 65MAGE. Machine-generated text detection in the wild (ACL 2024)
★ 229OpenScholar. This repository includes the official implementation of OpenScholar: Synthesizing Scientific Literature with Retrieval-augmented LMs.
★ 1.6kProactiveAgent. A LLM-based Agent that predict its tasks proactively.
★ 640quiet-star. Code for Quiet-STaR
★ 739Awesome-LLM-Reasoning. From Chain-of-Thought prompting to OpenAI o1 and DeepSeek-R1 🍓
★ 3.7kjittor. Jittor is a high-performance deep learning framework based on JIT compiling and meta-operators.
★ 3.2kface_recognition. The world's simplest facial recognition api for Python and the command line
★ 57kecnu-PGCourseShare. 华东师范大学研究生课程攻略共享计划
★ 282Langchain-Chatchat. Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain-ChatGLM), local knowledge based LLM (like ChatGLM, Qwen and Llama) RAG and Agent app with langchain
★ 38kanything-llm. Stop renting your intelligence. Own it with AnythingLLM. Everything you need for a powerful local-first agent experience
★ 64kAI-Scientist. The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery 🧑🔬
★ 14kgpt_academic. 为GPT/GLM等LLM大语言模型提供实用化交互接口,特别优化论文阅读/润色/写作体验,模块化设计,支持自定义快捷按钮&函数插件,支持Python和C++等项目剖析&自译解功能,PDF/LaTex论文翻译&总结功能,支持并行问询多种LLM模型,支持chatglm3等本地模型。接入通义千问, deepseekcoder, 讯飞星火, 文心一言, llama2, rwkv, claude2, moss等。
★ 71kawesome-chatgpt-zh. ChatGPT 中文指南🔥,ChatGPT 中文调教指南,指令指南,应用开发指南,精选资源清单,更好的使用 chatGPT 让你的生产力 up up up! 🚀
★ 12kGPT_API_free. Free ChatGPT&DeepSeek API Key,免费ChatGPT&DeepSeek API。免费接入DeepSeek API和GPT4 API,支持 gpt | deepseek | claude | gemini | grok 等排名靠前的常用大模型。
★ 39kPBR-White-Paper. ⚡️基于物理的渲染(PBR)白皮书 | White Paper of Physically Based Rendering(PBR)
★ 2kswiper. Most modern mobile touch slider with hardware accelerated transitions
★ 42kreact-device-mockups. 📱React Wrapper for html5-device-mockups
★ 17camera-controls. A camera control for three.js, similar to THREE.OrbitControls yet supports smooth transitions and more features.
★ 2.4kmodel-viewer. Easily display interactive 3D models on the web and in AR!
★ 8.2kthree-dat.gui. A package which create THREE.js controls on Dat.GUI
★ 37calendar. React Calendar
★ 1.7kdaycaca. 🐣 a pure JavaScript library to handle image via canvas
★ 57web-vitals. Essential metrics for a healthy site.
★ 8.6kloadable-components. The recommended Code Splitting library for React ✂️✨
★ 7.8kengine. A typescript interactive engine, support 2D, 3D, animation, physics, built on WebGL and glTF.
★ 5.9kShadowEditor. Cross-platform 3D scene editor based on three.js, golang and mongodb for desktop and web.
★ 1.7kwebglstudio.js. A full open source 3D graphics editor in the browser, with scene editor, coding pad, graph editor, virtual file system, and many features more.
★ 5.3kwhs.js. :rocket: 🌪 Super-fast 3D framework for Web Applications 🥇 & Games 🎮. Based on Three.js
★ 6.3kSketchbook. 3D playground built on three.js and cannon.js.
★ 1.7kreact-three-fiber. 🇨🇭 A React renderer for Three.js
★ 32kthree.js. JavaScript 3D Library.
★ 114kglTF-Tutorials. glTF Tutorials
★ 26Hilo3d. 🎮 A 3D WebGL Rendering Engine
★ 661vudio.js. 音频可视化展示模块
★ 255kBooks. kindle网络电子书资源
★ 68tailwindcss. A utility-first CSS framework for rapid UI development.
★ 96ktonovel-go. tonovel是一个简洁,干净的小说聚合系统
★ 88go-admin. 基于go+gin+vue+element admin 后台管理系统,支持用户管理,认证,内容管理等
★ 418react-use. React Hooks — 👍
★ 44kroomler. Roomler - Multi-party Video Conferencing & Team Collaboration Tool using WebRTC (Janus Gateway)
★ 293gp. General purpose multi-party video conferencing application using WebRTC and signalling over Firebase Firestore in a mesh topology.
★ 7webrtc-mesh-network. Demo for a webrtc mesh network, communicate and watch stuff with friends in real time.
★ 5meething. dWebRTC Video Meetings MESH/SFU hybrid using GunDB, MediaSoup and Beyond!
★ 463flutter-sfu-video-call-group. Group video call mobile app built with Flutter and WebRTC SFU architecture
★ 31edumeet. edumeet - multiparty web-meetings using mediasoup and WebRTC
★ 1.3kvideo-conference-webrtc. 🐀Complete client/server application demonstrating how to setup a video conference with multiple peers using WebRTC.
★ 126webrtc-stream. 🍧🍭😻包括但不局限于 WebRTC 的各种栗子
★ 404WiLearning. Multiparty meeting&e-learning using mediasoup, webrtc ,angular and ionic with powerful whiteboard support
★ 494Blog. Blogs about Technology - 一些技术文章
★ 44fe-code. 🍹🍰 愉快的写代码~(文章合集)
★ 361esbuild. An extremely fast bundler for the web
★ 40kantd-tools. 🔧 Cli Tools for Ant Design React
★ 407copyfiles. copy files on the command line
★ 422ant-design. An enterprise-class UI design language and React UI library
★ 99kastexplorer. A web tool to explore the ASTs generated by various parsers.
★ 6.5kTypeScript. TypeScript is a superset of JavaScript that compiles to clean JavaScript output.
★ 110kintermock. Mocking library to create mock objects with fake data for TypeScript interfaces
★ 1.1kvscode-extension-samples. Sample code illustrating the VS Code extension API.
★ 10kvscode-bookmarks. Bookmarks Extension for Visual Studio Code
★ 2.1khtmlparser2. The fast & forgiving HTML and XML parser
★ 4.8kvue3-mindmap. Mindmap component for Vue3
★ 454Vue.D3.tree. Vue component to display tree based on D3.js layout.
★ 897vue3-News. 🔥 Find the latest breaking Vue3、Vue CLI 3+ & Vite News. (2024/2025)
★ 2.9klerna. Lerna is a fast, modern build system for managing and publishing multiple JavaScript/TypeScript packages from the same repository.
★ 36kreact-native-video-project. 一个基于react-native的纯跨平台的影视项目,欢迎大家star
★ 356ReactNativeEveryDayRead. React Native实现的阅读器,模仿观止app
★ 22react-native-guide. React Native指南汇集了各类react-native学习资源、开源App和组件
★ 18kservice-workers. A collection of utilities for creating/testing/experimenting with service workers.
★ 1.3kvist. Virtual-list component build with react and rxjs
★ 90rxviz. Rx Visualizer - Animated playground for Rx Observables
★ 1.6krxmarbles. Interactive diagrams of Rx Observables
★ 4.2kzao-interview. 📖常见面试问题
★ 3dva. 🌱 React and redux based, lightweight and elm-style framework. (Inspired by elm and choo)
★ 16kleetcode-javascript. :beers: 喝杯小酒,一起做题。前端攻城狮从零入门算法的宝藏题库,根据知名算法老师的经验总结了 100+ 道 LeetCode 力扣的经典题型 JavaScript 题解和思路。已按题目类型分 label,一起加油。
★ 2.1kfree-api. 收集免费的接口服务,做一个api的搬运工
★ 16kSmileVue. 这是一个移动端的电商系统,前端使用了Vue,后端使用了Koa2.
★ 451JavaScript-Algorithms. 基础理论+JS框架应用+实践,从0到1构建整个前端算法体系
★ 5.7kreact-cloud-music. React 16.8打造精美音乐WebApp
★ 1.9kfucking-algorithm. Crack LeetCode, not only how, but also why.
★ 135kdaily_management. 日常管理系统
★ 3VeraBlog. HTML
★ 1Koa2-blog. node+koa2+mysql
★ 887React-Whole-barrels. webapck + react + react-router +dva + es6 + less + antd 实现的脚手架 👌👌
★ 24mywebpack. JavaScript
★ 108node-react-koa. 一个简单的前端react 后端node koa 的用户增删改查小web
★ 32blog. 博客
★ 844reactjs101. 从零开始学 ReactJS(ReactJS 101)是一本希望让初学者一看就懂的 ReactJS 中文入门教学书,由浅入深学习 ReactJS 生态系 (Flux, Redux, React Router, ImmutableJS, React Native, Relay/GraphQL etc.)。该分支为转码简体中文版
★ 792programmer-book. 公众号:普通程序员
★ 2kfe-interview. 前端面试每日 3+1,以面试题来驱动学习,提倡每日学习与思考,每天进步一点!每天早上5点纯手工发布面试题(死磕自己,愉悦大家),6000+道前端面试题全面覆盖,HTML/CSS/JavaScript/Vue/React/Nodejs/TypeScript/ECMAScritpt/Webpack/Jquery/小程序/软技能……
★ 26kalbum-express. node编写,express搭建的简单相册,功能有查看相册内图片,新建相册文件夹,上传图片。看这个代码可以简单了解下前后端交互的流程
★ 7react-projects. React从入门到放弃 -- 项目实战
★ 287verify-slide. 纯前端的滑动验证
★ 47jigsaw. canvas滑动验证码
★ 710DataStructure. C
★ 7