This is your work, valued
ECE PhD Research Assistent Homepage: liqunchen0606.github.io
Graph-Optimal-Transport. Code for ICML 2020 "Graph Optimal Transport for Cross-Domain Alignment"
★ 164OT-Seq2Seq. code for paper "Improving Sequence-to-Sequence Learning via Optimal Transport"
★ 68Triangle-GAN. implementation for NIPS paper Triangle Generative Adversarial Networks
★ 60Symmetric-VAE. https://arxiv.org/abs/1709.01846
★ 9Seq2Seq-OT. Tensorflow Implementation
★ 5Jointly-Discovering-Visual-Objects-and-Spoken-Words. an implementation for paper Jointly Discovering Visual Objects and Spoken Words
★ 5Resume. My Resume
★ 2data-science-question-answer. A repo for data science related questions and answers
★ 2PyTorch-GAN. PyTorch implementations of Generative Adversarial Networks.
★ 1HCL. ICLR 2021, Contrastive Learning with Hard Negative Samples
★ 1OpenAI-CLIP. Simple implementation of OpenAI CLIP model in PyTorch.
★ 1image-gpt. Python
★ 1LM-BFF. Python
★ 1SWEM. The Tensorflow code for this ACL 2018 paper: "Baseline Needs More Love: On Simple Word-Embedding-Based Models and Associated Pooling Mechanisms"
★ 1ud032. Data Wrangling with MongoDB class code
★ 1clawterminal-docs. OpenClaw Terminal — Guides & Tutorials
★ 11hermes-agent. The agent that grows with you
★ 223kskillforge. Quality layer for Agent Skills (agentskills.io) — lint, hash, sign, version, eval your SKILL.md files. 10 structural checks, Ed25519 signing, semantic diff, 3-stage eval pipeline. Bidirectional SKILL.md ↔ AIF conversion.
★ 1claw-code-parity. Join Discord: https://discord.gg/5TUQKqFWd / claw-code Rust port parity work - it is temporary work while claw-code repo is doing migration
★ 6.7kclaw-code. An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.
★ 195kmarkit. 🖍️ Convert anything to markdown. Mark it.
★ 1.3kfast-weight-attention. Implementation of Fast Weight Attention
★ 33RIM-pytorch. Implementation of Recurrent Independent Mechanisms in Pytorch
★ 27skills_public. My Public Skills for Agents
★ 2CLI-Anything. "CLI-Anything: Making ALL Software Agent-Native" -- CLI-Hub: https://clianything.cc/
★ 46kpi-multi-pass. Multi-subscription extension for pi -- use multiple OAuth accounts per provider (Anthropic, Codex, Copilot, Gemini, Antigravity)
★ 464OpenClaw-RL. OpenClaw-RL: Train any agent simply by talking
★ 5.6kllm_training_handbook. An open collection of methodologies to help with successful training of large language models.
★ 565autoresearch. AI agents running research on single-GPU nanochat training automatically
★ 93ksuperpowers. An agentic skills framework & software development methodology that works.
★ 264kLLMs-from-scratch. Implement a ChatGPT-like LLM in PyTorch from scratch, step by step
★ 100kMegatron-LM. Ongoing research training transformer models at scale
★ 17kopenclaw. Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
★ 385kawesome-claude-skills. A curated list of awesome Claude Skills, resources, and tools for customizing Claude AI workflows
★ 71ktiny8. A tiny CPU simulator written in Python
★ 1.4kgepa. Optimize prompts, code, and more with AI-powered Reflective Optimization
★ 5.9ktokenizer-ui. HTML
★ 48spacy-models. 💫 Models for the spaCy Natural Language Processing (NLP) library
★ 1.9kAI-Trader. "AI-Trader: 100% Fully-Automated Agent-Native Trading"
★ 21kopenfold-3. A fully open source biomolecular structure prediction model based on AlphaFold3
★ 808nof0. NOF0 - 开源的 AI 交易竞技场
★ 2.8kDeepSeek-OCR. Contexts Optical Compression
★ 24knanochat. The best ChatGPT that $100 can buy.
★ 57kdexter. An autonomous agent for deep financial research
★ 27ktinker-cookbook. Post-training with Tinker
★ 4kKronos. Kronos: A Foundation Model for the Language of Financial Markets
★ 35knpq-vit. [ICLR 2025] Binary Spherical Quantization + [CVPR 2026] Leech Spherical Quantization
★ 222lookahead-keys-attention. Causal Attention with Lookahead Keys
★ 28pyserini. Pyserini is a Python toolkit for reproducible information retrieval research with sparse and dense representations.
★ 2.1kDeepRetrieval. [COLM’25] DeepRetrieval — 🔥 Training Search Agent by RLVR with Retrieval Outcome
★ 715Qwen3-Coder. Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team.
★ 17ksmolagents. 🤗 smolagents: a barebones library for agents that think in code.
★ 29kASI-Arch. AlphaGo Moment for Model Architecture Discovery.
★ 1.2ksimple_GRPO. A very simple GRPO implement for reproducing r1-like LLM thinking.
★ 1.7kARMA-Attention. Code implementation of the paper "WAVE: Weighted Autoregressive Varying Gate for Time Series Forecasting" (ICML 2025)
★ 30gpt-oss. gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI
★ 20kCognitiveKernel-Pro. Deep Research Agent CognitiveKernel-Pro from Tencent AI Lab. Paper: https://arxiv.org/pdf/2508.00414
★ 526OLMoE. OLMoE: Open Mixture-of-Experts Language Models
★ 1knautilus_trader. Production-grade Rust-native trading engine with deterministic event-driven architecture
★ 25kWan2.2. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kTimeBase. Python
★ 37HRM. Exploration into the proposed architecture from Sapient Intelligence of Singapore 🇸🇬
★ 78HRM. Hierarchical Reasoning Model Official Release
★ 13kPruLong. Code for the preprint "Cache Me If You Can: How Many KVs Do You Need for Effective Long-Context LMs?"
★ 48SAITS. The official PyTorch implementation of the paper "SAITS: Self-Attention-based Imputation for Time Series". A fast and state-of-the-art (SOTA) deep-learning neural network model for efficient time-series imputation (impute multivariate incomplete time series containing NaN missing data/values with machine learning). https://arxiv.org/abs/2202.08516
★ 513hnet. H-Net: Hierarchical Network with Dynamic Chunking
★ 869esm. Evolutionary Scale Modeling (esm): Pretrained language models for proteins
★ 4.2kopenai-imo-2025-proofs.
★ 484x-transformers-rl. Implementation of a transformer for reinforcement learning using `x-transformers`
★ 72medical-imaging-datasets. A list of Medical imaging datasets.
★ 2.6knanoVLM. The simplest, fastest repository for training/finetuning small-sized VLMs.
★ 5kminirocket. MINIROCKET: A Very Fast (Almost) Deterministic Transform for Time Series Classification
★ 341ChinaTextbook. 所有小初高、大学PDF教材。
★ 76kUCR_Time_Series_Classification_Deep_Learning_Baseline. Fully Convlutional Neural Networks for state-of-the-art time series classification
★ 715TabPFN. ⚡ TabPFN: Foundation Model for Tabular Data ⚡
★ 7.7kFramePack. Lets make video diffusion practical!
★ 17kConvolutional-KANs. This project extends the idea of the innovative architecture of Kolmogorov-Arnold Networks (KAN) to the Convolutional Layers, changing the classic linear transformation of the convolution to learnable non linear activations in each pixel.
★ 922TimeXer. Official implementation for "TimeXer: Empowering Transformers for Time Series Forecasting with Exogenous Variables" (NeurIPS 2024)
★ 505OpenManus. No fortress, purely open ground. OpenManus is Coming.
★ 58kTimePFN. Official repository for "TimePFN: Effective Multivariate Time Series Forecasting with Synthetic Data" (AAAI 2025).
★ 39Q-Flow. Complete Reinforcement Learning Toolkit for Large Language Models!
★ 21KANbeFair. A More Fair and Comprehensive Comparison between KAN and MLP
★ 185OpenRLHF. An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
★ 9.9kKAN-Tutorial. Understanding Kolmogorov-Arnold Networks: A Tutorial Series on KAN using Toy Examples
★ 207deep-cross-attention. Implementation of the proposed DeepCrossAttention by Heddes et al at Google research, in Pytorch
★ 106convkan. Convolutional layer for Kolmogorov-Arnold Network (KAN)
★ 121efficient-kan. An efficient pure-PyTorch implementation of Kolmogorov-Arnold Network (KAN).
★ 4.7kagents-course. This repository contains the Hugging Face Agents Course.
★ 31krational_kat_cu. Python
★ 81e2-tts-pytorch. Implementation of E2-TTS, "Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS", in Pytorch
★ 516rational_activations. Rational Activation Functions - Replacing Padé Activation Units
★ 107kat. [ICLR2025] Kolmogorov-Arnold Transformer
★ 848flash-attention. Fast and memory-efficient exact attention
★ 25kNeural-Experts. This repo contains the official code release of the Neural Experts paper, published in NeurIPS 2024.
★ 14open-thoughts. Fully open data curation for reasoning models
★ 2.3kJanus. Janus-Series: Unified Multimodal Understanding and Generation Models
★ 18kTinyZero. Minimal reproduction of DeepSeek R1-Zero
★ 13kDeepSeek-R1.
★ 92kmicro_diffusion. Official repository for our work on micro-budget training of large-scale diffusion models.
★ 1.6kDeepSeek-V3. Python
★ 104kPRIME. Scalable RL solution for advanced reasoning of language models
★ 1.9kcoconut-pytorch. Implementation of 🥥 Coconut, Chain of Continuous Thought, in Pytorch
★ 184DiT. Official PyTorch Implementation of "Scalable Diffusion Models with Transformers"
★ 8.7klarge_concept_model. Large Concept Models: Language modeling in a sentence representation space
★ 2.4kLVMAE-pytorch. Implementation of the proposed LVMAE, from the paper, Extending Video Masked Autoencoders to 128 frames, in Pytorch
★ 55Protenix. Toward High-Accuracy Open-Source Biomolecular Structure Prediction.
★ 2kllm2vec. Code for 'LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders'
★ 1.7kO1-Journey. O1 Replication Journey
★ 2kdeep-vector-quantization. VQVAEs, GumbelSoftmaxes and friends
★ 654LLM101n. LLM101n: Let's build a Storyteller
★ 38kUSACO. Can Language Models Solve Olympiad Programming?
★ 125TimeSiam. Python
★ 43samformer. Official implementation of SAMformer, a transformer leveraging Sharpness-Aware Minimization and Channel-Wise Attention for Time Series Forecasting.
★ 190rectified-flow-pytorch. Implementation of rectified flow and some of its followup research / improvements in Pytorch
★ 477PDF-Extract-Kit. A Comprehensive Toolkit for High-Quality PDF Content Extraction
★ 9.9kPEER-pytorch. Pytorch implementation of the PEER block from the paper, Mixture of A Million Experts, by Xu Owen He at Deepmind
★ 137SparseTSF. [TPAMI 2026 & ICML 2024 Oral] Official repository of the SparseTSF paper: "SparseTSF: Modeling Long-term Time Series Forecasting with 1k Parameters". This work is developed by the Lab of Professor Weiwei Lin (linww@scut.edu.cn), South China University of Technology; Pengcheng Laboratory.
★ 273pg-is-all-you-need. Policy Gradient is all you need! A step-by-step tutorial for well-known PG methods.
★ 1kKeyBERT. Minimal keyword extraction with BERT
★ 4.2kbuild-nanogpt. Video+code lecture on building nanoGPT from scratch
★ 5.4ksumy. Module for automatic summarization of text documents and HTML pages.
★ 3.7kgranite-tsfm. Foundation Models for Time Series
★ 876c-style. My favorite C programming practices.
★ 2.2kqlib. Qlib is an AI-oriented Quant investment platform that aims to use AI tech to empower Quant Research, from exploring ideas to implementing productions. Qlib supports diverse ML modeling paradigms, including supervised learning, market dynamics modeling, and RL, and is now equipped with https://github.com/microsoft/RD-Agent to automate R&D process.
★ 47kTime-Series-Forecasting-and-Deep-Learning. Resources about time series forecasting and deep learning.
★ 800ToRA. ToRA is a series of Tool-integrated Reasoning LLM Agents designed to solve challenging mathematical reasoning problems by interacting with tools [ICLR'24].
★ 1.1kMAE-LM. [ICLR 2024] Representation Deficiency in Masked Language Modeling
★ 10alphafold3-pytorch. Implementation of Alphafold 3 from Google Deepmind in Pytorch
★ 1.7kLLaVA-pp. 🔥🔥 LLaVA++: Extending LLaVA with Phi-3 and LLaMA-3 (LLaVA LLaMA-3, LLaVA Phi-3)
★ 842chronos-forecasting. Chronos: Pretrained Models for Time Series Forecasting
★ 5.7kLLaVA-HR. [ICLR2025] LLaVA-HR: High-Resolution Large Language-Vision Assistant
★ 249selfcodealign. [NeurIPS'24] SelfCodeAlign: Self-Alignment for Code Generation
★ 324mistral-common. Official inference library for pre-processing of Mistral models
★ 927MGM. Official repo for "Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models"
★ 3.3kLlamaFactory. Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
★ 74ktwenty. The open alternative to Salesforce, designed for AI.
★ 54kllm.c. LLM training in simple, raw C/CUDA
★ 31kOpen-Sora-Plan. This project aim to reproduce Sora (Open AI T2V model), we wish the open source community contribute to this project.
★ 12kguanaco-lora. Instruct-tune LLaMA on consumer hardware
★ 71SWE-agent. SWE-agent takes a GitHub issue and tries to automatically fix it, using your LM of choice. It can also be employed for offensive cybersecurity or competitive coding challenges. [NeurIPS 2024]
★ 20kpytorch-vqvae. Vector Quantized VAEs - PyTorch Implementation
★ 955awesome-mixture-of-experts. A collection of AWESOME things about mixture-of-experts
★ 1.3kControlNet. Let us control diffusion models!
★ 34kgpt-llm-trainer. Jupyter Notebook
★ 4.2kmistral-inference. Official inference library for Mistral models
★ 11khackathon. Python
★ 444OpenHands. 🙌 OpenHands: AI-Driven Development
★ 83kOpen-Sora. Open-Sora: Democratizing Efficient Video Production for All
★ 29kgrok-1. Grok open release
★ 52klaunchables. Collection of notebook guides created by the Brev.dev team!
★ 1.8kannotated-mamba. Annotated version of the Mamba paper
★ 502LOBCAST. LOBCAST is a Python-based open-source framework for stock market trend forecasting using Limit Order Book (LOB) data. 🤖📈
★ 123ml-sigma-reparam. Python
★ 315mixture-of-experts. A Pytorch implementation of Sparsely-Gated Mixture of Experts, for massively increasing the parameter count of language models
★ 866LLMDataHub. A quick guide (especially) for trending instruction finetuning datasets
★ 3.4kCEPE. [ACL 2024] Long-Context Language Modeling with Parallel Encodings
★ 169flow_tutorials. Public tutorials of using Flow Forecast for forecasting and classifying time series data
★ 49flow-forecast. Deep learning PyTorch library for time series forecasting, classification, and anomaly detection (originally for flood forecasting).
★ 2.3kalignment-handbook. Robust recipes to align language models with human and AI preferences
★ 5.7kcoral-pytorch. CORAL and CORN implementations for ordinal regression with deep neural networks.
★ 277UniRepLKNet. [CVPR 2024 & TPAMI 2025] UniRepLKNet
★ 1.1kAutoAWQ. AutoAWQ implements the AWQ algorithm for 4-bit quantization with a 2x speedup during inference. Documentation:
★ 2.4ktimeseries-clustering-vae. Variational Recurrent Autoencoder for timeseries clustering in pytorch
★ 492meta-prompting. Meta-Prompting: Enhancing Language Models with Task-Agnostic Scaffolding
★ 421minbpe. Minimal, clean code for the Byte Pair Encoding (BPE) algorithm commonly used in LLM tokenization.
★ 11ktiktoken. tiktoken is a fast BPE tokeniser for use with OpenAI's models.
★ 19kLongMamba. Some preliminary explorations of Mamba's context scaling.
★ 222axolotl. Go ahead and axolotl questions
★ 12kMedusa. Medusa: Simple Framework for Accelerating LLM Generation with Multiple Decoding Heads
★ 2.8kcontrastors. Train Models Contrastively in Pytorch
★ 800OLMo. Modeling, training, eval, and inference code for OLMo
★ 6.6kMRL. Code repository for the paper - "Matryoshka Representation Learning"
★ 653AutoGPTQ. An easy-to-use LLMs quantization package with user-friendly apis, based on GPTQ algorithm.
★ 5.1kRL-Adventure. Pytorch Implementation of DQN / DDQN / Prioritized replay/ noisy networks/ distributional values/ Rainbow/ hierarchical RL
★ 3.2kDeep-reinforcement-learning-with-pytorch. PyTorch implementation of DQN, AC, ACER, A2C, A3C, PG, DDPG, TRPO, PPO, SAC, TD3 and ....
★ 4.6kpytorch-transformer-ts. Repository of Transformer based PyTorch Time Series Models
★ 322s4. Structured state space sequence models
★ 2.9kMASTER. This is the official code and supplementary materials for our AAAI-2024 paper: MASTER: Market-Guided Stock Transformer for Stock Price Forecasting. MASTER is a stock transformer for stock price forecasting, which models the momentary and cross-time stock correlation and guide feature selection with market information.
★ 517mamba-hf. Implementation of the Mamba SSM with hf_integration.
★ 56mamba. Mamba SSM architecture
★ 19kTime-Series-Library. A Library for Advanced Deep Time Series Models for General Time Series Analysis.
★ 13kTSMOM. Replication of Time Series Momentum strategy by Moskowtiz, Ooi, Pedersen, 2011.
★ 76alpaca_farm. A simulation framework for RLHF and alternatives. Develop your RLHF method without collecting human data.
★ 845x-transformers. A concise but complete full-attention transformer with a set of promising experimental features from various papers
★ 5.9ksoft-moe-pytorch. Implementation of Soft MoE, proposed by Brain's Vision team, in Pytorch
★ 348TimeSeries. Implementation of deep learning models for time series in PyTorch.
★ 395DeepAR-pytorch. Python
★ 274llama-cookbook. Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama model family and using them on various provider services
★ 19kDeepLOB-Deep-Convolutional-Neural-Networks-for-Limit-Order-Books. This jupyter notebook is used to demonstrate our recent work, "DeepLOB: Deep Convolutional Neural Networks for Limit Order Books", published in IEEE Transactions on Singal Processing. We use FI-2010 dataset and present how model architecture is constructed here. The FI-2010 is publicly avilable and interested readers can check out their paper.
★ 604Limit-Orderbook. Limit Order Book for high-frequency trading (HFT) strategies using data science approaches
★ 24PowerfulPromptFT. [NeurIPS 2023 Main Track] This is the repository for the paper titled "Don’t Stop Pretraining? Make Prompt-based Fine-tuning Powerful Learner"
★ 76xgen. Salesforce open-source LLMs with 8k sequence length.
★ 727neuralforecast. Scalable and user friendly neural :brain: forecasting algorithms.
★ 4.2kvector-quantize-pytorch. Vector (and Scalar) Quantization, in Pytorch
★ 4kPatchTST. An offical implementation of PatchTST: "A Time Series is Worth 64 Words: Long-term Forecasting with Transformers." (ICLR 2023) https://arxiv.org/abs/2211.14730
★ 2.7ksd-scripts. Python
★ 7.2ktransformers. 🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
★ 163ktoolformer-pytorch. Implementation of Toolformer, Language Models That Can Use Tools, by MetaAI
★ 2.1ktextgen. Open-source desktop app for local LLMs. Text, vision, tool-calling, OpenAI/Anthropic-compatible API. 100% private.
★ 48kpeft. 🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
★ 21kLLaMA-Adapter. [ICLR 2024] Fine-tuning LLaMA to follow Instructions within 1 Hour and 1.2M Parameters
★ 5.9kGPT-4-LLM. Instruction Tuning with GPT-4
★ 4.3kbaize-chatbot. Let ChatGPT teach your own chatbot in hours with a single GPU!
★ 3.2kFastChat. An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
★ 40kchatgpt-retrieval-plugin. The ChatGPT Retrieval Plugin lets you easily find personal or work documents by asking questions in natural language.
★ 21kdolly. Databricks’ Dolly, a large language model trained on the Databricks Machine Learning Platform
★ 11kopenai-cookbook. Examples and guides for using the OpenAI API
★ 75kai-pdf-chatbot-langchain. AI PDF chatbot agent built with LangChain & LangGraph
★ 17kevals. Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.
★ 19kpyllama. LLaMA: Open and Efficient Foundation Language Models
★ 2.8kself-instruct. Aligning pretrained language models with instruction data generated by themselves.
★ 4.6kstanford_alpaca. Code and documentation to train Stanford's Alpaca models, and generate the data.
★ 30klora. Using Low-rank adaptation to quickly fine-tune diffusion models.
★ 7.5kllama. Inference code for Llama models
★ 60kPrompt-Engineering-Guide. 🐙 Guides, papers, lessons, notebooks and resources for prompt engineering, context engineering, RAG, and AI Agents.
★ 77k