This is your work, valued
image-to-image-papers. 🦓<->🦒 🌃<->🌆 A collection of image to image papers with code (constantly updating)
★ 1.1kCool-Fashion-Papers. 👔👗🕶️🎩 Cool resources about Fashion + AI! (papers, datasets, workshops, companies, ...) (constantly updating)
★ 633arbitrary-text-to-image-papers. A collection of arbitrary text to image papers with code (constantly updating)
★ 211Clothes-3D. clothes research in 3D
★ 182metrics. IS, FID score Pytorch and TF implementation, TF implementation is a wrapper of the official ones.
★ 129deepfashion2-kps-agg-finetune. 1st place solution (Team StylingAI Inc. & PKU AIIC) for CVPR 2020 DeepFashion2 Clothes Landmark Detection Track. Aggregation and Finetuning for Clothes Landmark Detection
★ 103Awesome-BEV-Papers.
★ 19alias-free-gan-explanation. Trying to understand alias-free-gan.
★ 14Fast-Fourier-Transform. A C++ Implementation of Fast Fourier Transform (Project of Digital Signal Processing course)
★ 14RBM-DBN-theano-DL4J. Train a DBN to classify a set of test data similar to MNIST, Using DL4J & theano (Project of Pattern Recognition course)
★ 11Palm-Live-Detection. Palm Live Detection (Project of Digital Image Processing Course)
★ 9Resources. A curated list of resources in Software Engineering related field
★ 8bigdatacup2021. 2nd Place of BigData Cup 2021 Track1
★ 7MCMC. An introduction of MCMC, implementation of TAP, AIS, RTS algorithm to estimate the partition function of an RBM (Project of Stochastic Process course)
★ 6pysc2-minigame-ai. Rule based AI for StarCraft 2 Minigames
★ 5texture-synthesis-papers. about texture synthesis papers
★ 4SinGAN-pytorch. pytorch reimplementation of SinGAN
★ 4stylegan2-1. StyleGAN2 - Official TensorFlow Implementation with practical improvements
★ 3gem5-NVP-DFS. Dynamic Frequency Selection (Project of Modern Computer Architecture course)
★ 2DTMF. An implementation of DTMF detection and Goertzel Algorithm (Project of Digital Signal Processing course)
★ 2GAN-image-synthesis-papers-collection. about all kinds of image synthesis with GANs
★ 2DeepSeek-Coder. DeepSeek Coder: Let the Code Write Itself
★ 1OpenOccupancy. OpenOccupancy: A Large Scale Benchmark for Surrounding Semantic Occupancy Perception
★ 1gpr. Gaussian Process Regression: A practical overview (Project of Stochastic Process course)
★ 13-Classic-Operating-System-Problems. Implementation of 3 classic Operating System problems (Project of Operating System course)
★ 1libigl. Simple C++ geometry processing library.
★ 1FashionAI-1. FashionAI Global Challenge—Attributes Recognition of Apparel based on PyTorch.
★ 1Leetcode-Python. My solutions of Leetcode using Python
★ 1EffectiveTensorflow. TensorFlow tutorials and best practices.
★ 1IJCAI17_Tianchi_Rank4. Python
★ 1Kimi-K3. Open Frontier Intelligence
★ 7.5kflashlib. Fast and memory-efficient classical machine learning operators
★ 551sparser-faster-llms. Cuda kernels for leveraging LLM sparsity to improve throughput and decrease the memory requirements during inference and training.
★ 255autoresearch. AI agents running research on single-GPU nanochat training automatically
★ 92kMiMo-Code. MiMo Code: Where Models and Agents Co-Evolve
★ 13kdiffusion-pipe. A pipeline parallel training script for diffusion models.
★ 2kBernini. Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.
★ 1.2kPiD. PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion
★ 1kvggt-omega. [CVPR 2026 Oral] VGGT Omega
★ 3.8kseedance-2.0. Comprehensive production pipeline for quad-modal AI filmmaking with Seedance 2.0
★ 5.6kmodded-nanogpt. NanoGPT (124M) in 90 seconds
★ 5.6kpytorch-memory-fix. Two environment variables that fix PyTorch/glibc memory creep on Linux. Zero code changes. Zero performance cost.
★ 13pm-skills. PM Skills Marketplace: 100+ agentic skills, commands, and plugins — from discovery to strategy, execution, launch, and growth.
★ 25kms-swift. Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
★ 15kqust. HTML
★ 178waoowaoo. 首家工业级全流程 AI 影视生产平台。Industry-first professional AI Agent platform for controllable film & video production. From shorts to live-action with Hollywood-standard workflows.
★ 13kopencode. The open source coding agent.
★ 191kawesome-nanobanana-pro. 🚀 An awesome list of curated Nano Banana pro prompts and examples. Your go-to resource for mastering prompt engineering and exploring the creative potential of the Nano banana pro(Nano banana 2) AI image model.
★ 10kFastGen. NVIDIA FastGen: Fast Generation from Diffusion Models
★ 899openclaw. Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
★ 385kx-algorithm. Algorithm powering the For You feed on X
★ 27kEngram. Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models
★ 4.6kGLM-Image. GLM-Image: Auto-regressive for Dense-knowledge and High-fidelity Image Generation.
★ 1kAwesome-RL-for-Video-Generation. A curated list of papers on reinforcement learning for video generation
★ 575TurboDiffusion. TurboDiffusion: 100–200× Acceleration for Video Diffusion Models
★ 3.6kVTP. [ECCV 2026] Towards Scalable Pre-training of Visual Tokenizers for Generation
★ 496Wan-Move. [NeurIPS 2025] Wan-Move: Motion-controllable Video Generation via Latent Trajectory Guidance
★ 649FinRL. FinRL®: Financial Reinforcement Learning. 🔥
★ 16kbitcoin-arbitrage. Bitcoin arbitrage - opportunity detector
★ 2.6kgated_attention. The official implementation for [NeurIPS2025 Oral] Gated Attention for Large Language Models: Non-linearity, Sparsity, and Attention-Sink-Free
★ 973Z-Image. Python
★ 12kRepTok. [ICLR 2026] Adapting Self-Supervised Representations as a Latent Space for Efficient Generation
★ 60GenExam. [ICML 2026] GenExam: A Multidisciplinary Text-to-Image Exam
★ 70MiMo-Embodied. MiMo-Embodied
★ 399sam-3d-objects. SAM 3D Objects
★ 7.2ksam3. The repository provides code for running inference and finetuning with the Meta Segment Anything Model 3 (SAM 3), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 11kflame. 🔥 A minimal training framework for scaling FLA models
★ 408JiT. PyTorch implementation of JiT https://arxiv.org/abs/2511.13720
★ 2.5kVideoREPA. [NeurIPS 2025] VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models
★ 198lejepa. Python
★ 1.3knvg. [ICLR 2026] Code for our paper "Next Visual Granularity Generation".
★ 54flash-moba. C++
★ 252flashrnn. FlashRNN - Fast RNN Kernels with I/O Awareness
★ 188ml-atoken. Jupyter Notebook
★ 145omnilingual-asr. Omnilingual ASR Open-Source Multilingual SpeechRecognition for 1600+ Languages
★ 2.9kInfinityStar. [NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation
★ 774cat-transformer. Compress and Attend Transformers (CATs) 😸
★ 23PyPortfolioOpt. Financial portfolio optimization in python, including classical efficient frontier, Black-Litterman, Hierarchical Risk Parity
★ 5.9kadata. 免费开源A股量化交易数据库; 专注A股,专注量化,向阳而生; 开放、纯净、持续、为Ai(爱)发电。为个人量化交易而生,保卫3000点,珍惜底部机会......【股票数据,股票行情数据,股票量化数据,股票交易数据,k线行情数据,股票概念数据,股票数据接口,行情数据接口,量化交易数据】【多数据源融合,动态设置代理,保障数据高可用性】
★ 5kDanceGRPO. An official implementation of DanceGRPO: Unleashing GRPO on Visual Generation
★ 1.6kKimi-Linear.
★ 1.5kAwesome-Visual-Generation-Alignment-Survey. A survey for visual generation alignment
★ 144tinker-cookbook. Post-training with Tinker
★ 4kLongCat-Video. Python
★ 5.7kLongLive. Long Video Gen Infrastructure
★ 2.5kpico-banana-400k. Python
★ 1.8kDreamOmni2. This project is the official implementation of 'DreamOmni2: Multimodal Instruction-based Editing and Generation (CVPR2026 Highlight)''
★ 2kDeepSeek-OCR. Contexts Optical Compression
★ 24ktrading-signals. Technical indicators to run technical analysis with JavaScript & TypeScript. 📈
★ 973dexbotic. Dexbotic: Open-Source Vision-Language-Action Toolbox
★ 1.3kKandinsky-2. Kandinsky 2 — multilingual text2image latent diffusion model
★ 2.8kedm2. EDM2 and Autoguidance -- Official PyTorch implementation
★ 848RAE. Official PyTorch Implementation of "Diffusion Transformers with Representation Autoencoders"
★ 2knanochat. The best ChatGPT that $100 can buy.
★ 57kTinyRecursiveModels. Python
★ 6.6klmdeploy. LMDeploy is a toolkit for compressing, deploying, and serving LLMs.
★ 8kPromptEnhancer. [CVPR 2026] PromptEnhancer is a prompt-rewriting tool, refining prompts into clearer, structured versions for better image generation.
★ 3.7kBaichuan-Omni-1.5. Python
★ 194HuMo. HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning
★ 1.3kRobusTok. Image Tokenizer Needs Post-Training
★ 24Awesome-Nano-Banana-images. A curated collection of fun and creative examples generated with Nano Banana & Nano Banana Pro🍌, Gemini-2.5-flash-image based model. We also release Nano-consistent-150K openly to support the community's development of image generation and unified models(click to website to see our blog)
★ 23kWaver. Industry-level video foundation model for unified Text-to-Video (T2V) and Image-to-Video (I2V) generation.
★ 950FastVGGT. [ICLR 2026] FastVGGT: Fast Visual Geometry Transformer
★ 806USO. [CVPR 2026] 🔥🔥 Official Repo of USO: Unified Style and Subject-Driven Generation via Disentangled and Reward Learning
★ 1.2kRLinf. RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI
★ 4.3kHQ-Edit. [ICLR 2025] HQ-Edit: A High-Quality and High-Coverage Dataset for General Image Editing
★ 114ImageFolder. High-performance Image Tokenizers for VAR and AR
★ 307AttentionEngine. Python
★ 123dinov3. Reference PyTorch implementation and models for DINOv3
★ 11kVeOmni. VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo
★ 2.1kDreamVVT. DreamVVT: Mastering Realistic Video Virtual Try-On in the Wild via a Stage-Wise Diffusion Transformer Framework
★ 153gpt-oss. gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI
★ 20kQwen-Image. Qwen-Image is a powerful image generation foundation model capable of complex text rendering and precise image editing.
★ 8.2kDispLoss. Python
★ 111Wan2.2. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kPi3. [ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning
★ 2.1khnet. H-Net: Hierarchical Network with Dynamic Chunking
★ 869Kimi-K2. Kimi K2 is the large language model series developed by Moonshot AI team
★ 11kREG. [NeurIPS 2025 Oral] Representation Entanglement for Generation: Training Diffusion Transformers Is Much Easier Than You Think
★ 274trae-agent. Trae Agent is an LLM-based agent for general purpose software engineering tasks.
★ 12kperception_models. State-of-the-art Image & Video CLIP, Multimodal Large Language Models, and More!
★ 2.3kHeavyBall. Efficient optimizers
★ 336lightning-thunder. PyTorch compiler that accelerates training and inference. Get built-in optimizations for performance, memory, parallelism, and easily write your own.
★ 1.5kREPA-E. [ICCV 2025] Official implementation of the paper: REPA-E: Unlocking VAE for End-to-End Tuning of Latent Diffusion Transformers
★ 512LoRAEdit. We achieves high-quality first-frame guided video editing given a reference image, while maintaining flexibility for incorporating additional reference conditions.
★ 336vjepa2. PyTorch code and models for VJEPA2 self-supervised learning from video.
★ 4.4kSeedVR. Repo for SeedVR2 (ICLR2026) & SeedVR (CVPR2025 Highlight)
★ 1.3kMiniMax-M1. MiniMax-M1, the world's first open-weight, large-scale hybrid-attention reasoning model.
★ 3.2kflux. Official inference repo for FLUX.1 models
★ 26kOmniEdit. Official Repo for Paper "OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision" [ICLR2025]
★ 144zoology. Understand and test language model architectures on synthetic tasks.
★ 280Dream. Dream 7B, a large diffusion language model
★ 1.3kD-AR. the official repo for "D-AR: Diffusion via Autoregressive Models"
★ 138linear-attention-and-beyond-slides. TeX
★ 120BLIP3o. Official implementation of BLIP3o-Series
★ 1.7kLaCT. Code release for paper "Test-Time Training Done Right"
★ 500DetailFlow. 🔥 Official impl. of "DetailFlow: 1D Coarse-to-Fine Autoregressive Image Generation via Next-Detail Prediction"
★ 171RF-Solver-Edit. [🚀ICML 2025] "Taming Rectified Flow for Inversion and Editing" Using FLUX and HunyuanVideo for image and video editing!
★ 638Bagel. Open-source unified multimodal model
★ 6.1k3DDFA-V3. The official implementation of 3DDFA_V3 in CVPR2024 (Highlight).
★ 388deer-flow. An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, subagents and message gateway, it handles different levels of tasks that could take minutes to hours.
★ 78kgenesis-world. Simulation platform for general-purpose robotics & embodied AI learning.
★ 30kSuperEdit. [ICCV 2025] Code & Data for: SuperEdit - Rectifying and Facilitating Supervision for Instruction-Based Image Editing
★ 165Awesome-FLUX-DiT. A collection of diffusion models based on FLUX/DiT for image/video generation, editing, reconstruction, inpainting .etc.
★ 86ICEdit. [NeurIPS 2025] Image editing is worth a single LoRA! 0.1% training data for fantastic image editing! Surpasses GPT-4o in ID persistence~ MoE ckpt released! Only 4GB VRAM is enough to run!
★ 2.1kUniAnimate-DiT. UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
★ 854Step1X-Edit. A SOTA open-source image editing model, which aims to provide comparable performance against the closed-source models like GPT-4o and Gemini 2 Flash.
★ 2.2kDreamO. [SIGGRAPH Asia 2025] DreamO: A Unified Framework for Image Customization
★ 1.7kREPA. [ICLR'25 Oral] Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think
★ 1.7kHiDream-I1. Python
★ 2.5kMAGI-1. MAGI-1: Autoregressive Video Generation at Scale
★ 3.8kMiniMax-01. The official repo of MiniMax-Text-01 and MiniMax-VL-01, large-language-model & vision-language-model based on Linear Attention
★ 3.4kFramePack. Lets make video diffusion practical!
★ 17kEasyControl. Implementation of "EasyControl: Adding Efficient and Flexible Control for Diffusion Transformer"(ICCV2025)
★ 1.7kAnyEdit. 【CVPR 2025 Oral】Official Repo for Paper "AnyEdit: Mastering Unified High-Quality Image Editing for Any Idea"
★ 226UnifyEdit. Tuning-Free Image Editing with Fidelity and Editability via Unified Latent Diffusion Model
★ 13ttt-video-dit. Official PyTorch implementation of One-Minute Video Generation with Test-Time Training
★ 2.4kcuda-python. CUDA Python: Performance meets Productivity
★ 3.3kSana. SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
★ 8.6kLightningDiT. [CVPR 2025 Oral] Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models
★ 1.5kBlaGPT. Experimental playground for benchmarking language model (LM) architectures, layers, and tricks on smaller datasets. Designed for flexible experimentation and exploration.
★ 113diffusion-4k. [CVPR 2025] Diffusion-4K: Ultra-High-Resolution Image Synthesis with Latent Diffusion Models
★ 364Wan2.1. Wan: Open and Advanced Large-Scale Video Generative Models
★ 17kInfiniteYou. 🔥 [ICCV 2025 Highlight] InfiniteYou: Flexible Photo Recrafting While Preserving Your Identity
★ 2.7knative-sparse-attention. 🐳 Efficient Triton implementations for "Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention"
★ 1kvggt. [CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer
★ 14kEquiTriton. EquiTriton is a project that seeks to implement high-performance kernels for commonly used building blocks in equivariant neural networks, enabling compute efficient training and inference.
★ 74DeepGEMM. DeepGEMM: clean and efficient BLAS kernel library on GPU
★ 7.6kOpenManus. No fortress, purely open ground. OpenManus is Coming.
★ 58kktransformers. A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations
★ 19kFLASHNN. Python
★ 107triton-index. Cataloging released Triton kernels.
★ 310attorch. A subset of PyTorch's neural network modules, written in Python using OpenAI's Triton.
★ 604PAR. [CVPR2025 Highlight] PAR: Parallelized Autoregressive Visual Generation. https://yuqingwang1029.github.io/PAR-project
★ 186xAR. This repository includes the official implementation of our paper "Beyond Next-Token: Next-X Prediction for Autoregressive Visual Generation"
★ 251DoRA. [ICML2024 (Oral)] Official PyTorch implementation of DoRA: Weight-Decomposed Low-Rank Adaptation
★ 986llm-foundry. LLM training code for Databricks foundation models
★ 4.4kDeepEP. DeepEP: an efficient expert-parallel communication library
★ 9.9kDeepClaude. Unleash Next-Level AI! 🚀 💻 Code Generation: DeepSeek r1 + Claude 3.7 Sonnet - Unparalleled Performance! 📝 Content Creation: DeepSeek r1 + Gemini 2.5 Pro - Superior Quality! 🔌 OpenAI-Compatible. 🌊 Streaming & Non-Streaming Support. ✨ Experience the Future of AI – Today! Click to Try Now! ✨
★ 2.9kMoonlight. Muon is Scalable for LLM Training
★ 1.5kFlashMLA. FlashMLA: Efficient Multi-head Latent Attention Kernels
★ 13kTime-Series-Works-Conferences. Time-Series Work Summary in CS Top Conferences (NIPS, ICML, ICLR, KDD, AAAI, WWW, IJCAI, CIKM, ICDM, ICDE, etc.)
★ 967GatedDeltaNet. [ICLR 2025] Official PyTorch Implementation of Gated Delta Networks: Improving Mamba2 with Delta Rule
★ 636COAT. [ICLR 2025] COAT: Compressing Optimizer States and Activation for Memory-Efficient FP8 Training
★ 263Step-Video-T2V. Python
★ 3.2kQwen3. Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.
★ 27kMoBA. MoBA: Mixture of Block Attention for Long-Context LLMs
★ 2.2ktritonbench. Tritonbench is a collection of PyTorch custom operators with example inputs to measure their performance.
★ 363imagesize_py. Python
★ 253unsloth. Unsloth is a local UI for training and running Kimi K3, Gemma 4, Qwen3.6, DeepSeek, GLM and other models.
★ 69kTinyZero. Minimal reproduction of DeepSeek R1-Zero
★ 13kFlagAttention. A collection of memory efficient attention operators implemented in the Triton language.
★ 304extension-cpp. C++ extensions in PyTorch
★ 1.2kFlagGems. FlagGems is an operator library for large language models implemented in the Triton Language.
★ 1.1kLiger-Kernel. Efficient Triton Kernels for LLM Training
★ 6.5kDailyArXiv. Daily ArXiv Papers.
★ 446Qwen3-VL. Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.
★ 20kHunyuan3D-2. High-Resolution 3D Assets Generation with Large Scale Hunyuan3D Diffusion Models.
★ 14kKimi-k1.5.
★ 3.5kDeepSeek-R1.
★ 92kcosmos. NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.
★ 11kxDiT. xDiT: A Scalable Inference Engine for Diffusion Transformers (DiTs) with Massive Parallelism
★ 2.7kflash-linear-attention. 🚀 Efficient implementations for emerging model architectures
★ 5.5kmicro_diffusion. Official repository for our work on micro-budget training of large-scale diffusion models.
★ 1.6kDeepSeek-V3. Python
★ 104kVideoVAEPlus. [ICCV 2025] VideoVAE+: Large Motion Video Autoencoding with Cross-modal Video VAE
★ 410Infinity. [CVPR 2025 Oral]Infinity ∞ : Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
★ 1.6kRuyi-Models. Python
★ 518LlamaFactory. Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
★ 74kVAR. [NeurIPS 2024 Best Paper Award][GPT beats diffusion🔥] [scaling laws in visual generation📈] Official impl. of "Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction". An *ultra-simple, user-friendly yet state-of-the-art* codebase for autoregressive image generation!
★ 8.7kMARS. The official implementation of MARS: Unleashing the Power of Variance Reduction for Training Large Models
★ 723ai-toolkit. The ultimate training toolkit for finetuning diffusion models
★ 11kJanus. Janus-Series: Unified Multimodal Understanding and Generation Models
★ 18kaddit. Python
★ 390RAG-Diffusion. [ICCV 2025] Region-Aware Text-to-Image Generation via Hard Binding and Soft Refinement 🔥
★ 622alphafold3. AlphaFold 3 inference pipeline.
★ 8.4kEmbodiedCity. Python
★ 300mirage. Mirage Persistent Kernel: Compiling LLMs into a MegaKernel
★ 2.4kxformers. Hackable and optimized Transformers building blocks, supporting a composable construction.
★ 11kAwesome-Mamba-Papers. Awesome Papers related to Mamba.
★ 1.4ksapiens. High-resolution models for human tasks.
★ 5.4kControlNeXt. Controllable video and image Generation, SVD, Animate Anyone, ControlNet, ControlNeXt, LoRA
★ 1.6kmattermost. Mattermost is an open source platform for secure collaboration across the entire software development lifecycle..
★ 39kIDM-VTON. [ECCV2024] IDM-VTON : Improving Diffusion Models for Authentic Virtual Try-on in the Wild
★ 5.1ksam2. The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
★ 20kmem0. Universal memory layer for AI Agents
★ 62kmplfinance. Financial Markets Data Visualization using Matplotlib
★ 4.4kiTransformer. Official implementation for "iTransformer: Inverted Transformers Are Effective for Time Series Forecasting" (ICLR 2024 Spotlight)
★ 2.2kX-Pose. [ECCV 2024] Official implementation of the paper "X-Pose: Detecting Any Keypoints"
★ 815LivePortrait. Bring portraits to life!
★ 19k