so-vits-svc-4.0-v2. SoftVC VITS Singing Voice Conversion
581so-vits-svc-4.0. SoftVC VITS Singing Voice Conversion
269VQGAN-CLIP. VQGAN+CLIP Colab Notebook with user-friendly interface.
228FakeYou-Tacotron2-Notebook. Tacotron2 Training Notebook for FakeYou.com
169Wav2Lip. This repository contains the codes of "A Lip Sync Expert Is All You Need for Speech to Lip Generation In the Wild", published at ACM Multimedia 2020.
68MaskGCT-Windows. MaskGCT-Windows For Windows Users
64StyleFlow-Windows-10. StyleFlow: Attribute-conditioned Exploration of StyleGAN-generated Images using Conditional Continuous Normalizing Flows
51Audio-Splitter. Audio Splitter provides a user-friendly solution for splitting audio files based on silence detection.
18StyleGAN3-CLIP-ColabNB. Google Colab notebook for NVIDIA's StyleGAN3 and OpenAI's CLIP for a text-based guided image generation.
18TalkNET-colab. NVIDIA's TalkNET - Train and Synthesize on colab
15EverybodyDanceNow-Colab. Motion Retargeting Video Subjects, Modified Colab Version by Justin John
14TTS-TT2. Tacotron 2 - PyTorch implementation with faster-than-realtime inference
11SpeedScribe. High-performance ASR tool using Faster Whisper, supporting custom models, multi-language transcription, and real-time processing feedback.
10video-retalking. [SIGGRAPH Asia 2022] VideoReTalking: Audio-based Lip Synchronization for Talking Head Video Editing In the Wild
10local-llasa-tts-windows. Examples of using the llasa-tts models locally
8PF-AFN-Windows. Windows implementation by Justin John for "Parser-Free Virtual Try-on via Distilling Appearance Flows", CVPR 2021.
8ditto-talkinghead-windows. Ditto: Motion-Space Diffusion for Controllable Realtime Talking Head Synthesis
7LatentSync-windows. Taming Stable Diffusion for Lip Sync!
7efficient-vits-finetuning. Finetuning VITS Efficiently
6tacotron2. Tacotron 2 - PyTorch implementation with faster-than-realtime inference
5csm-windows. A Conversational Speech Generation Model
5Mangio-RVC-Fork-Simple-CLI. Python
5GPT-SoVITS-No-WebUI. No WebUI google colab implementation of GPT-SoVITS
4StyleGANEX. StyleGANEX: StyleGAN-Based Manipulation Beyond Cropped Aligned Faces
43d-photo-inpainting-Windows. [CVPR 2020] 3D Photography using Context-aware Layered Depth Inpainting
4e4s-windows. (CVPR 2023) E4S: Fine-grained Face Swapping via Regional GAN Inversion
4OpenMMD. OpenMMD is an OpenPose-based application that can convert real-person videos to the motion files (.vmd) which directly implement the 3D model (e.g. Miku, Anmicius) animated movies.
4so-vits-svc. 基于vits与softvc的歌声音色转换模型
4ARPAtaco2. Tacotron 2 - PyTorch implementation with faster-than-realtime inference and CMU Pronouncing Dictionary support
3Wav2Lip-Colab. A slightly modified version of Wav2Lip for newbies
2PCAVS-Colab. Pose-Controllable Talking Face Generation by Implicitly Modularized Audio-Visual Representation (CVPR 2021)
1EverybodyDanceNow-Windows10. EverybodyDanceNow implementation for Windows 10
1GPT-SoVITS-v1-No-WebUI-Colab. No WebUI google colab implementation of GPT-SoVITS
1talknet-hifi-gan. HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis
1AMLD2020-Dirty-GANcing. AMLD 2020
1Mano2Smpl-X. We provide a way to fuse MANO parameters into SMPLX.
1instant-ngp-Windows. Instant neural graphics primitives: lightning fast NeRF and more
1wav2lip-onnx-HQ. Full version of wav2lip-onnx including face alignment and face enhancement and more...
1Mocapy. Markerless motion capture: video/webcam -> VRM avatar bones -> BVH
1