Interested in Machine Learning(Image recognition/retrieval, Face 2D&3D, ASR/TTS/Speaker, OCR). Versed in FFmpeg. Also an iOS Developer
Speaker-Diarization. speaker diarization by uis-rnn and speaker embedding by vgg-speaker-recognition
501FaceConverter. Face swap and 3D alignment from a single image based on PRNet
113MachineLearningDOC. 图像、人脸、OCR、语音相关算法整理
66voicepuppet. Audio driven video synthesis
40Facenet-Caffe. facenet recognition and retrieve by using hnswlib and flask, convert tensorflow model to caffe
29FacePlayer. Face converter in video
17ghostvlad-speaker. An tensorflow implementation of ghostvlad for speaker recognition
15face_recognition_ios. face recognition and retrieve in ios
13AudioKWS. Audio Keyword Search
12ocrDetect. ocr detect using psenet
5ocrlite. text detect and recognition on iOS platform
4detectron2. Python
2GE2E. Speaker Verification
2FaceBasisBuild. Python
1speakerDiarization. Python
1imgFinderServer. image retrieve based on lsh and akaze feature
1pageScanner. C
1