Engineer
Diff-VC. Diffusion Model for Voice Conversion
72yolo8-tracking-counting-speed_estimation. Tracking, counting and speed estimation using yolo8
16handwritten-ocr. My personal implementation of SVTR model for handwritten OCR
14KWS-BCResnet. Keyword Spotting using BCResNet and Arcface Loss
13yolo-ncnn-cpp. everything to infer yolo with ncnn and cpp
10tflite-yamnet-audio-classification. Yamnet model using tflite_model_maker with esc-50 dataset
9ASR-based-KWS. QbE Keyword Spotting System based on ASR
8bark-voice-cloning. Personal customization of some bark-voice-cloning implementations
6Intent-Classification-and-Entity-Recognition. Intent classification and Entity Recoginition
6Paraphrasing-Generation-T5. Training paraphasing using huggingface T5
3OCR-Vietnamese-Text-Generator. A synthetic data generator for text recognition
3NER-Flair-ONNX. This code is used to convert and infer Flair-ONNX model
2voicemaker. Js Web for changing voice
2Hubert-Training. Python
2finetune-wav2vec2. Python
2ConvNextV2-Classification. Python
2Voice-Restoration. Speech Restoration
2VITS2-TTS. Python
2tflite-detection-yolov5. Java
2Wav2Vec2-Serving-API. Python
2pytorch-ml-utils. Some utility functions / decorators / modules related to Pytorch to help speed up coding
1OneFormer3D. My personal customization of OneFormer3D
1TiktokAutoUploader. Automatically Edits Videos and Uploads to Tiktok with CLI, Requests not Selenium.
1fastapi-course. Python
1buzz. Buzz transcribes and translates audio offline on your personal computer. Powered by OpenAI's Whisper.
1Document-Scanner. Simple Document Scanner using Semantic Segmentation
1interview-question-data-science-.
1Vietnamese_LLMs. Dự án bao gồm: 1. Xây dựng bộ dữ Instructions Vietnamese (chất lượng, nhiều, và đa dạng). 2.LLM Training, Finetuning, Evaluating & Testing trên Open-source mô hình ngôn ngữ: bloomz, OpenLLaMA, GPT-J pythia etc. 3. Ứng dụng và Giao diện Người dùng (UI)
1MEDIAR. (NeurIPS 2022 CellSeg Challenge - 1st Winner) Open source code for "MEDIAR: Harmony of Data-Centric and Model-Centric for Multi-Modality Microscopy"
1jetson-voice. ASR/NLP/TTS deep learning inference library for NVIDIA Jetson using PyTorch and TensorRT
1PaddleOCR-FastAPI. Python
1playtorch-study. Study playtorch for mobile
1wav2vec2-pretraining.
1bytetrack-yolo-ncnn-cpp. C++
1computer-vision-notebooks. Examples and tutorials on using SOTA computer vision models and techniques. Learn everything from old-school ResNet, through YOLO and object-detection transformers like DETR, to the latest models like Grounding DINO and SAM.
1Image-Colorizer-Pix2Pix. Jupyter Notebook
1