Hanoi, Vietnam

Tuan-Vu Trinh

Advanced
@trinhtuanvubk

Engineer

Diff-VC. Diffusion Model for Voice Conversion

72

yolo8-tracking-counting-speed_estimation. Tracking, counting and speed estimation using yolo8

16

handwritten-ocr. My personal implementation of SVTR model for handwritten OCR

14

KWS-BCResnet. Keyword Spotting using BCResNet and Arcface Loss

13

yolo-ncnn-cpp. everything to infer yolo with ncnn and cpp

10

tflite-yamnet-audio-classification. Yamnet model using tflite_model_maker with esc-50 dataset

9

ASR-based-KWS. QbE Keyword Spotting System based on ASR

8

bark-voice-cloning. Personal customization of some bark-voice-cloning implementations

6

Intent-Classification-and-Entity-Recognition. Intent classification and Entity Recoginition

6

Paraphrasing-Generation-T5. Training paraphasing using huggingface T5

3

OCR-Vietnamese-Text-Generator. A synthetic data generator for text recognition

3

NER-Flair-ONNX. This code is used to convert and infer Flair-ONNX model

2

voicemaker. Js Web for changing voice

2

Hubert-Training. Python

2

finetune-wav2vec2. Python

2

ConvNextV2-Classification. Python

2

Voice-Restoration. Speech Restoration

2

VITS2-TTS. Python

2

tflite-detection-yolov5. Java

2

Wav2Vec2-Serving-API. Python

2

pytorch-ml-utils. Some utility functions / decorators / modules related to Pytorch to help speed up coding

1

OneFormer3D. My personal customization of OneFormer3D

1

TiktokAutoUploader. Automatically Edits Videos and Uploads to Tiktok with CLI, Requests not Selenium.

1

fastapi-course. Python

1

buzz. Buzz transcribes and translates audio offline on your personal computer. Powered by OpenAI's Whisper.

1

Document-Scanner. Simple Document Scanner using Semantic Segmentation

1

interview-question-data-science-.

1

Vietnamese_LLMs. Dự án bao gồm: 1. Xây dựng bộ dữ Instructions Vietnamese (chất lượng, nhiều, và đa dạng). 2.LLM Training, Finetuning, Evaluating & Testing trên Open-source mô hình ngôn ngữ: bloomz, OpenLLaMA, GPT-J pythia etc. 3. Ứng dụng và Giao diện Người dùng (UI)

1

MEDIAR. (NeurIPS 2022 CellSeg Challenge - 1st Winner) Open source code for "MEDIAR: Harmony of Data-Centric and Model-Centric for Multi-Modality Microscopy"

1

jetson-voice. ASR/NLP/TTS deep learning inference library for NVIDIA Jetson using PyTorch and TensorRT

1

PaddleOCR-FastAPI. Python

1

playtorch-study. Study playtorch for mobile

1

wav2vec2-pretraining.

1

bytetrack-yolo-ncnn-cpp. C++

1

computer-vision-notebooks. Examples and tutorials on using SOTA computer vision models and techniques. Learn everything from old-school ResNet, through YOLO and object-detection transformers like DETR, to the latest models like Grounding DINO and SAM.

1

Image-Colorizer-Pix2Pix. Jupyter Notebook

1
36
Apply