DCM_vgg_transformer. Dual cross modality attention audio-visual speech recognition model based on vgg transformer with hybrid CTC/attention architecture using fairseq

github.com/LeeYongHyeok/DCM_vgg_transformer

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.