This is your work, valued
VindLU. Python
108TALLFormer. Python
53lipnet-replication. A replication of Google DeepMind's paper:LipNet: End-to-End Sentence-level Lipreading
28jpeg_toolbox_python. extract the dct coefficient of jpeg images. Provide c and python interface.
16DAM. Official code for DAM: Dynamic Adapter Merging for Continual Video QA Learning
15digitalImageForensics. 图像篡改检测
4visual-speaker-authentication. Visual speaker authentication with random prompt texts by a Multi-task CNN Framework
2airnet. airtnet source code
2str-model-analysis. Scene Text Recognition: A model Analysis
2video-cover-picker. pick the most appropriate frame in the video as the cover
2spring-learn. Java
1kyee-common-framework. this framework, whose purpose is to simplify configurations, was developed when I was an intern in kyee
1webcamera-stream-demo. stream webcamera to browser using websocket
1variational-transformer. variational transformer
1