Researcher at Shanghai AI Laboratory.
VL-BERT. Code for ICLR 2020 paper "VL-BERT: Pre-training of Generic Visual-Linguistic Representations".
742bottom-up-attention. Bottom-up attention model for image captioning and VQA, based on Faster R-CNN and Visual Genome
2google-access-helper. 谷歌访问助手破解版
1pythia. A modular framework for vision & language multimodal research from Facebook AI Research (FAIR)
1