Rare find

latr. Implementation of LaTr: Layout-aware transformer for scene-text VQA,a novel multimodal architecture for Scene Text Visual Question Answering (STVQA)

github.com/uakarsh/latr

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.