Rare find

ocr-vqgan. OCR-VQGAN, a discrete image encoder (tokenizer and detokenizer) for figure images in Paper2Fig100k dataset. Implementation of OCR Perceptual loss for clear text-within-image generation. Fork from VQGAN in CompVis/taming-transformers

github.com/joanrod/ocr-vqgan

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.