ViT-Assembler. This repository presents a self-supervised learning approach utilizing a Vision Transformer (ViT) Encoder to solve a jigsaw puzzle as a pretext task.

github.com/diegobonilla98/ViT-Assembler

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.