Rare find

LLaVA-Magvit2. LLaVA combines with Magvit Image tokenizer, training MLLM without an Vision Encoder. Unifying image understanding and generation.

github.com/lucasjinreal/LLaVA-Magvit2

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.