VILA. VILA - a multi-image visual language model with training, inference and evaluation recipe, deployable from cloud to edge (Jetson Orin and laptops)

github.com/dusty-nv/VILA

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.