jepa-llm. Fine-tuning causal language models with an additional JEPA-style representation regularisation loss as well as a plain Hugging Face trainer

github.com/mkurman/jepa-llm

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.