Rare find

LLM-self-play. Minimal implementation of the Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models paper (ArXiv 20232401.01335)

github.com/thomasgauthier/LLM-self-play

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.