mlx-cuda-distributed-pretraining. mlx-cuda-distributed-pretraining is a new repo forked from n8python/mlx-pretrain - the goal is to have an out of the box solution for MLX enthusiasts for 1) seamless distributed pretraining infrastructure that works Apple Silicon / CUDA devices, 2) easy way to test and bench different optimizers

github.com/arthurcolle/mlx-cuda-distributed-pretraining

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.