Rare find

AudioLM. Turn text into realistic speech and music using AI.

github.com/lucidrains/audiolm-pytorch

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

January 2025
  • Release —2.4.0
  • hyper connect the audio attention models
  • Release —2.3.1
  • needs to be zero centered GP
  • Release —2.3.0
  • new paper claims gradient penalty on fake images going into discrimin…
November 2024
  • Release —2.2.3
  • address https://github.com/lucidrains/audiolm-pytorch/issues/279 again
  • Release —2.2.2
  • update vq
  • Release —2.2.1
  • update vq again
  • Release —2.2.0
  • add value residual learning, proposed in iclr 2025
  • use future annotations
  • Release —2.1.4
  • update vq again
  • Release —2.1.2
  • update vq again
  • Release —2.1.1
  • fix an issue with quantize dropout for residual finite scalar quantiz…
October 2024
  • turn on a new finding that dramatically improves vector quantization
January 2024
  • patch
  • Merge pull request #265 from biendltb/fix/wrong_tensor_assignment
  • fix wrong tensor assignment of the output of attention
  • fix dep
  • patch
  • reports of instability using heinsen formulation
December 2023
  • patch
  • Merge pull request #260 from orrp/main