news.volyx.in

VALL-E: Neural codec language models are zero-shot text to speech synthesizers (valle-demo.github.io)

325 points by georgehill · 1339 days ago · 137 comments on HN

The AI summary for this story hasn't been generated yet — it's produced hourly. Check back soon. Meanwhile, read the discussion on HN.