[HN Gopher] Diffsound: Discrete Diffusion Model for Text-to-Soun...
___________________________________________________________________
Diffsound: Discrete Diffusion Model for Text-to-Sound Generation
Author : selimonder
Score : 41 points
Date : 2022-08-08 10:49 UTC (12 hours ago)
(HTM) web link (dongchaoyang.top)
(TXT) w3m dump (dongchaoyang.top)
| notorious-dto wrote:
| Twiddling around trying to get this to work. Looks exciting :)
| [deleted]
| pbronez wrote:
| Very cool. Need to dig around and figure out what the training
| dataset is. Could be a great way to get some sample fodder.
| notorious-dto wrote:
| There seems to be a missing python module called
| "image_synthesis", anyone know more about this?
___________________________________________________________________
(page generated 2022-08-08 23:02 UTC)