A fully featured audio diffusion library, for PyTorch. Includes models for unconditional audio generation, text-conditional audio generation, diffusion autoencoding, upsampling, and vocoding. The provided models are waveform-based, however, the U-Net (built using a-unet), DiffusionModel, diffusion method, and diffusion samplers are both generic to any dimension and highly customizable to work on other formats. Note: no pre-trained models are provided here, this library is meant for research purposes.

Features

  • Unconditional Generator
  • Text-Conditional Generator
  • Diffusion Upsampler
  • Diffusion Vocoder
  • Diffusion Autoencoder
  • Inpainting

Project Samples

Project Activity

See All Activity >

License

MIT License

Follow audio-diffusion-pytorch

audio-diffusion-pytorch Web Site

You Might Also Like
MongoDB Atlas runs apps anywhere Icon
MongoDB Atlas runs apps anywhere

Deploy in 115+ regions with the modern database for every enterprise.

MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
Start Free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of audio-diffusion-pytorch!

Additional Project Details

Programming Language

Python

Related Categories

Python AI Music Generators, Python Generative AI, Python Inpainting Tool

Registered

2023-03-28