A fully featured audio diffusion library, for PyTorch. Includes models for unconditional audio generation, text-conditional audio generation, diffusion autoencoding, upsampling, and vocoding. The provided models are waveform-based, however, the U-Net (built using a-unet), DiffusionModel, diffusion method, and diffusion samplers are both generic to any dimension and highly customizable to work on other formats. Note: no pre-trained models are provided here, this library is meant for research purposes.

Features

  • Unconditional Generator
  • Text-Conditional Generator
  • Diffusion Upsampler
  • Diffusion Vocoder
  • Diffusion Autoencoder
  • Inpainting

Project Samples

Project Activity

See All Activity >

License

MIT License

Follow audio-diffusion-pytorch

audio-diffusion-pytorch Web Site

Other Useful Business Software
MongoDB Atlas | Run databases anywhere Icon
MongoDB Atlas | Run databases anywhere

Ensure the availability of your data with coverage across AWS, Azure, and GCP on MongoDB Atlas—the multi-cloud database for every enterprise.

MongoDB Atlas allows you to build and run modern applications across 125+ cloud regions, spanning AWS, Azure, and Google Cloud. Its multi-cloud clusters enable seamless data distribution and automated failover between cloud providers, ensuring high availability and flexibility without added complexity.
Learn More
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of audio-diffusion-pytorch!

Additional Project Details

Programming Language

Python

Related Categories

Python AI Music Generators, Python Generative AI, Python Inpainting Tool

Registered

2023-03-28