The Patchbay_

ML & Generative Audio Tools, compared

All 19 ml & generative tools in the directory, side by side — language, license, and platforms. Open any tool for install steps and a code example.

ToolLanguageLicensePlatforms
ACE-StepFast open-source music-generation foundation model producing full tracks from a style prompt and lyrics.PythonApache-2.0Linux Windows macOS
AmphionOpen-source toolkit for reproducible audio, music and speech generation (TTS, SVS, VC, TTA).PythonMITLinux
AudioCraft (MusicGen)Meta's PyTorch library for audio generation, home of the MusicGen text-and-melody music model.PythonMITmacOS Windows Linux
AudioGenMeta's text-to-sound model for environmental audio and sound effects, shipped inside AudioCraft. PythonMIT / CC-BY-NC Linux Windows macOS
AudioLDM 2Latent diffusion model generating speech, sound effects and music from text. PythonCC-BY-NC-SA-4.0 Linux Windows macOS
BarkinactiveTransformer text-to-audio model generating speech, music, sound effects and nonverbal sounds.PythonMITmacOS Windows Linux
DDSPDifferentiable Digital Signal Processing: DSP modules (oscillators, filters) usable inside neural nets.PythonApache-2.0macOS Windows Linux
DemucsinactiveState-of-the-art deep-learning music source separation (vocals, drums, bass, other) by Meta.PythonMITmacOS Windows Linux
Fréchet Audio DistanceReference implementation of FAD and CLAP score for evaluating generated audio. PythonMIT Linux Windows macOS
MagentainactiveGoogle research project using ML (TensorFlow) to generate music, art, and drawings.PythonApache-2.0macOS Windows Linux
Magenta RealTimeGoogle's open-weights model for real-time music generation, streaming live on Apple Silicon. Python / C++Apache-2.0 / CC-BY-4.0 macOS Linux Windows
MAGNeTMeta's non-autoregressive masked transformer for text-to-music and text-to-sound — faster than MusicGen. PythonMIT / CC-BY-NC Linux Windows macOS
note-seqinactiveSerializable NoteSequence representation and utilities from Google Magenta for music ML.PythonApache-2.0macOS Windows Linux
RAVEIRCAM's Realtime Audio Variational autoEncoder for fast, high-quality neural audio synthesis.PythonCC-BY-NC-4.0macOS Windows Linux
RiffusioninactiveGenerates music by diffusing spectrogram images — the original open-source project, now unmaintained.PythonMITLinux Windows macOS
SpleeterDeezer's pretrained source-separation library (2/4/5 stems) built on TensorFlow.PythonMITmacOS Windows Linux
Stable Audio OpenOpen text-to-audio diffusion model for generating short instrumentals, loops and sound effects locally.PythonMITmacOS Windows Linux
stable-audio-toolsStability AI's training and fine-tuning toolkit for their audio generation models. PythonMIT Linux Windows macOS
YuEOpen foundation model that turns lyrics into full songs — vocals and backing — across many genres.PythonApache-2.0Linux Windows macOS

← All comparisons