Skip to content
#

audio-embeddings

Here are 17 public repositories matching this topic...

Self-hosted audio API in one Docker container. Stem separation, mastering, MIR (BPM/key/chords), audio→MIDI, text-to-music + SFX (MusicGen/Stable Audio/Riffusion/AudioLDM2), diarization, speech enhancement, restoration, embeddings + zero-shot tagging, workflows, async + webhooks. REST + MCP. CPU + CUDA.

  • Updated Aug 1, 2026
  • Python

Self-hosted semantic music search and audio tagging over a catalog on Backblaze B2. Essentia extracts BPM, key, genre, and mood; CLAP embeddings power find-similar and plain-English text-to-audio search. All local OSS models — B2 is the only credential.

  • Updated Sep 11, 2026
  • TypeScript

Music analysis for DJs — interpretable DSP features (BPM, four-on-floor, spectral balance, key) combined with MERT audio embeddings, UMAP genre clustering, Engine DJ playlist import and similar-track suggestions.

  • Updated Sep 9, 2026
  • Python

Add this topic to your repo

To associate your repository with the audio-embeddings topic, visit your repo's landing page and select "manage topics."

Learn more