Agentβ™₯︎Age
Catalog

Audiolla

Official

by psyb0t Β· Python

Self-hosted MCP server: audio stems, mastering, MIR analysis, DSP, MIDI, speech tools.

io.github.psyb0t/audiolla MCP Server

Self-hosted Model Context Protocol (MCP) server for audio workflows, including audio stems, mastering, MIR analysis, DSP, MIDI, and speech tools. The repository is distributed via Docker and lists support for GPU acceleration (CUDA) and Python-based tooling.

πŸ› οΈ Key Features

  • Audio stems and stem-separation
  • Mastering and MIR (music information retrieval) analysis
  • DSP and speech-related tools (speech-enhancement, speaker-diarization)
  • Audio-to-MIDI and related MIDI tooling
  • Audio embeddings and audio-to-audio generation topics (text-to-audio, text-to-music)
  • Mentions β€œThirty audio engines” and β€œOne port” in the readme excerpt

πŸš€ Use Cases

  • Music production workflows (music-production, mastering, stem-separation)
  • BPM detection and MIR-based analysis (bpm-detection)
  • Converting audio to MIDI (audio-to-midi)
  • Speech enhancement and diarization (speech-enhancement, speaker-diarization)

⚑ Developer Benefits

  • Self-hosted MCP integration (mcp)
  • Docker deployment (docker) with CI/version/license metadata
  • GPU-oriented setup indicated (cuda)
  • Fits LLM agent and audio automation contexts (llm-agents)

⚠️ Limitations

  • The provided data does not enumerate individual MCP tools or counts beyond the β€œthirty audio engines” claim.

Topics

bpm-detectiondemucsdockerfastapillm-agentsmasteringmcpmusic-productionself-hostedstem-separationaudio-embeddingsaudio-to-midicudapythonspeaker-diarizationspeech-enhancementmusicgenstable-audio-opentext-to-audiotext-to-music
Audiolla - agentage MCP Catalog