TECHNOLOGY

Audionaut Lets AI Agents Like Claude Edit Multitrack Audio Projects Directly

A new free, open-source desktop audio editor called Audionaut exposes its editing engine over MCP, letting AI assistants cut, arrange, and mix multitrack sessions alongside human users.

Audio engineer's desktop showing multitrack waveform editing software with a laptop running an AI chat assistant nearbyTECHNOLOGY

Image: Audionaut · Uploaded by IntraGoals — usage rights confirmed

A developer has released Audionaut, a free and open-source multitrack audio editor built to be operated not just by human sound engineers but by AI agents as well. The project, hosted on GitHub by user kvoltmer, is designed for music production, podcast editing, and multitrack recording, offering precise cutting, per-track playlists, and flexible multi-channel support without the overhead of a full digital audio workstation.

What sets Audionaut apart is its support for the Model Context Protocol, or MCP, which allows AI assistants such as Claude to drive the editing process directly. According to the project's documentation, a demo shows Claude cutting a song every sixteen bars, splitting the resulting clips across two tracks, closing the gaps, and applying crossfades and a fade-out, while a human user simultaneously runs the app's Auto Edit and stem-separation features.

To use Audionaut with Claude, users install the application alongside Node.js 18 or later and register the MCP server with a single command. Each AI-driven edit is logged as a single undo step, and on macOS, project files need to be kept in the Music folder for the sandboxed app to access them.

Audionaut is written in modern C++ on the JUCE framework and runs natively on Windows, macOS, and Linux. It is dual-licensed under GPL v3 and a commercial license. The software's analysis features, including beat tracking and onset detection, rely on a static build of the Essentia audio-analysis library, while stem separation is powered by demucs.cpp, a C++ port of Meta's Demucs model; the underlying model weights are downloaded on first use rather than bundled with the repository.

Beyond the graphical app, the project ships a command-line tool called audionaut-cli, intended to give scripts, continuous-integration pipelines, and AI agents headless access to Audionaut project files without needing a GUI or audio hardware. The tool supports a wide range of operations, from creating and importing audio to splitting clips, setting fades, separating stems, and exporting final mixes, with every command able to output a single machine-readable JSON result for easy automation. A typical automated workflow, the documentation suggests, might involve an agent creating a project, importing audio, analyzing it, auto-editing or assembling segments, and exporting the result, checking for success at each step.

The tool also includes an opt-in anonymous usage-reporting feature, which can be disabled entirely via an environment variable for use in continuous-integration environments or scripts. The project welcomes external contributions under a standard Contributor License Agreement and accepts sponsorship to help fund continued development.

Audionaut was highlighted on Hacker News, where it was shared as a tool allowing AI agents to perform multitrack editing tasks previously reserved for human engineers using full-featured DAWs.

Sources and further readingGitHub - kvoltmer/Audionaut: Audionaut professional audio editing and audio recording ↗
ABOUT THE DESK

IntraGoals News Desk

IntraGoals reports on important changes in technology and work. We check each story for clear writing, trusted sources and useful information before it is published.

KEEP READING

Latest from IntraGoals.

All latest stories ↗