// AI TOOLS · AUDIONAUT

Audionaut

A free, open-source multitrack audio editor and recorder for macOS, Windows, and Linux — built in C++ on JUCE, with local stem separation and AI-agent control over MCP.

Last verified · 2026-10-02 · by Moe Ameen

What Audionaut is

Audionaut is a free, open-source multitrack audio editor and recorder for macOS, Windows, and Linux, created by a developer who goes by kvoltmer. Its code base went open source on August 15, 2026 under the GPLv3 license — the project's dual-license model also names a commercial license, but its own license notes say that option isn't actually available yet, pending agreements for the GPL/AGPL-linked third-party libraries it uses. It is built in modern C++ on the JUCE framework and runs natively on all three desktop platforms, positioned as a focused, lightweight editor rather than a full digital audio workstation.

The workflow stays close to a regions-to-playlist-to-export model: record or import audio, mark named regions, drop them into a playlist, and export. Underneath are precise cutting and trimming, multi-channel support, flexible I/O routing, time-stretch from roughly ×0.25 to ×4, and an open JSON project format. Export covers WAV, AIFF, MP3 (via LAME), FLAC, and Ogg Vorbis.

Two features set it apart. Stem separation splits a clip into drums, bass, vocals, and other using demucs.cpp, running entirely on your machine so the audio never leaves your computer. And an Auto Edit feature uses on-device musical analysis — beat tracking, onset detection, and segmentation — to cut and arrange clips.

The headline capability is that AI agents can drive it. Audionaut ships a Model Context Protocol (MCP) server with roughly two dozen tools, so Claude and other MCP-capable agents can operate the editor conversationally while it stays open on screen, with each edit landing as a single, visible undo step.

What you can make with it

  • Edited multitrack audio — podcasts, interviews, and multi-channel recordings cut, arranged, and exported
  • Clean voiceovers and narration, trimmed and time-stretched for a target length
  • Separated stems (drums, bass, vocals, other) for remixes, karaoke tracks, or isolating a voice — all processed locally
  • Auto-edited arrangements built from on-device beat and segmentation analysis
  • Agent-driven edits where Claude splits clips, adjusts gain and fades, names regions, and assembles a session over MCP
  • Exports in WAV, AIFF, MP3, FLAC, or Ogg Vorbis, saved alongside an open JSON project file

How Kompozy turns Audionaut output into content

Audionaut is where the audio gets made; [Kompozy](/) is where that audio becomes a published content campaign. Treat them as two halves of one pipeline. Audionaut's real strength is a fast, local, scriptable editor — and because it is agent-drivable over MCP, you can even have Claude do the repetitive cutting, fading, and arranging for you. But once you have a clean episode, voiceover, or track, Audionaut has done its job: it exports a file and nothing more. The step that actually earns reach — turning that recording into clips, video, and posts across every platform — is exactly what Kompozy automates.

Hand Kompozy the finished recording (or the video you captured alongside it) and it fans the ideas into formats an audio editor can't make: [Clipped Shorts](/glossary/clipped-short) cut vertical highlights from the long recording, [Persona Shorts](/glossary/persona-shorts) and HeyGen avatar video build a talking-head trailer, brand-exact [Carousel Posts](/glossary/hyperframes) and Quote Graphics surface the strongest lines, and the transcript becomes a Blog Article of show notes plus an Email Newsletter. A [Persona Brief](/glossary/persona-brief) holds one voice and your banned words across all of it, and [Autopilot](/glossary/autopilot) with a per-post review step schedules and publishes the batch across eight social platforms plus blog and email from a single queue. Audionaut produces the audio; Kompozy produces and ships everything around it.

  1. Record and edit in Audionaut — cut, trim, arrange, and separate stems locally, optionally letting Claude drive the repetitive edits over MCP, then export the finished audio (and keep any video you captured).
  2. Bring the recording into Kompozy as a source, and set a Persona Brief so every output holds one voice and your banned words.
  3. Fan it into short-form: cut Clipped Shorts and Listicle Video from the long recording, and generate a HeyGen persona or avatar trailer, auto-captioned and reframed to 9:16, 1:1, and 16:9.
  4. Spin the ideas into a brand-exact Carousel and Quote Graphics via HyperFrames, plus a Blog Article of show notes and an Email Newsletter from the transcript.
  5. Schedule and publish the whole batch across eight social platforms plus blog and email from one queue with Autopilot and a per-post review pass.

Frequently asked questions

What is Audionaut?

Audionaut is a free, open-source multitrack audio editor and recorder for macOS, Windows, and Linux, built in modern C++ on the JUCE framework. It handles recording, cutting, regions and playlists, multi-channel editing, time-stretch, local stem separation, and broad export — and it can be driven by AI agents over the Model Context Protocol (MCP).

How do I use Claude to edit audio in Audionaut?

With the app open and Node.js 18+ installed, run `claude mcp add audionaut -- npx -y audionaut-mcp`. Claude can then create projects, split and move clips, adjust gain, speed, and fades, name regions, run beat detection, separate stems, and assemble arrangements conversationally, with each edit appearing as a single undo step in the UI.

Is Audionaut free, and what are the limits?

Yes — it went open source on August 15, 2026 under GPLv3 and is free across all three desktop platforms. A commercial license is on the project's roadmap but, per its own license notes, not yet actually available. The honest limits are that it is a young project with a leaner feature set than a mature DAW, and it is an editor only: it produces audio files, not captions, video, or published posts.

How do I turn Audionaut audio into social content?

Audionaut exports the audio but does not publish it. Bring the finished recording into Kompozy to cut Clipped Shorts and Listicle Video, generate a HeyGen persona trailer, build a brand-exact carousel and quote graphics, write show notes as a blog and a newsletter from the transcript, and schedule and publish across eight social platforms plus blog and email from one queue.

Related tools

  • Transcribe.cpp — Open-source C/C++ library that runs 16+ speech-to-text model families locally on your own GPU via the ggml runtime — accurate, offline transcription with no per-minute API bill.
  • ElevenLabs — ElevenLabs is the leading AI voice platform — ultra-realistic text-to-speech, voice cloning, dubbing, sound effects, speech-to-text, and conversational voice agents.
  • HeyGen Video Podcast — HeyGen's app that turns a topic, URL, PDF, or audio track into a two-host video podcast — a shared studio scene, multi-camera cuts, B-roll, and captions, rendered in minutes.
  • Kokoro TTS — An open-weight, 82-million-parameter text-to-speech model that runs high-quality narration locally on a CPU — free, offline, and Apache-2.0 licensed for commercial use.
  • Fish Audio — An AI voice platform for expressive real-time text-to-speech and fast voice cloning — with an open-source model family (Fish Speech) and a hosted flagship, S2.1 Pro, aimed at creators, developers, and enterprises.

← All AI tools · Get started →