Hi, i finally feel confident enough to announce Audionaut. It's been 3-4 years in the making and my initial motivation was to edit my own multi-channel recordings with the old Sound Designer II workflow (create regions, drop to a playlist, export playlist, done). It got a little bigger now and the most current feature is the agent editing, which is imho very useful since you have the UI which always provides transparency.<p>To try the agent part with Claude Code (app and Node 18+ installed):<p><pre><code> claude mcp add audionaut -- npx -y audionaut-mcp
</code></pre>
Edits land in the open project as one undo step each, and saving stays with you. On macOS the app is sandboxed, so keep projects in ~/Music.<p>Downloads: <a href="https://audionaut.app/download" rel="nofollow">https://audionaut.app/download</a>.<p>any feedback is most appreciated
Congrats on the launch!<p>I tried to built something similar in go and wails but quickly hit roadblock with some critical features that required high time precision.<p>I wonder if I could use it for my use case, do you have any roadmap?
Does it edit music/podcasts automatically, or still needs a human proxy? If it does automatically, how does AI know when to stop?
Congrats on the launch.<p>Some feedback after using it for a few minutes; It's a good effort and the software is useable. I understand there is a lot more to come. it would be nice to have:<p>- ability to import files into tracks from a menu or context menu<p>- keyboard short cut keys for for splitting, etc.<p>- automatic crossfade when moving clips into each other<p>- envelopes<p>- effects<p>- I think Sony Vegas had one of the most intuitive audio editing experiences when working with tracks and clips, maybe adopt what worked from it.
thanks so much for your feedback!<p>- context menu file import is a no brainer, added to my backlog.<p>- keyboard shortcut for split is command+e<p>- automatic crossfade when dragging usually works with the shift modifier key, added to my backlog<p>- envelops: this feature was requested by another user already, so top of my backlog.<p>- effects... yes, of course. i will not implement the in place processing but insert and send effect like it's done regular DAW. so stay tuned.
Task I want but have been too lazy - I have a favorite podcast, that I often fall asleep to. Certain parts with music etc. wake me up. I'd like to chop those parts out as automatically as possible. Can I do that here by e.g. giving some example edits to a few files, and asking to remove similar stuff from other files?
Sounds like something a better model armed with ffmpeg would already be able to do. Run an analysis on loudness along the track, detect when speech begins/ends (with whisper) around loud segments, and compress or cut those parts out
an interesting use case... the analysis could detect rhythmic or not and then edit the rhythmic parts away. i will look into, added to my backlog. thanks for your feedback!
Wow, thanks! I don't have time to look at it now, but it'll go high on my list. I had a quick look though the docs and I see it supports regions. Does it have labels too? Can I import/export labels and/or regions from file?<p>Perhaps an idea to create a docs/features.md to mention what it can do. Might even help LLMs picking up your project. Thanks again, I'm very exited!
Thanks! Regions yes, labels not yet.<p>Good idea on features.md. The manual has it all, but a one-page list is easier to find, for people and for LLMs. I'll add it.
JUCE is open source, but it has a lot of barriers for commercial products. Why did you choose JUCE over iPlug2?
Concrats for finding the confidence. Looks really interesting. Put it on my list of projects to view during my down time. Thanks a ton for sharing. And congrats for getting something out the door.
How does this compare with audiomass?
Main difference is; this a desktop app, whereas audiomass is web.<p>In theory native apps have a higher ceiling for long processing tasks, since they can offload buffers on the disk and reuse - versus having to fit everything in browser memory.
Is it just me but developers started creating video/audio editors after generative AI?<p>Most devs I see creating a video editor or best markdown editor out there with claude/codex.
nice to see a proper open-source multitrack editor that isn't tied to one OS. how does latency hold up with lots of tracks?
is this a DAW? or primarily oriented towards podcasts/non-music. in any case, congrats!
This is a pretty neat looking platform, especially the production pipeline possibilities in it. This possibly should be front and center a bit more. Adding to my stack to try out.
This looks really compelling! Bookmarking to give it a try the next time I have a round of podcasts edits in my todo list.
flatpak please
nice!
[flagged]
[dead]
[dead]
[dead]
[flagged]