ClutchCut Audio Help

Professional macOS DAW — reference guide.

Questions not covered here? Email support@clutchcut.studio.

Getting Started

ClutchCut Audio opens to an empty project. To begin, either import audio (toolbar → Import button or ⌘O), record from your mic (arm a track then hit Record), or generate speech with the TTS button.

  • Import Audio / MIDI — accepts WAV, MP3, M4A, AIFF, FLAC, MID, MIDI. MIDI files create a MIDI track automatically.
  • Add Track — adds a blank audio track. Arm (R button) to record into it.
  • Save / Open projects — projects are saved as .clutchproject JSON bundles. Auto-save runs every 30 seconds.

Clip Editing

All edits support unlimited undo/redo (⌘Z / ⌘⇧Z).

  • Split clip — position playhead over clip, press S or right-click → Split.
  • Trim — drag the left or right edge of a clip to trim. Hold ⌥ for ripple trim.
  • Move clip — drag the clip body. Hold ⌘ to disable snap.
  • Ripple delete — select clip, press for regular delete or toolbar Ripple Delete to close the gap.
  • Multi-select — click then Shift-click or drag a selection box. Nudge all selected with , / ..
  • Snap to grid — enabled by default. BPM field in toolbar controls beat grid. Disable snap by holding ⌘ while dragging.
  • Fade in/out — drag the small triangle handles at the top-left / top-right of a clip.
  • Remove Silence — right-click clip → Remove Silence. Set threshold and minimum gap duration.
  • Insert Silence — toolbar Media group or right-click → Insert Silence. Optionally ripple all tracks.

Music Library

A built-in catalog of royalty-free music — open it from the Insert menu → Music Library (⌘⇧M) or the toolbar. It opens in its own window so you can keep browsing while you edit, and every track is free to use.

  • Browse & filter — search by title, artist, or mood, and narrow by mood, genre, and length. Curated playlists (Vlog & Daily, Epic & Montage, Sleep, Focus & Study, and more) group tracks by use case.
  • Preview — click the play button on any track to audition it before adding.
  • Favorite & Recent — heart tracks you like and jump back to them via the Favorites and Recent filters; save your own playlists to reuse per chapter.
  • Add to timeline — click + Add to drop a track onto your timeline as a new audio track at the playhead.

Keyboard Shortcuts

Playback & Transport

SpacePlay / Pause
LShuttle forward (1× → 2× → 4× → 8×, repeating press cycles)
KStop shuttle / reset to 1×
JShuttle backward; when paused nudges −5 s
← / →Move playhead ±0.1 s
⇧← / ⇧→Move playhead ±1 s
⌥← / ⌥→Move playhead ±5 s
Home / Fn+←Jump to start (t = 0)
End / Fn+→Jump to end of last clip
⌘← / ⌘→Jump to previous / next clip boundary or marker

Clip Operations

SSplit clip at playhead
,Nudge selected clips −0.1 s
.Nudge selected clips +0.1 s
⇧,Nudge selected clips −1 s
⇧.Nudge selected clips +1 s
;Move selected clips so earliest starts at playhead
⌘ZUndo
⌘⇧ZRedo
⌘OOpen project
⌘SSave project
⌘⇧SSave As…
⌘⇧EExport…

Timecode Entry

Click the timecode display in the toolbar to enter a target time. Formats accepted:

  • +5 or -2:30 — relative offset in seconds or MM:SS
  • 1:30 — absolute MM:SS
  • 1:30:00 — HH:MM:SS
  • 00013000 — FCP 8-digit HHMMSSFF at 30 fps

You can also drag up/down on the timecode to scrub at 0.1 s per point.

Mixer & Effects

Click any track header to open the Mixer Inspector panel on the right.

  • Volume / Pan — sliders at the top of the inspector. Volume automation is written to the yellow lane below the track in W/T/L modes.
  • 3-Band EQ — low shelf, mid peak, high shelf with gain ± 12 dB and frequency knobs.
  • Reverb — wet/dry mix and room size.
  • Effects chain — serial Delay and Distortion nodes, add/remove/bypass individually.
  • Pitch & Time — non-destructive pitch shift (±24 semitones) and playback rate (0.5×–2×) via AVAudioUnitTimePitch.
  • Bus routing — (Pro) send the track output to Bus A, B, or C for shared reverb / compression.
  • AUv3 Plugins — (Pro) click Add Effect → AU Browser. Loads third-party Audio Unit v3 plugins with up to 8 parameter sliders.

Auto-Ducking

Toolbar AI Tools → Auto-Duck. Select one or more trigger tracks (e.g. voice-over), set threshold, duck amount, attack/release. Generates volume automation on all non-trigger tracks.

Mastering & Auto-Master

Give any mix a finished, release-ready sound in one place — toolbar AI Tools → Mastering (⌘M). It opens in its own window. Turn on Master Processing, then dial in the halves below.

Character (the enhancement)

Layered on top of loudness, Character adds the “mastered” feel — presence + air (detail), stereo width (space), glue compression and warmth. Pick a preset and set Strength (0–100%):

  • None — loudness only, no tonal change.
  • Natural — transparent polish.
  • Warm — fuller low end + gentle saturation.
  • Bright — presence + air for vocal clarity/detail.
  • Open — wider stereo image / bigger space.
  • Punchy — glue + density for a louder feel.

Loudness target & Auto-Gain

Choose a target — Spotify / YouTube (−14 LUFS), Apple (−16), or Archival (none). Analyze Loudness renders the full mix offline (ITU-R BS.1770-4) and computes the exact gain to hit it; a true-peak limiter catches peaks at −1.0 dB.

Compare (A/B) & export

Click Render Preview, then toggle Original vs Enhanced to hear the difference, with before/after LUFS and peak shown. Turn on Match loudnessto cancel the “louder = better” bias so you judge tone, not level. When it sounds right, click Save Mastered WAV… for a 24-bit delivery file. Auto-master is also applied per-episode inside the Batch Podcast Builder.

Text-to-Speech

Access TTS from the toolbar Media group:

  • TTS button — generates a short clip (≤10,000 chars) placed at the playhead. Kokoro runs on-device with 54 voices; no internet needed for English.
  • Batch Voice Generation — converts multiple .txt / .md / .pdf files to M4A or WAV in the background. Output folder is user-selectable; processing continues while you work and progress shows in the status bar.

Voice source — Local or Cloud (free)

Both single-clip TTS and Batch have a Voice Source toggle: Local — free uses the bundled on-device voices (no internet), and Cloud — freeuses a higher-quality cloud voice at no cost — no credits, no Pro. Cloud needs an internet connection and falls back to the on-device voice automatically if it's ever unavailable.

Engine priority

LanguageDirect Build
Any languageAzure TTS (if API key set in Settings)
EnglishKokoro CoreML → Kokoro CLI → System TTS
CJK (ZH/JP/KR)EdgeTTS (network) → MeloTTS → CosyVoice2 → System TTS
Other 25+ langsEdgeTTS (network) → System TTS

Azure TTS setup

In Settings, enter your Azure Cognitive Services API key and region (e.g. eastus). Azure takes priority over all local engines and supports all 25+ languages with neural voices.

Fix Pronunciation

In the Mixer Inspector → SOURCE TEXT panel, right-click any sentence → Fix Pronunciation. Enter a phonetic spelling and click Regenerate — the app splices in the new audio at frame-level precision.

Transcript Alignment

Import a .srt or .vtt file onto an external podcast track via Mixer Inspector → Import Transcript. For TTS drift correction, use Fix Drift — uses STT binary search to locate and correct misaligned sentences.

Batch Podcast Builder

Turn a folder of narration into finished, publish-ready episodes in one unattended run — Insert menu / toolbar → Batch Podcast Builder (Pro). It opens in its own window; everything runs in the background with a progress bar.

  • Narration — pick individual audio files or a whole folder (each file becomes one episode).
  • Background music — from a local folder or the built-in Music Library; clips are looped and shuffled under each reading.
  • Recipe — Music length (Fit to narration, or Fixed “thick music” for a set length), narration start offset, end fade, shuffle, Auto-fix transcript drift, and Auto-master to your loudness target.
  • Output — WAV / M4A / MP3 with bitrate, and optionally save an editable .clutchaudio project next to every episode so you can tweak it later.
  • Estimate — shows expected output size vs. free disk space before you start.

Each episode is assembled on its own throwaway project and rendered offline, so it never touches your live timeline. Transcript sidecars (.tts.json / .srt / .vtt) are written when a narration has a matching transcript.

AI Feature Setup Guide

Some AI features require external tools installed by the user. All of these apply to the Direct Download build only.

Demucs 6-stem (CLI Stem Separation)

Requires the demucs Python package in a venv at ~/.local/share/demucs-venv or discoverable via PATH.

  1. python3 -m venv ~/dev/tools/demucs-venv
  2. ~/dev/tools/demucs-venv/bin/pip install demucs
  3. In ClutchCut Audio Settings → Stem Separation, set the CLI path if needed.

CoreML Stem Models (on-device, no CLI)

Place converted models in ~/Library/Application Support/ClutchCutAudio/StemModels/:

  • DemucsV2.mlmodelc — 4-stem CoreML Demucs v2
  • OpenUnmix.mlmodelc — 4-stem Open-Unmix

Conversion scripts are in the app's Scripts/ folder.

EdgeTTS (multilingual neural voices, free, requires internet)

  1. Activate your Python venv: source ~/dev/tools/rvc-venv/bin/activate
  2. pip install edge-tts

RAVE Timbre Transfer (Pro)

Place .ts TorchScript RAVE model files in ~/Library/Application Support/ClutchCutAudio/RAVEModels/.

Requires torch, soundfile, scipy in a Python env. Auto-detected from demucs-venv, Homebrew, or system python3.

License Activation (Direct Build)

After purchasing ClutchCut Audio Pro:

  1. You will be redirected to a page showing your CCA-… license key. Copy it (or click the key to copy).
  2. Open ClutchCut Audio. Click the ClutchCut wordmark in the toolbar to open the About sheet.
  3. Paste the key into the license field and click Activate Key.
  4. The app validates the key against our server and unlocks Pro features permanently on that machine.

Your key is tied to your purchase, not to a specific Mac. You can activate it on any Mac you personally own. Contact support@clutchcut.studio if you need to transfer your license.