Vocal Slice — cut audio by selecting text

Vocal Slice

4 min read Original article ↗

Word-accurate voice slicing for Mac & Windows

Highlight a phrase and the waveform jumps straight to it — export it as a named, ready-to-deliver clip. Your audio never leaves your device.

Every feature, free for seven days — you don't need an account or a card to start. Also for macOS. Also for Windows.

The Vocal Slice app: a phrase is selected in the transcript and the waveform below highlights exactly that span of audio, showing its start, end and duration.

How it works

Three steps, take to delivery.

Load a recording, select the words you want, export the clips. It all runs locally, so transcribing still works with the wifi off.

1 · Load your audio

Open a WAV, MP3, FLAC, M4A, AAC or OGG file, of any length. Vocal Slice transcribes it locally with Whisper, producing timestamps down to the individual word.

2 · Select the words

Highlight a phrase in the transcript. The waveform zooms to exactly that region, with draggable start and end handles for frame-accurate trimming.

3 · Export named files

Preview on loop, then export. A WAV source is cut byte-perfect, so what you hand over is the original audio rather than a re-encode of it, and every file is named from your own template.

Features

Built for people who cut voice recordings all day.

Podcasters, video editors, voiceover artists and content creators — anyone who works through hours of recordings to find the few parts that matter.

Named to your convention

Every slice is written from your own template — {source}_{index}_{slug}, timestamps, or whatever your project expects. A session's worth of clips comes out already matching the convention you deliver in, rather than needing an evening of renaming first.

Source-quality exports

WAV files are sliced losslessly at the byte level, so a slice carries exactly the source's channel count, sample rate and bit depth — stereo and multichannel come through untouched. Other formats decode to 24-bit WAV.

Re-trim without starting over

Open any slice back up in the list and move its boundaries, either by dragging its own waveform or by typing exact start and end times. The transcription is still loaded, so a second pass costs you nothing but the adjustment itself.

Speaks your language

English and multilingual Whisper models, from fast and light to slow and accurate, with languages from Spanish and German through to Japanese, Arabic and Hindi.

GPU accelerated

Uses WebGPU where your hardware supports it — including the GPU built into every Apple Silicon Mac — and falls back to CPU automatically, so it runs on the machine you already have.

Nothing leaves your machine

Transcription and slicing both happen on your own hardware. There's no account to create, and your audio is never sent anywhere — which is what makes it usable for unreleased episodes, dialogue under NDA, or an interview you promised to protect.

Pricing

One subscription. Everything in.

$29 a year covers every feature, all updates, and up to three machines. The seven-day trial comes first, and doesn't ask for a card.

$29 per year

  • Every feature, unlocked. Transcription, text-selection slicing and lossless export — nothing held back for a higher tier.
  • Updates included. Every improvement and new feature, for as long as your subscription is active.
  • Three activations. Use it on your desktop, your laptop, and one more.
  • A 7-day trial first. Everything unlocked, before you decide anything.

Runs entirely on your machine. Your audio never leaves it — Vocal Slice only checks your subscription now and then. If it lapses, transcribing and exporting pause until you resubscribe; your files and settings stay exactly where they are.

Secure checkout via Polar · tax included

Requirements

What you'll need.

Operating system
Windows 10 or 11 (64-bit), or macOS 11 Big Sur and later

Graphics
Any WebGPU-capable GPU — including the one built into every Apple Silicon Mac, and most modern integrated graphics. CPU fallback included

Disk space
~200 MB for the app, plus 75–500 MB per transcription model

Internet
Needed once to download a model and activate your licence — not for transcribing

Download

Try it on your own audio.

Free for seven days with every feature available. Runs on Windows and macOS, Apple Silicon and Intel alike.

Windows 10/11 64-bit · macOS 11 and later (Apple Silicon and Intel). Full requirements.

Windows may warn you the first time. Vocal Slice isn't yet signed with a code-signing certificate, so SmartScreen shows “Windows protected your PC — unknown publisher”. Choose More info → Run anyway to continue.