Skip to content
Cadence
sec · 01 / heroturn-taking · resolved
Mac · iPhone · iPad

Who said what, and when.

Native transcription with speaker diarization, on Mac, iPhone and iPad. Drop in an interview or share a voice memo, and get a timestamped transcript that knows who was talking — on the Neural Engine, no account, no upload, no per-minute meter.

Apple Silicon nativeSpeakers = 22 MBv1.4.1
research-call-04.m4aParakeet TDT v2 · 3 speakers · 47 min
00:00 / 01:56Live demo — click the waveform
Maya · interviewer
Dan · participant
Priya · participant
DER 10.6% · Diarization 22 MB
ExportTXTSRTVTT · ProMD · ProJSON · ProCSV · Pro
Click a line to jump
Core framework
Neural Engine + Core ML
Speech engines
Five engines
Hour of audio, on a Mac
Under a minute
Data leaves device
Never
sec · 02 / screenshots
Screenshots

A reading room, not a waiting room.

Drop a file in the library, pick an engine, and the transcript builds while you keep working. Play it back and the transcript follows along, word by word. Five screens, one at a time: the library, the engine manager, the queue with its parallel lanes, a search that lands on the line, and the side-by-side comparison of two engines on the same recording.

The library — Every recording, its status, its engine — and the transcript beside it.
01The libraryEvery recording, its status, its engine — and the transcript beside it.
Engine manager — Five engines, their sizes and licences. Download once, run offline.
02Engine managerFive engines, their sizes and licences. Download once, run offline.
The queue — Transcription and speakers run as parallel lanes, each with a countdown.
03The queueTranscription and speakers run as parallel lanes, each with a countdown.
Model comparison — Two engines on one recording, with the words they disagree about marked. Pro.
05Model comparisonTwo engines on one recording, with the words they disagree about marked. Pro.
sec · 03 / voice
“Four people on a call and every other transcript reads like one person talking to themselves. This one hands me the turns, stamped, and never asks me to upload anything.”
— Priya Shah
Staff researcher · Lisbon
sec · 04 / engines
Five engines

Tell it what you care about.

Accuracy, download size and language coverage pull against each other, and no single engine wins all three. Cadence ships five and lets you switch per recording — each engine keeps its own transcript, so nothing is overwritten when you change your mind. Apple Speech and Parakeet v3 compact are free and unlimited; the three full-size models come with Pro.

I care most about
→ Use
Parakeet TDT v2. The lowest word error rate Cadence can reach, 5.4 on English, still entirely on the Neural Engine. If the recording isn't English, take the next option instead.Whisper large-v3-turbo. One hundred languages, detected automatically. It is a little slower and a little less accurate than Parakeet on English, and it is the only engine that will not be surprised by your recording.Apple Speech. Already inside macOS, so the app transcribes the moment it opens. It uses whichever locales the system has installed — add more in System Settings and Cadence picks them up.v3 compact. 299 MB for all 25 European languages and the fastest throughput of the downloadable engines. Accuracy sits a step below full v3, which most interview audio will not notice.
e-01DEFAULT
Apple Speech
no download

Built into macOS. Transcribes on first launch with nothing to fetch, in whichever locales the system already has.

Accuracygood
Speedinstant
Languagessystem
Footprintnone
e-02Recommended
Parakeet TDT v2
464 MB · ENPro

The most accurate engine Cadence ships. English only — reach for v3 or Whisper if the recording isn't.

Accuracy5.4 WER
Speedvery fast
LanguagesEnglish
Footprint464 MB
e-03Recommended
Parakeet TDT v3
483 MB · 25 EUPro

Nearly v2's accuracy across 25 European languages. The default choice for multilingual European work.

Accuracy5.7 WER
Speedvery fast
Languages25
Footprint483 MB
e-04Recommended
v3 compact
299 MB · 25 EU

Two thirds the download of full v3 with the same language set. The sensible middle when disk is the constraint.

Accuracygood
Speedfastest
Languages25
Footprint299 MB
e-05Recommended
Whisper large-v3-turbo
627 MB · 100 langPro

For recordings in something Parakeet never saw. It detects the language itself — nothing to configure.

Accuracy7.0 WER
Speedfast
Languages100
Footprint627 MB

Apple Speech has no language count on purpose: it transcribes whichever locales the system has downloaded, you can add more in the system's settings, and the set changes under the app. Cadence asks the system at launch rather than printing a number that goes stale.

Parakeet TDT 0.6B — NVIDIA (CC-BY-4.0) · Whisper large-v3-turbo — OpenAI, Core ML conversion by Argmax (MIT) · pyannote Community-1 — packaged by FluidInference (MIT)

sec · 05 / speakers
Diarization

Turn-taking is the subject.

Cadence separates the voices before it writes anything down, so the transcript arrives already attributed. 10.6% diarization error rate, 22 MB, any language — and it is part of the free app, not a paywalled extra.

Fig. 01 — speaker turns over 47 minutes
Maya
Dan
Priya
00:00:0000:23:3100:47:12
What that buys you
Attribution, not guesswork. Each line carries the speaker it came from, in colours that stay consistent between the legend and the transcript.
Timecodes you can trust. Every turn is stamped, so a quote can be found again in the audio in seconds. Click anywhere in a turn to play from there.
Play it back and read along. A transport under the transcript plays the file from 0.75× to 2×, marking the word as it is spoken and tinting the turn it belongs to. Click any word in the turn being played to jump straight to it.
Any language. Diarization works on the sound of the voices, so it doesn't care what they're saying.
Video too. Screen recordings of meetings go in the same window as audio files.
sec · 06 / on-device
Stays on your device

Every other transcriber wants your audio. This one wants 22 MB.

The only network request Cadence ever makes is fetching the engine you chose, from Hugging Face, once. After that it works on a plane, on a locked-down network, and with a recording you are contractually not allowed to hand to anybody's cloud.

No account. Nothing to sign in to before you can transcribe.
No per-minute billing, no quota, no queue behind other people's files.
No telemetry, no crash reporter, no “help improve the product”.
Sandboxed, notarized, App-Store-shipped. Your files stay in your container.
Fig. 02 — where the audio goes
File inresearch-call-04.m4a
ANEdiarize → transcribe → align
Disktranscript, on your machine
NetworkNot used
sec · 07 / devices
Mac, iPhone and iPad

One app, three screens.

The same engines, the same speaker separation and the same transcripts, each running entirely on the device in your hand. The Mac does the most, the iPad most of that, and the iPhone the core. Each keeps its own library: nothing syncs between them, because nothing leaves the device. One Pro purchase unlocks all three.

Cadence on iPad comparing Apple Speech and Parakeet transcripts of one recording side by side
Cadence on iPhone showing a transcript with speakers labelled and playback controls
What Cadence does on each device
FeatureMaciPadiPhone
Apple Speech and Parakeet v3 compactIncluded, freeIncluded, freeIncluded, free
Full-size Parakeet and WhisperIncluded with ProIncluded with ProIncluded with Pro
Speakers identified, and renamedIncluded, freeIncluded, freeIncluded, free
Playback that follows the wordsIncluded, freeIncluded, freeIncluded, free
Search names and transcriptsIncluded, freeIncluded, freeIncluded, free
Send recordings in from other appsIncluded, freeIncluded, freeIncluded, free
Several recordings queued at onceIncluded with ProIncluded with ProIncluded with Pro
Drag and dropIncluded, freeIncluded, freeNot on this device
TXT, SRT and clipboard exportIncluded, freeIncluded, freeNot on this device
WebVTT, Markdown, JSON and CSVIncluded with ProIncluded with ProNot on this device
BriefsIncluded with ProIncluded with ProNot on this device
Model comparisonIncluded with ProIncluded with ProNot on this device
Transcribe AllIncluded with ProNot on this deviceNot on this device
Dictation into any appIncluded with ProNot on this deviceNot on this device

✓ free · PRO with the one-time upgrade · — not on that device. The Mac needs Apple Silicon; iPad and iPhone need iPadOS or iOS 26. Briefs, and the verdict inside model comparison, need Apple Intelligence.

sec · 08 / pricing
Free, with a Pro unlock

Pay once. Own the tool.

No subscription, no seats, no minutes. The free app is a complete transcriber — two engines, whole files, unlimited, speakers included. Pro adds the models that squeeze out the last of the accuracy, and the tools for thirty recordings instead of one.

Free
$0always

A whole transcriber, not a trial. No watermark, no time limit, no sign-up.

+Two engines: Apple Speech and Parakeet v3 compact
+Unlimited transcription, 25 European languages
+Speaker diarization, with renaming
+Playback with word-level follow-along
+Search your library and your transcripts
+TXT, SRT and clipboard export
+Free updates, forever
Download free
Pro · one-timeEarly launch
$9.99$29.99pay once

Early launch price through 1.x, rising to $29.99 at 2.0

One purchase covers Mac, iPhone and iPad ·what each can do

Full-precision Parakeet and Whisper, plus the tools for a backlog rather than a file. Queue the library, compare engines, export in whatever format your pipeline eats, and on the Mac, dictate into any app.

+Everything in Free
+WebVTT export, with options
+Full-size Parakeet and Whisper
+Markdown export
+Transcribe All
+JSON export
+Briefs, on device
+CSV export
+Model comparison
+Copy a brief as Markdown
+Dictation into any app, on the Mac
+Dictate and send, by key or mouse
Get Pro — $9.99 on the App Store
sec · 09 / faq
Honest answers

Questions, answered.

The ones that come up most. If yours isn't here, write to us.

Q · 01

Do I need an internet connection?

Only to install the app and to download an engine the first time you pick one. The default engine is already built into macOS and iOS, so a fresh install transcribes with nothing downloaded and nothing online.

Q · 02

What does it run on?

macOS 26 or later, on an Apple Silicon Mac: Cadence has been an arm64-only build since 1.2.0. Within that, the downloadable engines run on the Neural Engine and Apple's built-in engine runs on any Mac the app installs on. Cadence is also on iPhone and iPad with iOS 26 or later, and one Pro purchase covers all three.

Q · 03

Can I use it on confidential recordings?

That is the reason it exists. Nothing is uploaded, so there is no processor to add to a DPA and no terms of service quietly reserving the right to train on your interviews.

Q · 04

Does Cadence record?

No. Transcription reads files you already have — audio or video, dropped into the window. Keep recording with whatever you record with. The one time Cadence opens the microphone is dictation on the Mac, and only while you dictate: the words are typed where your cursor is, and the audio is never saved or sent anywhere.

Q · 05

What is actually in Pro?

On the Mac, six things: the full-size speech models — Parakeet TDT v3 at full precision, the English-tuned v2, and Whisper large-v3-turbo; Transcribe All, which works through every untranscribed recording in the library one at a time; briefs, which summarize a transcript into decisions, action items and open questions; model comparison, which runs one recording through several engines in a single action, shows two of them side by side and has the on-device model judge which one read each disagreement correctly; dictation into any app, which types what you say wherever your cursor is, from a keyboard shortcut or a mouse button; and the full export set — WebVTT, Markdown, JSON and CSV, with their options. On iPad, Pro adds the full-size models, queueing several recordings at once, briefs, model comparison and the full export set. On iPhone it adds the full-size models and queueing several recordings at once, and one purchase covers all three.

Q · 06

Can Cadence type what I say into other apps?

On the Mac, with Cadence Pro. Press a shortcut, speak, and Cadence types the words wherever your cursor is: hold it while you talk and let go, or press it once and it inserts when you pause. A second shortcut dictates and then presses Return, for a chat or a terminal prompt, and saying "slash commit" types /commit. A mouse button can do either. Speech is recognized on your Mac, by Parakeet TDT v3 when it is downloaded and Apple Speech otherwise. To type for you, Cadence needs permission under Accessibility; without it, the text is copied and you press ⌘V. Dictation is off until you switch it on in Settings ▸ Voice.

Q · 07

What do briefs need?

They run on Apple's on-device model, so they need Apple Intelligence, and nothing is uploaded to produce one. That model reads 15 languages, against the 25 Cadence transcribes free and the 100 Whisper handles — where it cannot read a transcript, Cadence says so rather than offering a button that fails, and it checks that before it checks whether you have Pro. It also declines the occasional passage; those are listed with their timestamps instead of being quietly dropped, so a partial brief tells you it is partial.

Q · 08

What can the free app actually do?

Transcribe as much as you like on two engines: Apple Speech, built into macOS, and Parakeet TDT v3 compact, a 299 MB download covering 25 European languages. Whole files, not samples, with the speakers identified, playback that follows the words, search, and export to plain text, SRT and the clipboard. No trial timer, no watermark, no account, and no cap on minutes. Pro adds the full-size models and the tools for working through a backlog.

Q · 09

Why is speaker identification free when it downloads a model?

Because it is what the app is for. Diarization is a 22 MB download that adds to whichever engine you are using rather than replacing it, so the free tier produces speaker-attributed turns rather than a wall of text. Parakeet v3 compact is free for a related reason: with only one free engine you could never have two models installed, so model comparison would be a feature we sell that you could never try.

Q · 10

What's coming next?

A speaker library that recognises the same voices between recordings; Shortcuts actions; per-recording settings that override the defaults; and editing the transcript text itself. None of it ships today — listed because it's where Cadence is going. Search and speaker renaming used to be on this list and now ship, both free.

sec · 10 / also
Also from Magenta Creations
Silhouette — clean cutouts, in bulk.

Batch background removal for Mac, with marketplace presets, real edge control and readiness checks before you upload.

Contour — segment anything, on your Mac.

Prompt-driven image segmentation with COCO, YOLO and per-frame video masks. Dataset-ready, entirely on-device.

sec · 11 / download

Who said what, and when. On your own devices.

Free forever on two engines, 25 European languages included, speakers identified. Unlock full-size Parakeet, Whisper and the backlog with a one-time Pro upgrade, and one purchase covers your Mac, iPhone and iPad.

macOS 26 · Apple SiliconiOS 26v1.4.1