Whisperkit

Whisperkit

On-Device AI Transcription

0 ratings
Free

Rating summary

Details

  • Released
  • Updated
  • June 9, 2026
  • August 28, 2026

Features

Whisperkit screenshot #1 for iPhone
Whisperkit screenshot #2 for iPhone
Whisperkit screenshot #3 for iPhone
Whisperkit screenshot #4 for iPhone
iphone
ipad
🖼️Get Icon
Icons↘︎

About

Whisper turns speech into accurate text — entirely on your device. Powered by Apple's Neural Engine and an open-source multilingual speech recognition model, Whisper transcribes audio recordings, voice memos, lectures, meetings, podcasts and videos directly on your iPhone. Nothing is uploaded. Nothing is tracked. No accounts. KEY FEATURES • On-device transcription — your audio never leaves your iPhone • Record audio directly in the app, or import files from Files, iCloud, or Photos • Supports many languages including English, Chinese, Spanish, French, German, Japanese, Korean, Russian, Portuguese, and Italian — plus auto-detection • Real-time progress with live segment-by-segment streaming • Multiple model sizes — pick speed or accuracy to match your needs • Performance monitor shows CPU, memory and Neural Engine usage in real time • Works offline once a model is downloaded • Copy or share transcripts with one tap • Beautiful, modern interface designed for iOS and iPadOS WHY ON-DEVICE MATTERS Most transcription apps send your audio to a server. Whisper does not. Your meetings, voice notes, interviews and personal recordings stay private because everything runs locally on your iPhone's Neural Engine. • No accounts to create • No subscriptions, no ads • No analytics or tracking SDKs • Works in airplane mode after the first model download • Safe to use with confidential and sensitive recordings PERFECT FOR • Journalists transcribing interviews • Students capturing lectures • Professionals transcribing meetings • Podcasters generating show notes • Writers dictating drafts • Anyone who values privacy Whisper is open speech recognition done right — fast, accurate, and yours alone.
Show more

What's New in Whisperkit

1.1

August 28, 2026

Version 1.1 is a redesign of the whole app, plus a substantial rebuild of what happens under it. A NEW LOOK, BUILT FOR THE WORK Whisper has been redesigned around the recording itself. Every session now shows its real waveform — drawn from the actual audio, so silences are gaps and speech has shape — and you can tap anywhere on it to jump straight there. Transcripts are laid out as timestamped lines rather than a wall of text; tap a line and the audio follows. On iPad, past sessions live in a rail down the side, so moving between recordings takes one tap. On iPhone, recent work sits on the home screen. Dark mode is supported properly throughout. EXPORT YOU CAN SEE BEFORE YOU SEND Choose TXT, SRT, VTT, Markdown or JSON and the app shows you the actual file before you share it, with its size. Speaker labels can be included or left out. Subtitle timings have also been corrected — they were rounding a millisecond early. BETTER MODELS, SMALLER DOWNLOADS The model line-up is rebuilt around Neural Engine–optimised builds, with sizes that are measured rather than estimated: • Small — 217 MB. Genuinely multilingual. • Turbo — 646 MB. Recommended for most people, and less than half the size of the model it replaces. • Large V3 — 948 MB, and Large V3 Turbo — 1.1 GB. Previously these pulled over 3 GB each. Downloads show real progress, can be cancelled, and resume where they left off. Long-press a model to delete it. CHINESE, JAPANESE AND KOREAN Transcribing these languages could previously return nothing at all. An internal repetition check misread dense CJK text as a decoding failure and discarded good results. That is fixed. Quiet and far-field recordings are also far less likely to be thrown away. READY WHEN YOU ARE Whisper now runs on the GPU by default and is ready in seconds, rather than spending minutes preparing a model on first use. You can start recording while a model is still loading, and importing a long video no longer stalls. Live text now accumulates as the audio is read, instead of resetting every thirty seconds. Everything still runs entirely on your device. No account, no analytics, and nothing you record ever leaves your iPhone or iPad.

More

Developer apps