Features

Built for serious transcription

ScribeZapp combines research-grade AI models with a clean Windows interface — everything runs locally, privately, with no usage limits.

faster-whisper + WhisperX

ScribeZapp uses the most accurate open-source transcription stack available — the same models used by enterprise cloud providers, running on your own hardware.

All plans

faster-whisper

CTranslate2-optimised version of OpenAI Whisper. INT8 quantised for fast CPU inference without sacrificing accuracy.

Base, Small, Medium, Large-v3 model sizes
Pro

WhisperX

Adds forced alignment and word-level timestamps to Whisper output. Enables accurate speaker diarization tied to exact word boundaries.

Word-level timestamps · Forced alignment
Pro

pyannote.audio

Speaker diarization pipeline bundled locally — no Hugging Face account needed. Automatically labels who spoke when.

Offline · MIT-licensed community models

All models run entirely on your CPU using CTranslate2 INT8 quantisation. A modern Intel/AMD 64-bit processor is all you need — no GPU, no CUDA, no special hardware.

Know who said what — 100% offline

Most transcription tools require you to upload audio to a cloud service to get speaker labels. ScribeZapp bundles the diarization models directly — no Hugging Face account, no internet connection needed at runtime.

  • pyannote.audio community models bundled locally
  • Automatically detects number of speakers
  • Speaker labels tied to word-level timestamps
  • Exported in transcript with [Speaker 1], [Speaker 2]… format
  • Works on interviews, meetings, podcasts, and panels
  • No speaker enrollment or training required
Get ScribeZapp Pro →
🧑‍🤝‍🧑
Audio / Video file (local)
WhisperX — word timestamps
pyannote — speaker segments
Speaker-labelled transcript

Everything runs on your machine. Nothing is sent anywhere.

Turn recordings into actionable, shareable documents

Pro adds the professional workflow on top of Core — anonymize, summarize, export, and translate, all without leaving your machine.

Pro

🛡️ PII Detection & Anonymization

ScribeZapp automatically scans your transcript for names, emails, phone numbers, and other identifying information and replaces them with placeholders. Restore the original values (deanonymize) any time, entirely on your machine.

Microsoft Presidio + multilingual NER, fully offline
Pro

✨ AI Summary & Action Items

Send your anonymized transcript to the AI assistant you already use — Claude, ChatGPT, or another — and get back a summary, key decisions, and action items. Names are restored automatically once the result comes back to your machine.

Anonymize locally → your AI assistant → restore locally
Pro

📄 DOCX & Anonymized PDF Export

Export a properly formatted Word document with speaker labels and timestamps, or a anonymized PDF with PII removed — ready to share externally without manual cleanup.

.docx · anonymized .pdf
Pro

🌐 Offline Transcript Translation

Transcribe a recording in one language and translate the transcript into another — fully offline, powered by the same CTranslate2 engine behind Langzapp's document translation.

Same engine as Langzapp · offline CT2 pipeline
Pro

📚 Custom Vocabulary / Glossary

Teach ScribeZapp your domain's jargon — legal terms, medical terminology, product names — for more accurate transcription on specialized recordings.

Per-project glossary management
Pro

🧑‍🤝‍🧑 Speaker Diarization

Automatically identify who spoke when, fully offline. See the dedicated section above for details.

pyannote.audio · bundled locally

How AI Summary stays private: ScribeZapp anonymizes your transcript locally before anything is sent anywhere. The connection to your AI assistant happens from your own machine, using your own access — ScribeZapp does not operate or have visibility into that connection. When the result comes back, names are restored locally using your machine's anonymization map.

Input and output formats

Input formats (audio & video)

.mp3

MP3 Audio

The most common audio format. Works with all bitrates and sample rates.

.mp4 / .m4a

MP4 / M4A

Video files and Apple audio. Audio track extracted automatically via FFmpeg.

.wav

WAV Audio

Uncompressed audio. Ideal for high-quality field recordings and studio audio.

.flac

FLAC Audio

Lossless audio compression. Common in archival and academic research contexts.

.mkv / .mov

MKV / MOV Video

Container formats for video. Audio track extracted automatically.

.ogg / .webm

OGG / WebM

Web and open-source formats. Common in browser recordings and streaming.

Export formats (transcript output)

.txt

Plain Text

Clean transcript text. Copy into Word, Google Docs, or any text editor.

  • Speaker labels (Pro)
  • Timestamps optional
.srt

SRT Subtitles

Standard subtitle format. Drop directly into video editors like Premiere, DaVinci, or Final Cut.

  • Word-level timing (Pro)
  • Speaker IDs in subtitles
.vtt

VTT Subtitles

WebVTT format for web video. Compatible with HTML5 video players and YouTube caption uploads.

.json

JSON

Structured transcript with full timing, confidence scores, and speaker data. For developers and custom workflows.

.docx

Word Document Pro

A properly formatted Word document with speaker labels and timestamps — ready to edit or share.

.pdf

Anonymized PDF Pro

An anonymized PDF with PII removed, ready to share externally without manual cleanup.

Your recordings never leave your PC

Cloud transcription services require you to upload your audio to their servers — where it may be stored, reviewed, or used to train AI models. ScribeZapp is fundamentally different.

  • Audio and video files never leave your machine
  • No API calls during transcription
  • AI models stored locally on your PC
  • No telemetry or usage tracking in the app
  • Works completely offline after installation
  • Compliant with attorney-client privilege, HIPAA contexts, and IRB requirements
🔒
Your audio file
ScribeZapp (local)
AI models (local)
Transcript (local)

Zero network requests during transcription.

Transcribe in 99+ languages

OpenAI Whisper supports virtually every major language. ScribeZapp gives you access to all of them — without sending a single audio byte to OpenAI's servers.

🇺🇸 English
🇪🇸 Spanish
🇫🇷 French
🇩🇪 German
🇮🇹 Italian
🇵🇹 Portuguese
🇳🇱 Dutch
🇷🇺 Russian
🇵🇱 Polish
🇯🇵 Japanese
🇨🇳 Chinese
🇰🇷 Korean
🇮🇳 Hindi
🇸🇦 Arabic
🇹🇷 Turkish
🇺🇦 Ukrainian
🇸🇪 Swedish
🇳🇴 Norwegian
🇩🇰 Danish
🌐 +80 more

All 99+ Whisper-supported languages work with the same download. Language detection is automatic — ScribeZapp can detect the language from the audio if you don't specify one.

Ready to transcribe privately?

Free forever. Pro from $19/year — 30-day free trial, no credit card required.