Scribe v2 is a speech-to-text (transcription) model built specifically for handling large-scale, complex audio.
Previous transcription tools struggled with:
Imagine trying to transcribe a 2-hour podcast with multiple guests, background noise, and technical jargon. Basic tools fail here. Scribe v2 is engineered precisely for this challenge.
| Version | Optimized For |
|---|---|
| Scribe v2 | Long, complex batch recordings |
| Scribe v2 Realtime | Ultra-low latency, live agent use |
Standard transcription tools use fixed vocabulary lists — they always transcribe a word the same way regardless of context.
Keyterm Prompting is smarter. It uses the surrounding context of the transcript to decide when a specific term actually applies.
Without context awareness:
"The patient needs an MRI scan" → might transcribe as "em-ar-eye"
With Keyterm Prompting:
The model recognizes the medical context and correctly applies "MRI"
Beyond just transcribing words, Scribe v2 can automatically identify and locate sensitive information within your audio.
An entity is a specific category of meaningful information, such as:
Scribe v2 supports up to 56 detection categories and provides:
A compliance team processing recorded customer calls can automatically flag and redact credit card numbers spoken aloud — without manually reviewing hours of audio.
Audio Recording → Transcription → Entity Detection → Flag/Redact/Review
Real-world audio doesn't always stay in one language. Scribe v2 handles multiple languages within a single audio file automatically.
A multinational company records a meeting where speakers switch between English, Spanish, and French. Scribe v2 handles all three in a single pass.
Scribe v2 isn't just accurate — it's built for real production environments with features that developers and enterprises actually need.
Scribe v2 meets major security and privacy standards:
| Standard | What It Covers |
|---|---|
| SOC 2 | Data security controls |
| ISO 27001 | Information security management |
| PCI DSS L1 | Payment data protection |
| HIPAA | Health information privacy |
| GDPR | EU data privacy |
Plus: EU and India data residency and zero retention mode (your data isn't stored).
SCRIBE V2
│
├── Core Strength → Accurate batch transcription at scale
│
├── Keyterm Prompting → Context-aware custom vocabulary
│
├── Entity Detection → Automatic sensitive data identification + timestamps
│
├── Multi-Language → Single file, automatic language switching
│
└── Production Features
├── Speaker diarization
├── Word-level timestamps
├── Audio event tagging
└── Enterprise compliance
Scribe v2 transforms raw, complex, real-world audio into accurate, structured, compliant, and actionable transcripts — at scale.