Pinna detects who-spoke-when and turns your call and meeting audio into speaker-labeled, time-aligned transcripts. Submit by web, email, or API — and get back a transcript that tells the voices apart.
Detects how many speakers are in the audio and labels who-spoke-when automatically — no manual tagging, no guessing.
Whisper-class accuracy in Hebrew, Arabic, Russian and English — including the languages cloud tools get wrong. Speaker diarization built in.
Hebrew, Arabic, English, and Russian — including right-to-left scripts. Auto-detected, or force a language when you know it.
Upload in the web app, email an attachment, or call the API. Same engine, same speaker-labeled result, whichever you use.
The core runs with no third-party cloud-API dependency — the foundation that makes on-prem and airgap deployment possible (available with Enterprise).
Every speaker turn carries a timestamp, so you can jump straight to the moment in the recording a line came from.
Drag a file into the app, watch the job progress, and download the transcript when it's done. The simplest way to get started.
Send the audio as an attachment and get the transcript back — no app needed. Perfect for forwarding a recording the moment a call ends.
<you>@in.pinna.im. The sending address has to be on your account allowlist.Submit jobs programmatically and pull results into your own systems. Built API-first, so automation is a first-class path, not an afterthought.
Every plan includes all languages and diarization. You pay for hours and delivery mode, not features.
The core diarization and transcription run on a self-contained engine. Your audio doesn't get handed to a third-party transcription API to do the work.
Because the engine is offline-capable, Enterprise can run Pinna entirely on infrastructure you own — including airgapped networks — so audio never leaves your perimeter.
Email ingress is allowlisted per account, and the authenticated app manages API keys and job access — so only the people you authorize can submit and read.