Hear it. Read it. Cite it.
Upload audio or video: an interview, a meeting, a dictation. Akoé writes it out, labels who said what, and ties every word to its second in the recording: check anything against the sound, fix mistakes, and download the transcript in whatever form you choose.
Upload. Check. Cite.
Three steps, no manual needed. Built for people who work with recordings, not for people who work with software.
Drop in your recording
Audio or video, minutes or hours. From video we keep only the sound. The images are dropped during processing and never kept.
Watch the transcript arrive
The first minutes appear while the rest is still transcribing and the voices are getting labelled. We email you when it's done.
Check it & export it
Fix anything in the editor, then export Word, subtitles, text, data or TEI XML, with your corrections in every format.
Automatic transcription is never perfect. The akoé platform is built for checking effectively.
Every automated transcript needs human checking. What matters is how fast you find and fix the mistakes.
Hear what you're fixing
The audio plays from the word you're checking at your preferred speed, so you hear the moment while you fix it at your own pace. Words the AI is unsure about stay flagged until you've looked; find & replace clears a repeated mistake everywhere.
Voices, untangled
Automatic speaker separation shows who said what. Wrong name, wrong voice, missed change of speaker—each is a one-click repair.
See who and what is mentioned
Akoé scans the transcript for the people, places, organisations and dates mentioned. It builds an automatic index and even lets you define your own categories for things to scan. Click a mention to jump to it.
Subtitles
Add subtitles to your videos and podcasts (SRT and VTT timed to the word) in the original language and in translation.
Translate, then verify
Machine-translated drafts open beside the original, line by line, editable, timings intact and ready to check.
Dictate straight in the browser
Tired of typing? Press record, speak, stop. Your words come back as text so you don't have to write from scratch.
Forgot what was said exactly? Search semantically.
Search all your recordings at once, by meaning—not just exact words, and even in a different language than the one of recording. Every result is a real moment from your own transcripts: click it and it plays.
One Product, Many Use Cases
akoe.ai does automated transcription (converts speech to text), speaker separation, dictation, translation and entity recognition—and because every transcript in your workspace is searchable and every word is timed, it works not only as a transcription service but also as a research tool: find who said what, cite it, or pull the exact soundbite for your edit.
Researchers & Universities
Interview and focus-group transcription your ethics committee can approve: EU-only processing, deletion guarantees, and a data-processing agreement (DPA) upon request.
Oral History & Archives
Speaker-labelled, timecoded transcripts. Recorded memory becomes a searchable, citable record.
Professional Transcribers
Start from a strong draft instead of a blank page. Playback, shortcuts and speaker tools built for high-volume correction. Your skill goes into perfecting, not typing.
Journalists
Sound in, quotable text out, every word checkable against the tape. Retention and processing on EU servers, gone on your schedule.
NGOs & Advocacy
Field interviews, testimonies and case documentation, deleted on your schedule.
Professionals Who Dictate
Tired of typing? Speak, then simply edit your dictated text. A dictation platform for confidential work.
Councils, Boards & Committees
Meeting recordings become draft minutes: who spoke, what was said, timestamped to the second—the secretary will check, not retype from scratch. Invoicing for public bodies and businesses available. Contact us.
Podcast & Video Makers
Word-timed SRT/VTT subtitles with free translated versions of them.
Speech Corpus Developers
Gold-standard transcripts, faster: correct a strong draft in a verbatim-first editor, then export word-timed, speaker-labelled JSON or CSV.
About 100 languages, auto-detected
Upload in whatever language the recording is in—akoé.ai detects it automatically, or you set it yourself.
English · Spanish · French · Portuguese · German · Italian · Polish · Dutch
Grouped by how automatic transcription tends to perform on clear audio. Accuracy varies depending on the quality of the recording and the language, which is exactly why the editor exists. Use your free 5 minutes on a real file to see what to expect.
We respect your material
Transcription without subscription.
€5 per hour of audio, pro-rated. That's it.
Billed to the second
A 90-minute interview is €7.50. A 45-minute lecture, €3.75. A 20-minute memo, €1.67. No tiers, no rounding up to the hour.
The exact price, before you pay
Your file's duration is known before checkout, so the price you see is the price you're charged. Every time. If a file can't be processed, the charge refunds itself automatically.
No subscription required
Nothing to prepay, no top-ups, nothing to cancel. Your first 5 minutes are free, without a card.
What would my recording cost?
minutes
Fair questions before your first upload
How do I convert audio to text?
How accurate is the automatic transcription?
When will automatic transcription struggle? What makes automatic speech recognition hard?
Does it detect laughter, pauses, or emotion?
Can it tell who said what?
Which languages does each feature support?
How long does a transcription take?
Can I transcribe video, and what happens to the footage?
What file types can I upload?
Can I dictate instead of uploading a file?
What does it cost?
Why is there a €1 minimum on checkouts?
Couldn't I just translate my transcript with an AI chatbot?
Where is my audio processed? Can I use this under GDPR?
Will my recordings be used to train AI models?
What can I export?
Can I include the transcript in a thesis, paper or report?
Can I search across my recordings?
Do I need a subscription?
Your first 5 minutes are free
A real recording or just your voice—see how it works. No card details needed.
Upload a recording Start dictating