Narakeet: text to speech and narrated PowerPoint videos, billed by the minute
Updated on September 6, 2026
Narakeet is a text-to-speech and video automation platform that converts scripts, documents and slide decks into audio files or narrated MP4 videos. The first 20 files cost nothing and no account is required. After that, credits come in one-off packs, $6 for 30 minutes of output, with no subscription in sight. The library spans 900 voices in 100 languages, which is why teachers, e-learning teams and marketing departments rely on it for batch narration.
- 900 voices covering around 100 languages
- Automatic PowerPoint-to-video conversion with subtitles
- Pay-as-you-go credits that never expire
- 20 trial files without creating an account
- Free unlimited voice previews
- No voice cloning of any kind
- Slide animations and transitions get flattened
- Less expressive than premium voice engines
Speaker notes in, finished MP4 out, the Narakeet pipeline
Narakeet reads the presenter notes of a PowerPoint deck and generates the narration track of an MP4 video, with sync and subtitles handled automatically through neural speech synthesis. You write the script under each slide, upload the PPTX, and download a video ready for YouTube. Google Slides and Keynote decks work too, once exported to PowerPoint format.
The same engine goes well beyond slides. A Markdown script becomes a screencast, a translated SRT file becomes a dubbing track locked to its timestamps, and developers can drive everything through the command-line tools or the API for batch production.
| You upload | Narakeet returns |
|---|---|
| Text, Word, PDF, EPUB | MP3, WAV or M4A audio file |
| PowerPoint, Google Slides, Keynote | Narrated, subtitled MP4 video |
| Markdown script with images and clips | Screencast or video tutorial |
| Translated SRT or VTT subtitles | Synchronized dubbing track |
| Audio or video file | Text transcription |
Credits metered by the second, with no expiry date
Six dollars covers 30 minutes of generated media, $100 covers 1,000 minutes, and every file is deducted by the second. Credits never expire, a real advantage for burst production, one big course this month and then nothing for a while. The free tier caps each file at roughly 1,000 characters and keeps the output strictly personal.
One habit worth picking up early, the preview button costs nothing (you can audition twenty voices before committing, nobody is counting). Rates do shift over time, so the official pricing page remains the only number worth trusting.
Where Narakeet draws the line
Narakeet skips voice cloning entirely, every account draws from the same shared library, a field handled by dedicated voice cloning tools. Embedded PowerPoint animations are not carried over either, slides render as static images.
For explainer videos, internal training and faceless YouTube channels, that trade is easy to accept. The voices favor clarity over drama, and the pay-per-use model means a single tutorial costs cents rather than a monthly plan.
Frequently asked questions
Is Narakeet free?
Partly. The first 20 audio or video files are free, with no registration and a limit of roughly 1,000 characters per file, for personal use only. Beyond that, you buy duration-based credits, $6 for 30 minutes of output. Voice previews stay free at all times, no matter your plan.
Can I use Narakeet voiceovers commercially?
Yes, on a paid account. Commercial plans include full rights to everything you generate, with no copyright restrictions on the voices, YouTube monetization and client work included. Files made on the free tier cannot be monetized on social platforms or reused in commercial projects.
Narakeet or ElevenLabs, which one fits?
It depends on the job. ElevenLabs leads on emotional realism and cloning, sold through monthly subscriptions. Narakeet answers with one-off credits and automatic presentation-to-video conversion, a combination its rival does not attempt. For e-learning produced in bursts, the math usually lands on Narakeet's side.
Verdict: A full training module narrated and subtitled before your coffee goes cold, that is the everyday reality here. Teachers, trainers and e-learning teams producing in waves will feel at home, and anyone set on a cloned voice can simply feed an external recording into the same pipeline.
