vsref
Narakeet logo

Narakeet Review

Indie batch-content workhorse: aggregates a very broad voice/language catalog for voiceovers, audiobooks and video narration, priced per output minute (prepaid packs, no subscription) rather than per character; not aimed at real-time agent use.

Among the 46 text-to-speech tools we track, Narakeet has the 3rd-widest language coverage and the 2nd-largest voice library - a fit for multilingual and localization projects and content and character work.

See pricing

Facts verified Jul 20, 2026Try Narakeet →

If you sign up through links on this page, vsref may earn a commission; programs exist on both sides of most comparisons, and commissions never change verdicts. How we make money

What we know about Narakeet

Narakeet sits in the text-to-speech apis category, where it is indie batch-content workhorse: aggregates a very broad voice/language catalog for voiceovers, audiobooks and video narration, priced per output minute (prepaid packs, no subscription) rather than per character; not aimed at real-time agent use. We keep this Narakeet profile grounded in primary sources, each fact dated to when we last confirmed it.

Narakeet does not publish a public per-minute rate we have been able to verify, so the price figure here stays blank until we can confirm one. We would rather show nothing than a guessed number; if you have a current Narakeet quote, it is the fastest way for us to close that gap.

On capabilities, Narakeet covers streaming audio output, emotion / style controls, and ssml support, and does not offer realtime websocket api, pronunciation dictionaries, and commercial use on free tier. Each of those is verified against Narakeet's own docs or dashboard, not marketing copy.

Placed against the 46 text-to-speech tools we track, Narakeet's strongest showing is the 3rd-widest language coverage, while it trails at 2nd on stock voices - a spread worth weighing against your own priorities.

In total we track 13 verified facts for Narakeet today, each linking the primary source it came from so you can check our work - and vendor claims we have not measured ourselves are labeled as such on the Narakeet fact sheet below.

Reviewed by vsref Editorialfacts verified Jul 20, 2026Methodology →

Fact sheet

Pricing
Pricing facts
Pricing modelPrepaid pay-as-you-go minute packs (one-time purchasesJul 20
Free tier quota20 free conversions (max 1 KB audio scriptJul 20
Capabilities
Capabilities facts
Streaming audio output✓ YesJul 20
Realtime websocket API✗ NoJul 20
Voice library size900 voicesJul 20
Languages supported100 languagesJul 20
Emotion / style controls✓ YesJul 20
SSML support✓ YesJul 20
Pronunciation dictionaries✗ NoJul 20
Compliance & trust
Compliance & trust facts
Commercial use on free tier✗ NoJul 20
Build experience
Build experience facts
Official SDKsREST API with per-language examples (e.g. C# guide)Jul 20
Output formatsWAV (16-bit PCM, long-content API only), MP3, M4AJul 20
Max input per requestStreaming API: 1 KB scriptJul 20

Considering a switch? Best Narakeet alternatives →

STT in this stack

Text-to-speech is half of a voice pipeline: the other half is the speech-to-text that listens. Compare transcription engines on accuracy, streaming latency, and per-minute price: Deepgram, AssemblyAI, GPT-4o Transcribe, or the full speech-to-text comparison →

Voice agents in this stack

The engine is one layer: a voice agent speaks through its text-to-speech engine, but orchestration, telephony, and turn-taking come from the agent platform. Compare the platforms builders pair Narakeet with: Pipecat, OpenAI Realtime API, Twilio ConversationRelay, or the full voice-agent comparison →

Distribute it

Most Narakeet voiceover ends up in short-form video, and publishing that video across TikTok, YouTube, and Instagram is a scheduling problem with real per-channel pricing. Compare the schedulers creators actually run: Buffer, Postiz, Mixpost, or the social media scheduling platforms compared →

Avatar video in this stack

A cloned or bring-your-own Narakeet voice does not have to stay audio-only: AI avatar video platforms lip-sync it onto a talking avatar for finished video. Compare the platforms: HeyGen, Synthesia, Hedra, or the full avatar-video comparison →