vsref
Handy logo

Handy Review

Private local-first open-source dictation

Among the 43 speech-to-text tools we track, Handy has the 12th-widest language coverage - a fit for multilingual and localization projects.

See pricing

Facts verified Jul 20, 2026Try Handy

If you sign up through links on this page, vsref may earn a commission; programs exist on both sides of most comparisons, and commissions never change verdicts. How we make money

What we know about Handy

This is our verified profile of Handy, a speech-to-text apis platform - private local-first open-source dictation. Every fact about Handy below carries the source it came from and the day we checked it.

Handy does not publish a public per-minute rate we have been able to verify, so the price figure here stays blank until we can confirm one. We would rather show nothing than a guessed number; if you have a current Handy quote, it is the fastest way for us to close that gap.

On capabilities, Handy covers language auto-detection, custom vocabulary / keyterm boosting, speech translation, self-host / on-prem option, and ai formatting / edit features, and does not offer word-level timestamps, entity detection, and sentiment analysis. Each of those is verified against Handy's own docs or dashboard, not marketing copy.

For compliance, with Handy: SOC 2 Type II is not reported. If you are in a regulated space, confirm the current posture with Handy before you commit, since these change plan by plan.

Handy ranks the 12th-widest language coverage of the 43 speech-to-text tools we track, so where it lands for you depends on which of those matters more.

In total we track 29 verified facts for Handy today, each linking the primary source it came from so you can check our work - and vendor claims we have not measured ourselves are labeled as such on the Handy fact sheet below.

Reviewed by vsref Editorialfacts verified Jul 20, 2026Methodology →

Fact sheet

Pricing
Pricing facts
Pricing modelone-timeJul 20
Free tier quotaUnlimited (free app)Jul 20
Capabilities
Capabilities facts
Languages supported99 languagesJul 20
Language auto-detection✓ YesJul 20
Speaker diarization✗ Not availableJul 20
PII redaction✗ Not availableJul 20
Word-level timestamps✗ NoJul 20
Custom vocabulary / keyterm boosting✓ YesJul 20
Entity detection✗ NoJul 20
Sentiment analysis✗ Not offered.Jul 20
Summarization endpoint✗ NoJul 20
Speech translation✓ YesJul 20
Priced audio-intelligence add-onsAI post-processing (reformat/translate/grammar)Jul 20
Compliance & trust
Compliance & trust facts
HIPAA BAA available✗ Not availableJul 20
SOC 2 Type II✗ NoJul 20
GDPR / EU data residency✗ NoJul 20
Self-host / on-prem option✓ YesJul 20
Build experience
Build experience facts
Supported audio formatsMicrophone input onlyJul 20
Websocket streaming API✗ NoJul 20
Commercial
Commercial facts
Model size (parameters)Model sizes 31MB-1.6GB (Whisper, Parakeet V3, +more)Jul 20
Hardware to self-hostParakeet: CPU (Skylake+); Whisper: GPU/Apple SiliconJul 20
Hosted API availableNo (local desktop app only)Jul 20
Project maintenance statusActive - v0.9.3 released 2026-07Jul 20
GitHub stars27,007 starsJul 20
PlatformsmacOS, Windows, LinuxJul 20
Local vs cloud processingFully local/offline STTJul 20
App pricing (one-time vs subscription)Free (open-source, MIT)Jul 20
AI formatting / edit features✓ YesJul 20
App integrationsRaycast extension; OpenAI-compatible LLM endpointsJul 20

Considering a switch? Best Handy alternatives →