# JigsawStack Speech-to-Text Review

> JigsawStack Speech-to-Text review: Whisper-based STT endpoint inside an indie all-in-one small-model AI API platform. Verified pricing, features, and the.

![JigsawStack Speech-to-Text logo](https://www.versusref.com/logos/jigsawstack-stt.png)

Whisper-based STT endpoint inside an indie all-in-one small-model AI API platform

Among the 43 speech-to-text tools we track, JigsawStack Speech-to-Text has the 5th-widest language coverage - a fit for multilingual and localization projects.

From $27/mo · facts verified Jul 20, 2026

## What we know about JigsawStack Speech-to-Text

This is our verified profile of JigsawStack Speech-to-Text, a speech-to-text apis platform - whisper-based STT endpoint inside an indie all-in-one small-model AI API platform. Every fact about JigsawStack Speech-to-Text below carries the source it came from and the day we checked it.

On pricing, JigsawStack Speech-to-Text starts at $27 per month for its entry tier. That is the sticker rate: real production cost usually runs higher once you add a language model, a voice provider, and telephony minutes.

On capabilities, JigsawStack Speech-to-Text covers language auto-detection, word-level timestamps, sentiment analysis, summarization endpoint, speech translation, and soc 2 type ii, and does not offer websocket streaming api. Each of those is verified against JigsawStack Speech-to-Text's own docs or dashboard, not marketing copy.

For compliance, with JigsawStack Speech-to-Text: SOC 2 Type II is in place. If you are in a regulated space, confirm the current posture with JigsawStack Speech-to-Text before you commit, since these change plan by plan.

Placed against the 43 speech-to-text tools we track, JigsawStack Speech-to-Text's strongest showing is the 5th-widest language coverage - a spread worth weighing against your own priorities.

In total we track 16 verified facts for JigsawStack Speech-to-Text today, each linking the primary source it came from so you can check our work - and vendor claims we have not measured ourselves are labeled as such on the JigsawStack Speech-to-Text fact sheet below.

## Fact sheet

### Pricing

| Fact | Value | Verified | Source |
| --- | --- | --- | --- |
| Pricing model | hybrid | Jul 20 | [source](https://jigsawstack.com/pricing) |
| Cheapest paid plan | $27 | Jul 20 | [source](https://jigsawstack.com/pricing) |
| Free tier quota | 1M tokens free per month | Jul 20 | [source](https://jigsawstack.com/pricing) |

### Capabilities

| Fact | Value | Verified | Source |
| --- | --- | --- | --- |
| WER (vendor-claimed) | ~10.3 % WER | Jul 20 | [source](https://jigsawstack.com/blog/jigsawstack-vs-groq-vs-assemblyai-vs-openai-speech-to-text-benchmark-comparison) |
| Languages supported | ~100 languages | Jul 20 | [source](https://jigsawstack.com/speech-to-text) |
| Language auto-detection | ✓  Yes | Jul 20 | [source](https://jigsawstack.com/docs/api-reference/ai/speech-to-text) |
| Speaker diarization | ✓  Included | Jul 20 | [source](https://jigsawstack.com/docs/api-reference/ai/speech-to-text) |
| Word-level timestamps | ✓  Yes | Jul 20 | [source](https://jigsawstack.com/docs/api-reference/ai/speech-to-text) |
| Sentiment analysis | ✓  Yes | Jul 20 | [source](https://jigsawstack.com/docs/api-reference/ai/sentiment) |
| Summarization endpoint | ✓  Yes | Jul 20 | [source](https://jigsawstack.com/docs/api-reference/ai/summary) |
| Speech translation | ✓  Yes | Jul 20 | [source](https://jigsawstack.com/docs/api-reference/ai/speech-to-text) |

### Compliance & trust

| Fact | Value | Verified | Source |
| --- | --- | --- | --- |
| SOC 2 Type II | ✓  Yes | Jul 20 | [source](https://jigsawstack.com/pricing) |

### Build experience

| Fact | Value | Verified | Source |
| --- | --- | --- | --- |
| Official SDKs | JS/TS (Node.js), Python | Jul 20 | [source](https://jigsawstack.com/docs/introduction) |
| Max file size / duration | 100MB max file, 4 hours max duration | Jul 20 | [source](https://jigsawstack.com/docs/api-reference/ai/speech-to-text) |
| Supported audio formats | MP3, WAV, M4A, FLAC, AAC, OGG, WEBM | Jul 20 | [source](https://jigsawstack.com/docs/api-reference/ai/speech-to-text) |
| Websocket streaming API | ✗  No | Jul 20 | [source](https://jigsawstack.com/docs/api-reference/ai/speech-to-text) |

Source: https://www.versusref.com/stt/tools/jigsawstack-stt/
