# Methodology

> How we verify facts, how we compute real production costs, how we make money, and the live verdict-balance stats behind it.

How we verify facts, how we compute real production costs, and how we make money. This page is the contract behind every number on the site.

## 01 · Sourced: Every fact links to a primary source

Pricing pages, docs, changelogs. A fact without a source URL and a last-verified date does not render publicly. Where a number is a vendor claim (latency figures, for example), it is labeled as a claim on the page until we can measure it independently. Facts still flagged as sample data never render in production at all; you see "verification pending" instead.

## 02 · Fresh: Re-verified on a weekly cycle

Every number carries the date we last checked it, and the "What changed recently" section on each comparison is generated from the fact history, not written by hand. Stale facts get flagged for review before they get republished.

## 03 · Honest: We earn either way, so verdicts stay honest

vsref may earn a commission when you start a trial through our links, and every such link is marked and disclosed on the page it appears on. Programs exist on both sides of most comparisons, so no verdict is worth more to us than its opposite. As a structural check, no single platform should win more than 45% of its matchup verdicts site-wide; the live numbers are published below, including when they breach that line.

## 04 · Costed: True cost, not sticker price

Advertised per-minute rates rarely survive contact with production. Our cost tiers model all-in monthly spend (platform fee + per-minute + telephony) at 1K, 10K, and 100K minutes per month, using the measured real cost range where we have one.

## 05 · Normalized: One unit per category, converted in code

Vendors price the same work in different units: per character, per credit, per second, or a subscription with a quota. We normalize every rate to one canonical unit per category with deterministic math, never estimates. For text-to-speech that unit is dollars per 1M characters and dollars per audio minute, assuming ~950 characters of English text per minute of finished audio; subscription plans resolve to plan fee divided by included quota, plus published overage. For speech-to-text the unit is dollars per audio minute (and per 1,000 minutes), batch and streaming kept separate: per-hour rates divide by 60, per-second rates multiply by 60, and toggling an add-on like diarization or PII redaction adds its published per-minute rate, or renders as not computable when the vendor prices it but publishes no rate. Credit systems are converted only when the vendor states the credit-to-unit mapping, and a plan quota we cannot convert honestly renders as not computable rather than a guess.

## 06 · Reviewed: Who reviews this

We publish under the vsref Editorial name, not a personal byline, and we will not invent one. Comparisons and verdicts are drafted from the verified fact layer by the verdict engine, then published through an editorial review flow: a publish guard checks the generated prose against the same sourced, dated facts the tables show and blocks anything that fails the quality gate. That is what "reviewed by vsref Editorial" means on each page, the process is accountable to a primary source rather than to an opinion. If and when a named human reviewer joins, they are credited here by name; until then vsref carries the editorial responsibility as a publisher.

## 07 · In the open: Verdict balance, live

Win rate per platform across all published use-case verdicts, straight from the same database the pages render from. A platform over the 45% line is a signal to us to widen its matchup coverage, not to soften verdicts.

| Platform | Verdicts won | Win rate | Over 45%? |
| --- | --- | --- | --- |
| a0.dev | 0 of 7 | 0% | no |
| Abby Connect | 1 of 6 | 17% | no |
| Agorapulse | 7 of 11 | 64% | yes, flagged |
| Aider | 12 of 21 | 57% | yes, flagged |
| Akool | 2 of 12 | 17% | no |
| Alex (fka Apriora) | 1 of 7 | 14% | no |
| Amazon Polly | 2 of 14 | 14% | no |
| Amazon Transcribe | 4 of 9 | 44% | no |
| Amp | 3 of 7 | 43% | no |
| Argil | 5 of 18 | 28% | no |
| Arini | 0 of 7 | 0% | no |
| ASAPP | 0 of 6 | 0% | no |
| AssemblyAI | 25 of 43 | 58% | yes, flagged |
| Assort Health | 1 of 6 | 17% | no |
| Augment Code | 3 of 7 | 43% | no |
| Avoca | 3 of 14 | 21% | no |
| Avoma | 19 of 28 | 68% | yes, flagged |
| Azure Speech | 19 of 24 | 79% | yes, flagged |
| Azure AI Speech (STT) | 5 of 13 | 39% | no |
| Base44 | 12 of 21 | 57% | yes, flagged |
| Bland AI | 36 of 53 | 68% | yes, flagged |
| Bluedot | 9 of 14 | 64% | yes, flagged |
| Bolna | 1 of 4 | 25% | no |
| bolt.diy | 5 of 13 | 39% | no |
| Bolt.new | 22 of 48 | 46% | yes, flagged |
| BrightHire | 3 of 7 | 43% | no |
| Brooke.ai | 0 of 7 | 0% | no |
| Bubble | 6 of 21 | 29% | no |
| Buffer | 23 of 44 | 52% | yes, flagged |
| CAMB.AI | 2 of 3 | 67% | yes, flagged |
| Canva Code | 0 of 7 | 0% | no |
| Captions | 7 of 18 | 39% | no |
| Cartesia | 25 of 38 | 66% | yes, flagged |
| Cartesia Ink | 1 of 7 | 14% | no |
| CatDoes | 4 of 7 | 57% | yes, flagged |
| ChatAndBuild | 0 of 7 | 0% | no |
| Chorus by ZoomInfo | 7 of 21 | 33% | no |
| Circleback | 5 of 7 | 71% | yes, flagged |
| Clari Copilot | 1 of 21 | 5% | no |
| Claude Code | 61 of 133 | 46% | yes, flagged |
| Cline | 9 of 14 | 64% | yes, flagged |
| Cognigy (NiCE) | 2 of 12 | 17% | no |
| Cohere Transcribe | 1 of 7 | 14% | no |
| Colibri.ai | 0 of 7 | 0% | no |
| Colossyan | 5 of 18 | 28% | no |
| Command Code | 4 of 7 | 57% | yes, flagged |
| ContentStudio | 3 of 5 | 60% | yes, flagged |
| ConverzAI | 1 of 7 | 14% | no |
| Microsoft Copilot in Teams | 1 of 7 | 14% | no |
| CoSchedule | 1 of 4 | 25% | no |
| Cosine | 2 of 7 | 29% | no |
| CosyVoice | 1 of 4 | 25% | no |
| Create.xyz | 5 of 14 | 36% | no |
| Creatify | 2 of 12 | 17% | no |
| Cresta | 3 of 12 | 25% | no |
| Crush | 0 of 7 | 0% | no |
| Cursor | 29 of 77 | 38% | no |
| D-ID | 16 of 48 | 33% | no |
| Dasha | 2 of 24 | 8% | no |
| Decagon | 3 of 6 | 50% | yes, flagged |
| DeepBrain AI (AI Studios) | 10 of 18 | 56% | yes, flagged |
| Deepgram | 50 of 95 | 53% | yes, flagged |
| Deepgram Aura-2 | 3 of 10 | 30% | no |
| Deepgram Voice Agent API | 0 of 5 | 0% | no |
| DeepSeek Harness (dsh) | 3 of 7 | 43% | no |
| Dentina | 4 of 7 | 57% | yes, flagged |
| Devin | 8 of 28 | 29% | no |
| Devin Desktop (formerly Windsurf) | 4 of 21 | 19% | no |
| Dia / Dia2 | 1 of 6 | 17% | no |
| Dialzara | 1 of 6 | 17% | no |
| Dograh | 0 of 6 | 0% | no |
| Domu | 1 of 7 | 14% | no |
| Dualite | 1 of 7 | 14% | no |
| Dyad | 4 of 14 | 29% | no |
| Elai.io | 3 of 12 | 25% | no |
| ElevenLabs | 32 of 55 | 58% | yes, flagged |
| ElevenLabs Agents | 29 of 57 | 51% | yes, flagged |
| ElevenLabs Scribe | 8 of 22 | 36% | no |
| Emergent | 4 of 7 | 57% | yes, flagged |
| Factory | 5 of 7 | 71% | yes, flagged |
| Famous.ai | 0 of 7 | 0% | no |
| Fathom | 19 of 49 | 39% | no |
| Fellow | 2 of 7 | 29% | no |
| Figma Make | 11 of 14 | 79% | yes, flagged |
| Fireflies.ai | 105 of 119 | 88% | yes, flagged |
| Fish Audio | 5 of 11 | 46% | yes, flagged |
| Fish Speech | 3 of 4 | 75% | yes, flagged |
| Floot | 1 of 7 | 14% | no |
| FlutterFlow | 2 of 7 | 29% | no |
| Freebuff | 4 of 7 | 57% | yes, flagged |
| Gemini CLI | 3 of 7 | 43% | no |
| Google Gemini Live API | 0 of 6 | 0% | no |
| Gemini in Google Meet (Take notes for me) | 4 of 14 | 29% | no |
| GitHub Copilot | 34 of 56 | 61% | yes, flagged |
| GitLab Duo Agent Platform | 2 of 7 | 29% | no |
| Gladia | 4 of 21 | 19% | no |
| Gong | 20 of 49 | 41% | no |
| Goodcall | 1 of 3 | 33% | no |
| Google Antigravity | 15 of 21 | 71% | yes, flagged |
| Google Cloud Speech-to-Text | 0 of 19 | 0% | no |
| Google Cloud TTS | 8 of 18 | 44% | no |
| Goose | 3 of 7 | 43% | no |
| OpenAI gpt-4o-transcribe | 3 of 6 | 50% | yes, flagged |
| gptme | 1 of 7 | 14% | no |
| Grain | 10 of 21 | 48% | yes, flagged |
| Granola | 20 of 49 | 41% | no |
| Grok Build | 3 of 7 | 43% | no |
| xAI Grok Speech-to-Text | 0 of 4 | 0% | no |
| Groq (hosted Whisper) | 4 of 7 | 57% | yes, flagged |
| Hedra | 8 of 36 | 22% | no |
| Hello Patient | 1 of 6 | 17% | no |
| HeyBoss | 2 of 7 | 29% | no |
| HeyGen | 66 of 78 | 85% | yes, flagged |
| Higgsfield (LipSync Studio) | 4 of 6 | 67% | yes, flagged |
| Hootsuite | 0 of 28 | 0% | no |
| Hostinger Horizons | 2 of 7 | 29% | no |
| Hume EVI | 5 of 12 | 42% | no |
| Hypefury | 2 of 12 | 17% | no |
| IBM watsonx Speech to Text | 0 of 4 | 0% | no |
| Inworld TTS | 6 of 12 | 50% | yes, flagged |
| Jamie | 6 of 14 | 43% | no |
| Jiminny | 6 of 14 | 43% | no |
| JoggAI | 4 of 12 | 33% | no |
| Jump | 0 of 7 | 0% | no |
| JetBrains Junie | 2 of 7 | 29% | no |
| Kilo Code | 6 of 14 | 43% | no |
| Kimi Code | 3 of 7 | 43% | no |
| Kiro | 5 of 14 | 36% | no |
| Kore.ai | 2 of 6 | 33% | no |
| Krisp | 3 of 7 | 43% | no |
| Later | 3 of 22 | 14% | no |
| Laxis | 3 of 14 | 21% | no |
| Letta Code | 3 of 7 | 43% | no |
| Level AI | 4 of 6 | 67% | yes, flagged |
| lipsync.studio | 1 of 6 | 17% | no |
| LiveKit Agents | 7 of 29 | 24% | no |
| LMNT | 1 of 6 | 17% | no |
| Loman AI | 2 of 7 | 29% | no |
| Loomly | 1 of 6 | 17% | no |
| Lovable | 50 of 84 | 60% | yes, flagged |
| MakeUGC | 4 of 12 | 33% | no |
| MeetGeek | 4 of 7 | 57% | yes, flagged |
| Meku | 0 of 7 | 0% | no |
| Metaview | 3 of 14 | 21% | no |
| Metricool | 1 of 4 | 25% | no |
| Millis AI | 5 of 24 | 21% | no |
| MiniMax Speech | 2 of 6 | 33% | no |
| Mistral Vibe | 4 of 7 | 57% | yes, flagged |
| Mixpost | 6 of 11 | 55% | yes, flagged |
| Moonshine | 3 of 7 | 43% | no |
| Murf API | 6 of 15 | 40% | no |
| Netic | 0 of 7 | 0% | no |
| Notta | 5 of 7 | 71% | yes, flagged |
| Numa | 3 of 7 | 43% | no |
| NVIDIA Parakeet / Riva | 5 of 12 | 42% | no |
| Observe.AI | 1 of 6 | 17% | no |
| Omilia | 1 of 6 | 17% | no |
| Onlook | 2 of 7 | 29% | no |
| OpenAI Codex | 15 of 28 | 54% | yes, flagged |
| OpenAI Realtime API | 3 of 22 | 14% | no |
| OpenAI TTS | 6 of 27 | 22% | no |
| OpenCode | 16 of 21 | 76% | yes, flagged |
| OpenHands | 3 of 7 | 43% | no |
| Otter.ai | 72 of 112 | 64% | yes, flagged |
| Pallyy | 4 of 6 | 67% | yes, flagged |
| Parloa | 5 of 12 | 42% | no |
| Patter | 1 of 6 | 17% | no |
| Phonely | 13 of 30 | 43% | no |
| Pi | 3 of 7 | 43% | no |
| Pipecat | 5 of 17 | 29% | no |
| Planable | 6 of 12 | 50% | yes, flagged |
| PolyAI | 3 of 12 | 25% | no |
| Postiz | 7 of 26 | 27% | no |
| Publer | 8 of 15 | 53% | yes, flagged |
| Qwen Code | 3 of 7 | 43% | no |
| Qwen3-ASR | 4 of 7 | 57% | yes, flagged |
| Read AI | 3 of 14 | 21% | no |
| RecurPost | 7 of 12 | 58% | yes, flagged |
| Reflex Build | 1 of 7 | 14% | no |
| Replicant | 3 of 6 | 50% | yes, flagged |
| Replit | 13 of 35 | 37% | no |
| Retell AI | 41 of 58 | 71% | yes, flagged |
| Rev AI | 1 of 9 | 11% | no |
| Rime | 5 of 11 | 46% | yes, flagged |
| Ringly.io | 6 of 22 | 27% | no |
| Rocket.new | 9 of 14 | 64% | yes, flagged |
| Rork | 11 of 21 | 52% | yes, flagged |
| Rosie | 1 of 11 | 9% | no |
| SadTalker | 5 of 6 | 83% | yes, flagged |
| Salient | 2 of 7 | 29% | no |
| Sameday | 2 of 7 | 29% | no |
| Sembly AI | 5 of 14 | 36% | no |
| Sendible | 4 of 6 | 67% | yes, flagged |
| Shhots AI | 2 of 6 | 33% | no |
| Sierra | 2 of 6 | 33% | no |
| Slang AI | 0 of 7 | 0% | no |
| Smallest.ai Pulse | 2 of 6 | 33% | no |
| Smith.ai | 3 of 9 | 33% | no |
| SocialBee | 1 of 16 | 6% | no |
| SocialPilot | 2 of 5 | 40% | no |
| Softgen | 0 of 7 | 0% | no |
| Softr | 3 of 7 | 43% | no |
| Soniox | 3 of 12 | 25% | no |
| Speechify API | 1 of 6 | 17% | no |
| Speechmatics | 3 of 12 | 25% | no |
| Spinach | 1 of 7 | 14% | no |
| Sprout Social | 4 of 16 | 25% | no |
| STELLA Automotive AI | 4 of 14 | 29% | no |
| Supernormal | 1 of 7 | 14% | no |
| Superpowered | 0 of 7 | 0% | no |
| Sybill | 5 of 21 | 24% | no |
| Sync.so | 11 of 18 | 61% | yes, flagged |
| Synthesia | 26 of 42 | 62% | yes, flagged |
| Tabby | 3 of 7 | 43% | no |
| Tabnine | 2 of 7 | 29% | no |
| Tactiq | 3 of 7 | 43% | no |
| Tavus | 10 of 24 | 42% | no |
| Telnyx Voice AI Agents | 6 of 6 | 100% | yes, flagged |
| Tempo | 3 of 7 | 43% | no |
| Thoughtly | 9 of 23 | 39% | no |
| tl;dv | 3 of 28 | 11% | no |
| Toma | 0 of 14 | 0% | no |
| Trae | 2 of 7 | 29% | no |
| Twilio ConversationRelay | 1 of 10 | 10% | no |
| Typefully | 2 of 12 | 17% | no |
| Ultravox | 1 of 24 | 4% | no |
| Upfirst | 2 of 6 | 33% | no |
| v0 | 14 of 42 | 33% | no |
| Val Town (Townie) | 0 of 7 | 0% | no |
| Vapi | 65 of 95 | 68% | yes, flagged |
| VEED | 5 of 12 | 42% | no |
| Verdent | 1 of 7 | 14% | no |
| Vidnoz | 4 of 6 | 67% | yes, flagged |
| Vista Social | 1 of 5 | 20% | no |
| Vocode | 8 of 23 | 35% | no |
| Voiceflow | 13 of 24 | 54% | yes, flagged |
| Mistral Voxtral Transcribe | 10 of 20 | 50% | yes, flagged |
| Voxtral TTS | 5 of 16 | 31% | no |
| Warp | 3 of 7 | 43% | no |
| Wav2Lip | 0 of 6 | 0% | no |
| OpenAI Whisper (API) | 13 of 68 | 19% | no |
| Wudpecker | 0 of 7 | 0% | no |
| Yepic AI | 1 of 6 | 17% | no |
| YouWare | 2 of 7 | 29% | no |
| Zed | 4 of 7 | 57% | yes, flagged |
| Zoom AI Companion | 11 of 28 | 39% | no |
| Zoom Virtual Agent Receptionist | 1 of 5 | 20% | no |

Small denominators swing hard: a platform at 2 of 3 reads as 67% until more matchups publish. The flag drives coverage, and we leave it visible on purpose.

Tool logos are the property of their respective owners and are shown here for identification only; their use does not imply any affiliation or endorsement.

Source: https://www.versusref.com/methodology/
