# MegaTTS3 Review

> MegaTTS3 review: Research-grade Apache-2.0 TTS whose practical cloning is gated: the WaveVAE encoder is not released, so users must submit audio to ByteDance.

![MegaTTS3 logo](https://www.versusref.com/logos/megatts3.png)

Research-grade Apache-2.0 TTS whose practical cloning is gated: the WaveVAE encoder is not released, so users must submit audio to ByteDance channels to obtain pre-extracted speaker latents (.npy) for cloning.

Among the 46 text-to-speech tools we track, MegaTTS3 has the 38th-widest language coverage.

See pricing · Open source · facts verified Jul 20, 2026

## What we know about MegaTTS3

MegaTTS3 sits in the text-to-speech apis category, where it is research-grade Apache-2.0 TTS whose practical cloning is gated: the WaveVAE encoder is not released, so users must submit audio to ByteDance channels to obtain pre-extracted speaker latents (.npy) for cloning. We keep this MegaTTS3 profile grounded in primary sources, each fact dated to when we last confirmed it.

MegaTTS3 does not publish a public per-minute rate we have been able to verify, so the price figure here stays blank until we can confirm one. We would rather show nothing than a guessed number; if you have a current MegaTTS3 quote, it is the fastest way for us to close that gap.

On capabilities, MegaTTS3 covers instant voice cloning, emotion / style controls, and self-host / on-prem option. Each of those is verified against MegaTTS3's own docs or dashboard, not marketing copy.

MegaTTS3 ranks the 38th-widest language coverage of the 46 text-to-speech tools we track, so where it lands for you depends on which of those matters more.

In total we track 11 verified facts for MegaTTS3 today, each linking the primary source it came from so you can check our work - and vendor claims we have not measured ourselves are labeled as such on the MegaTTS3 fact sheet below.

## Fact sheet

### Capabilities

| Fact | Value | Verified | Source |
| --- | --- | --- | --- |
| Instant voice cloning | ✓  Yes | Jul 20 | [source](https://github.com/bytedance/MegaTTS3) |
| Languages supported | 2 languages | Jul 20 | [source](https://github.com/bytedance/MegaTTS3) |
| Emotion / style controls | ✓  Yes | Jul 20 | [source](https://github.com/bytedance/MegaTTS3) |

### Compliance & trust

| Fact | Value | Verified | Source |
| --- | --- | --- | --- |
| Self-host / on-prem option | ✓  Yes | Jul 20 | [source](https://github.com/bytedance/MegaTTS3) |
| Model weights license | Apache-2.0 | Jul 20 | [source](https://github.com/bytedance/MegaTTS3) |

### Build experience

| Fact | Value | Verified | Source |
| --- | --- | --- | --- |
| Official SDKs | Python (CLI + Gradio WebUI) | Jul 20 | [source](https://github.com/bytedance/MegaTTS3) |
| Output formats | WAV | Jul 20 | [source](https://github.com/bytedance/MegaTTS3) |

### Commercial

| Fact | Value | Verified | Source |
| --- | --- | --- | --- |
| Model size (parameters) | 0.45B (backbone) | Jul 20 | [source](https://github.com/bytedance/MegaTTS3) |
| Hardware to self-host | CPU inference supported (~30 s per generation at 10 steps) | Jul 20 | [source](https://github.com/bytedance/MegaTTS3) |
| Project maintenance status | active (low activity) | Jul 20 | [source](https://github.com/bytedance/MegaTTS3) |
| GitHub stars | 6,083 stars | Jul 20 | [source](https://github.com/bytedance/MegaTTS3) |

Source: https://www.versusref.com/tts/tools/megatts3/
