Introducing Song Creator Pro — create music with AI, locally on your device. Try it now →

Try MOSS-TTS-Nano Online for Free

Multilingual text-to-speech that runs entirely in your browser. No signup, no install, completely free.

Try MOSS-TTS-Nano Free
18 Languages
100% Private
Free and Unlimited

Why MOSS

Why Use MOSS-TTS-Nano

MOSS-TTS-Nano is a compact, open-source TTS model with just 100 million parameters, released under the Apache 2.0 license by MOSI.AI and the OpenMOSS team. Despite its small size (~715 MB), it delivers multilingual voice cloning in 18 languages with 48kHz stereo output.

MOSS pairs compact size with multilingual reach, making it a strong choice for projects that need natural speech in more than one language.

Multilingual by Design

Built for speech across languages including English, Chinese, Japanese, Korean, German, Spanish, French, Arabic, and more.

48kHz Stereo Output

Higher quality than most TTS models. Produces rich, natural-sounding stereo audio at 48kHz sample rate.

Real-Time Streaming

Hear audio as it generates with streaming output. Designed to run in real time, even on CPU.

Runs in Your Browser

No download, no server, no signup. Runs locally via WebGPU/WASM with full privacy.

Get Started

How It Works

MOSS-TTS-Nano works with built-in voices: pick a voice, type your text, and generate.

1

Open the free tool and choose your mode

Go to the MOSS-TTS-Nano tool in your browser and pick a built-in voice. No download or signup required.

2

Type your text

Enter the text you want spoken. Built-in voices support English, Chinese, and Japanese.

3

Generate and listen

Hit generate and your audio streams in real time. Download the result as a high-quality 48kHz stereo audio file.

Use Cases

Who Is MOSS-TTS-Nano For?

Multilingual Content Creators

Create voiceovers in multiple languages. Produce content in English, Chinese, and Japanese with consistent built-in voices.

Localization Teams

Generate localized audio for apps, games, and media. For a consistent cloned brand voice across regions, Voice Creator Pro Cloud offers multilingual cloning.

Developers and Researchers

Prototype multilingual voice features, test TTS pipelines, or experiment with a compact open-source model you can inspect and build on. Apache 2.0 licensed for maximum flexibility.

Accessibility

Create personalized synthetic voices for people who speak different languages, helping them communicate in their native language with a familiar voice.

Getting the Most Out of MOSS

Tips for Best Results

Try each built-in voice

MOSS includes built-in voices for English, Chinese, and Japanese. Try different voices to find the best match for your content.

Try different built-in voices for each language

MOSS includes built-in voices for English, Chinese, and Japanese. Try different voices to find the best match for your content before turning to cloning.

Streaming works best in Chrome

For the smoothest real-time streaming experience, use Chrome or another Chromium-based browser. Firefox and Safari may have limited support for some features.

Keep sentences natural

Conversational, sentence-length text produces the most natural pacing. Break very long passages into paragraphs.

Match text to a supported language

Built-in voice TTS is available in English, Chinese, and Japanese. For other languages, Voice Creator Pro Cloud covers 600+.

Voice Creator Pro

Need more languages, speed, or a commercial license?

Voice Creator Pro gives you higher quality models, voice cloning in 600+ languages, emotional speech, and a commercial use license. The Cloud free plan includes 1 hour 45 minutes of text to speech with basic voices or 10 minutes of voice cloning every month, no card required.

Voice Cloning in 600+ Languages

Go beyond 18 languages with additional open-source models for TTS across 600+ languages

No GPU Required

Generation runs on our hardware, so it works on any laptop, tablet, or phone

Voice Design from Text

Describe a voice in plain text and the AI creates it. No audio samples needed

Advanced Voice Cloning

Multiple cloning models to choose from, clone any voice from a few seconds of audio

Emotional Speech

13 emotions at 5 intensities with Qwen3-TTS

Commercial License

Full rights to use generated audio in commercial projects

FAQ

Common Questions

MOSS-TTS-Nano is a compact, open-source text-to-speech model with 100 million parameters. In the free tool it powers multilingual text-to-speech with built-in voices, running entirely in your browser.

MOSS-TTS-Nano was developed by MOSI.AI and the OpenMOSS team. It is released under the Apache 2.0 license, making it freely available for both personal and commercial use.

The free tool's built-in voices support English, Chinese, and Japanese. The underlying MOSS model family is built for 18 languages. For wider coverage, Voice Creator Pro Cloud spans 600+ languages.

The MOSS model family supports voice cloning, but the free browser tool focuses on text-to-speech with built-in voices. Voice cloning is part of Voice Creator Pro Cloud, with a free plan of 5,000 tokens every month, up to 10 minutes of voice cloning.

MOSS-TTS-Nano produces 48kHz stereo audio, which is higher quality than many other TTS models that output mono audio at 22kHz or 24kHz. The result is richer, more natural-sounding speech.

Yes. MOSS-TTS-Nano is open-source software released under the Apache 2.0 license. On this site, it runs directly in your browser with no account, no signup, and no usage limits.

MOSS-TTS-Nano is a smaller, more efficient version designed to run in real time on CPUs and in the browser. The full MOSS-TTS model is larger and may produce higher-fidelity output in some cases, but MOSS-TTS-Nano retains the core capabilities including multilingual voice cloning and 48kHz stereo output.

Chrome, Edge, and other Chromium-based browsers work best. Firefox and Safari have limited support for the WebGPU and WASM features that MOSS-TTS-Nano uses for acceleration. For the best streaming experience, use Chrome.