MODEL · OPENMOSS / MOSI.AI · 100M (0.1B)
MOSS-TTS-Nano (100M)
A 100M streaming-TTS that closes the multilingual gap Kokoro doesn't cover — 20 languages including English, Japanese, Korean, Spanish, French, Arabic, Mandarin, plus voice cloning from a short audio reference. 48 kHz stereo output, neural-audio-tokenizer + autoregressive LLM pipeline, runs real-time on 4 CPU cores. The ONNX build drops PyTorch entirely and gets ~2× the inference efficiency of the original.
License: Apache 2.0 · Context: n/a · Released: April 10, 2026 (PyTorch); April 17, 2026 (ONNX-CPU port)
The decision in five lines
- The call
- Consider — runnable locally, family reference
- Best for
- Local evaluation and family reference
- Runs on
- 23 hardware picks fit (cheapest: Intel Arc B580 12 GB · $310)
- Watch out
- English-only narration where you don't need cloning — Kokoro-82M is the proven default, ranks #1 on TTS Arena, and is even smaller.
- Evidence
- Estimated
- 100M (0.1B)
- PARAMETERS
- TTS + MULTILINGUAL VOICE CLONE
- TYPE
- —
- CONTEXT
- <400 MB (CPU-only, 4 cores enough)
- VRAM AT Q4
Where we recommend this
This model isn’t currently in an active planner slot. See the runner notes below if you’re running it anyway.
The call
A 100M streaming-TTS that closes the multilingual gap Kokoro doesn't cover — 20 languages including English, Japanese, Korean, Spanish, French, Arabic, Mandarin, plus voice cloning from a short audio reference. 48 kHz stereo output, neural-audio-tokenizer + autoregressive LLM pipeline, runs real-time on 4 CPU cores. The ONNX build drops PyTorch entirely and gets ~2× the inference efficiency of the original.
When not to use: English-only narration where you don't need cloning — Kokoro-82M is the proven default, ranks #1 on TTS Arena, and is even smaller. Use MOSS-TTS-Nano when you actually need multilingual coverage or voice cloning on hardware too small for Chatterbox/VoxCPM2.
Runner notes
GitHub `OpenMOSS/MOSS-TTS-Nano` for PyTorch path; HuggingFace `OpenMOSS-Team/MOSS-TTS-Nano-100M-ONNX` for the CPU-friendly route. Companion `MOSS-Audio-Tokenizer-Nano-ONNX` handles the audio tokenizer. No Ollama path (non-LM). Sibling MOSS-TTSD-v0.5 (2B, ZH/EN dialogue) covers multi-speaker if you outgrow Nano. Step-up: `OpenMOSS-Team/MOSS-TTS-v1.5` (8B, Apache 2.0, May 2026) is the flagship voice-cloning tier — community-reported to beat Fish Audio S2 Pro / Qwen3-TTS on English cloning; runs ~11 GB at Q4 on a 24 GB GPU. Use v1.5 when you have the VRAM and want top quality; Nano for CPU/edge.
Hardware that fits
Every hardware pick whose memory fits this model at the quant we recommend. Sorted cheapest-first — the top row is your best-value fit. Click through for the full buyer’s guide.
- Intel Arc B580 12 GBPerfect · 8.5× 12 GB · $310–$429 new / $190–$260 refurb
- NVIDIA RTX 3060 12 GBPerfect · 8.5× 12 GB · $480–$660
- Minisforum UM890 ProPerfect · 17.1× 32 GB DDR5 (shared) · ~$580 barebone (no RAM / SSD); ~2× that built with 32 GB — every variant sold out on Minisforum's store Sep 7 2026
- RTX 5060 Ti 16 GBPerfect · 11.4× 16 GB · $700 (refurb / open box) – $970 (new)
- AMD Radeon RX 9070 XTPerfect · 11.4× 16 GB · $750–$870
- AMD Radeon RX 7900 XTXPerfect · 17.1× 24 GB · $850 refurb / $1,400–$1,500 new
- Mac Mini M4 16 GBPerfect · 7.6× 16 GB unified · $899 (M6 successor, 16 GB / 256 GB, pre-order) / M4 residuals $499–$799
- NVIDIA RTX 3090 (used, single)Perfect · 17.1× 24 GB · $950–$1,200
- NVIDIA RTX 5070 TiPerfect · 11.4× 16 GB · $1,150–$1,300 new; open-box / refurb from ~$1,020
- NVIDIA RTX 5080Perfect · 11.4× 16 GB · $1,430–$1,700
- MacBook Air M5 24 GBPerfect · 11.4× 24 GB unified · $1,499–$1,899
- Mac Mini M4 Pro 24 GBPerfect · 11.4× 24 GB unified · $1,699 (M5 Pro successor, 24 GB / 512 GB, pre-order; M4 Pro discontinued Aug 25 2026)
- Dual RTX 3090 (used)Perfect · 34.2× 48 GB · $1,800–$2,500 all-in
- NVIDIA RTX 4090Perfect · 17.1× 24 GB · $2,200–$2,800
- M5 Pro MacBook Pro 48 GBPerfect · 22.9× 48 GB unified · $2,999–$3,599
- Framework Desktop (Ryzen AI Max+ 395)Perfect · 61.0× 128 GB unified · $3,449 (128 GB config)
- NVIDIA RTX A6000 (48 GB, used)Perfect · 34.2× 48 GB ECC · $3,500–$4,500
- Mac Studio M4 Max 64 GBPerfect · 30.5× 64 GB unified · $3,799 (M5 Max successor, 64 GB / 1 TB, pre-order; M4 Max discontinued Aug 25 2026)
- NVIDIA DGX SparkPerfect · 61.0× 128 GB unified · $4,699
- M5 Max MacBook Pro 64 GBPerfect · 30.5× 64 GB unified · ~$5,199 (est.; June 25 2026 increase)
- NVIDIA RTX 5090Perfect · 22.8× 32 GB · $5,300–$5,500 (new, in stock)
- Mac Studio M3 Ultra 96 GBPerfect · 45.8× 96 GB unified · $5,499 (M5 Ultra successor, 96 GB / 1 TB, pre-order; M3 Ultra discontinued Aug 25 2026)
- Dual RTX 5090Perfect · 45.5× 64 GB (2×32) · $11,000–$11,700 all-in (two cards alone are ~$10,600 at the Sep 2026 floor)
Next step
Find-by-model — see what hardware runs this→