uncloseai.com/public/tts/multilingual-cpu.html

182 lines
8.6 KiB
HTML

<!--
This is free software for the public good of a permacomputer hosted at
permacomputer.com, an always-on computer by the people, for the people.
One which is durable, easy to repair, & distributed like tap water
for machine learning intelligence.
The permacomputer is community-owned infrastructure optimized around
four values:
TRUTH First principles, math & science, open source code freely distributed
FREEDOM Voluntary partnerships, freedom from tyranny & corporate control
HARMONY Minimal waste, self-renewing systems with diverse thriving connections
LOVE Be yourself without hurting others, cooperation through natural law
This software contributes to that vision by making machine learning
accessible to everyone through a free, open, embeddable chat interface.
Code is seeds to sprout on any abandoned technology.
-->
<!DOCTYPE html>
<html lang="en">
<head>
<meta charset="UTF-8">
<meta name="viewport" content="width=device-width, initial-scale=1.0">
<meta name="theme-color" content="#43a047">
<meta name="color-scheme" content="light dark">
<title>Silero TTS | uncloseai-speech | uncloseai.com</title>
<meta name="description" content="Silero TTS: 148 voices across 5 languages, all on CPU. PyTorch, 48kHz output. Self-host with uncloseai-speech.">
<link rel="stylesheet" href="/css/pico.classless.min.css">
<link rel="stylesheet" href="/css/chunkfive/stylesheet.css" type="text/css" charset="utf-8" />
<link rel="stylesheet" href="/css/sidebar-theme.css">
<link rel="stylesheet" href="https://cdnjs.cloudflare.com/ajax/libs/highlight.js/11.10.0/styles/a11y-dark.min.css" />
<script src="https://cdnjs.cloudflare.com/ajax/libs/highlight.js/11.6.0/highlight.min.js"></script>
<script>
function switchTheme(theme) {
if (theme === "auto") {
document.documentElement.removeAttribute('data-theme');
} else {
document.documentElement.setAttribute('data-theme', theme);
}
}
window.UNCLOSEAI_CUSTOM_STYLING = false;
</script>
<script src="https://uncloseai.com/uncloseai.js" type="module"></script>
</head>
<body>
<button class="sidebar-toggle" onclick="document.querySelector('.sidebar').classList.toggle('open')">
</button>
<aside class="sidebar">
<div class="table-of-contents">
<nav>
<ul>
<li><a href="/">Home</a></li>
<li><a href="/c-examples.html">C Examples</a></li>
<li><a href="/csharp-examples.html">C# Examples</a></li>
<li><a href="/dart-examples.html">Dart Examples</a></li>
<li><a href="/elixir-examples.html">Elixir Examples</a></li>
<li><a href="/go-examples.html">Go Examples</a></li>
<li><a href="/java-examples.html">Java Examples</a></li>
<li><a href="/kotlin-examples.html">Kotlin Examples</a></li>
<li><a href="/nodejs-examples.html">Node.js Examples</a></li>
<li><a href="/php-examples.html">PHP Examples</a></li>
<li><a href="/python-examples.html">Python Examples</a></li>
<li><a href="/ruby-examples.html">Ruby Examples</a></li>
<li><a href="/rust-examples.html">Rust Examples</a></li>
<li><a href="/swift-examples.html">Swift Examples</a></li>
<li><a href="/uncloseai-js.html">uncloseai.js Docs</a></li>
<li><a href="/uncloseai-js-styleguide.html">Styleguide</a></li>
<li><a href="/cli.html">uncloseai-cli</a></li>
<li><a href="/browser-toys.html">Browser Toys</a></li>
<li><a href="/inference.html">Inference Setup</a></li>
<li><a href="/text-to-speech.html">Text-to-Speech</a></li>
<li><a href="/tts/voice-cloning.html" style="padding-left:2em">Qwen3-TTS</a></li>
<li><a href="/tts/fast-synthesis.html" style="padding-left:2em">Piper TTS</a></li>
<li><a href="/tts/hd-cloning.html" style="padding-left:2em">XTTS v2</a></li>
<li><a href="/tts/multilingual-cpu.html" class="active" style="padding-left:2em">Silero TTS</a></li>
<li><a href="/tts/lightweight.html" style="padding-left:2em">Kokoro TTS</a></li>
<li><a href="/crawler.html">Our Crawler</a></li>
<li><a href="/reverse-retrieval-augmented-generations-rag.html">Reverse RAG</a></li>
<li><a href="/languages" target="_blank">All Languages</a></li>
<li><a href="https://shop.unturf.com/p/8486f492-a93e-11f0-b477-02dfe05770ee/uncloseai-machine-learning-reference-guide-to-inference-clients" target="_blank">Book</a></li>
</ul>
</nav>
</div>
</aside>
<main>
<header>
<hgroup>
<a href="https://uncloseai.com"><h1 class="unturf" style="font-family: 'ChunkFiveRegular';">uncloseai.</h1></a>
<p>Silero TTS</p>
</hgroup>
<nav>
<ul>
<li><a href="#" onclick="switchTheme('auto')">Auto</a></li>
<li><a href="#" onclick="switchTheme('light')">Light</a></li>
<li><a href="#" onclick="switchTheme('dark')">Dark</a></li>
</ul>
</nav>
</header>
<p><strong>Self-host</strong> &mdash; Not on the public endpoint. Clone the repo to use this engine.</p>
<h2>What It Does</h2>
<p>148 voices across 5 languages, all running on CPU. English, Russian, German, Spanish, and French &mdash; each language with its own set of distinct speakers.</p>
<p>The models are small and efficient. They download automatically the first time you use them, no manual setup required. Output is high-quality 48kHz audio.</p>
<p>This is the go-to engine when you need multilingual support without a GPU. The Silero team actively maintains it, and we keep it integrated and tested in the raccoon dumpster.</p>
<h2>Example</h2>
<p>Once self-hosted and enabled, it works through the same OpenAI-compatible API:</p>
<pre><code class="python">from openai import OpenAI
client = OpenAI(
api_key="not-needed",
base_url="http://localhost:8000/v1"
)
# Silero uses native voice names like en_0, en_50, ru_0, de_0
client.audio.speech.create(
model="tts-1-silero",
voice="en_50",
input="One hundred and forty-eight voices across five languages, all running on CPU. No GPU required. The raccoons found this one and it just works."
).stream_to_file("silero.mp3")</code></pre>
<h2>Available Languages</h2>
<ul>
<li><strong>English:</strong> ~120 voices (en_0 through en_117)</li>
<li><strong>Russian:</strong> ~10 voices</li>
<li><strong>German:</strong> ~10 voices</li>
<li><strong>Spanish:</strong> ~5 voices</li>
<li><strong>French:</strong> ~5 voices</li>
</ul>
<h2>Technical Details</h2>
<ul>
<li><strong>Voices:</strong> 148 across 5 languages</li>
<li><strong>Sample rate:</strong> 48kHz</li>
<li><strong>Runtime:</strong> PyTorch (torch.hub)</li>
<li><strong>Hardware:</strong> CPU only, no GPU needed</li>
<li><strong>Model size:</strong> ~50-100MB per language</li>
<li><strong>Upstream:</strong> <a href="https://github.com/snakers4/silero-models" target="_blank">Silero TTS</a>, actively maintained</li>
</ul>
<h2>Self-Hosting</h2>
<p>Models auto-download on first use. Enable by adding Silero voices to your config:</p>
<pre><code class="bash">git clone https://git.unturf.com/engineering/unturf/uncloseai-speech.git
cd uncloseai-speech
make deploy
make voices-silero</code></pre>
<p>Add Silero voices to <code>voice_to_speaker.default.yaml</code> and restart.</p>
<p><a href="/text-to-speech.html"><strong>&larr; Back to Text-to-Speech overview</strong></a></p>
<script>hljs.highlightAll();</script>
<footer>
<small>&copy; uncloseai. 2025</small>
<br>
<small>Stylesheets by <a href="https://picocss.com" target="_blank">PicoCSS</a></small>
<small>& <a href="https://highlightjs.org/" target="_blank">highlight.js</a></small>
<br>
<small><a href="/privacy-policy.html">Privacy Policy</a> | <a href="/terms-of-use.html">Terms of Use</a></small>
</footer>
</main>
</body>
</html>