クリック数
69Ranked in AIForest
読み込み中...
A lightweight text-to-speech model (4 billion parameters) capable of generating realistic and expressive speech in 9 languages, with support for various dialects. Available via API and can be tested directly in the studio
Ranked in AIForest
Directory views
音声合成
Audio, Music & Speech, Text to Speech
Voxtral TTS utilizes a 4-billion parameter model, which is designed to be more lightweight than massive LLM-based speech engines. This focus allows for realistic and expressive speech generation across nine languages and various dialects. It is particularly useful for users who need a balance between high-quality vocal output and efficient processing speeds, whether using the studio or the API.
The platform provides a dedicated studio environment where users can input text and generate speech directly in the browser. This allows you to evaluate the realism, expressiveness, and dialect accuracy of the nine supported languages. It is a recommended first step for buyers to ensure the vocal characteristics align with their specific project needs before committing to API development.
自動翻訳が利用できない場合、一部のツールの説明は英語で表示されることがあります。