情報 スタイルTTS 2
StyleTTS 2 achieves human-level text-to-speech synthesis through style diffusion and adversarial training. It generates highly natural speech that rivals real human recordings. StyleTTS 2 represents the state-of-the-art in TTS quality and naturalness.
主要な特徴
人間レベルの品質
人間の音声とは区別できない音声を生成する。
表現的スタイル
モデルの話し方はテキストと自然に変化する。
自然韻律
拡散モデルを用いた完全リズム,ストレス,音調のモデル化を行った。
スタジオ・ヴォイス
音声ライブラリの既製 StyleTTS 2 音声から選択します。
ファストインフェクション
また,自動回帰モデルよりも高速であり,品質を維持できる。
オープンソース
MITライセンスで商用利用権を持つ。
ユースケース
スタイルTTS 2 Voices
View All 6StyleTTS2 Default
ENStyleTTS2 Expressive
ENStyleTTS2 Fast
ENStyleTTS2 Natural
ENStyleTTS2 Neutral
ENStyleTTS2 Quality
EN使い方 スタイルTTS 2
-
1
無料で登録
無料でTextToSpeechAIアカウントを作成して クレジットを得る
-
2
StyleTTS2 エンジンを選択
音声ライブラリから StyleTTS2 音声を選択します。
-
3
テキストを入力
ナレーションを行うスクリプトを貼り付けまたはタイプします。 StyleTTS2 は英語に優れ、長いパートにおいて自然な韻律、強調、音調を提供します。
-
4
音声を生成
生成をクリックすると、TextToSpeechAI が StyleTTS2 オーディオを GPU 上でレンダリングします。
-
5
API をダウンロードまたは使用
完成した StyleTTS2 オーディオを MP3、WAV、OGG としてダウンロードするか、自動生成のために StyleTTS2 音声で TextToSpeechAI API を呼び出す。
スタイルTTS 2 API
Generate speech programmatically using the TextToSpeechAI REST API.
curl -X POST "https://api.texttospeechai.com/v1/generate/" \
-H "Authorization: Bearer YOUR_API_TOKEN" \
-H "Content-Type: application/json" \
-d '{
"text": "スタイルTTS2は,プロの人間の録音と競うほど自然な音声を生成する。",
"voice": "styletts2-default"
}'
よくある質問
Technical Specs
- Generation Speed Moderate
- Output Quality Excellent
- Voice Cloning Not Supported
- Languages 1
- GPU VRAM 4-6GB
- Credits/1000 chars 50