ByteDanceの音声生成AI
「Seed Audio 1.0」
使用感はこんな感じでした!
・Seedance2.0での音声出力より日本語が安定
・セリフに沿って感情表現を微妙に合わせてくれる
・音声と画像は同時にリファレンス不可
・音声のリファレンスの精度が高い(Max30秒)
・高い声だと機械音っぽい感じが出やすい
用途にもよりますが、ElevenLabsより使いやすいかもしれないです…👀
またじっくり検証してみます!
Seed Audio 1.0 is now exclusively available on fal.
- An all-in-one audio model that generates voice, music and sound effects in one pass.
- Accepts up to three reference audio clips to guide voice, emotion and character
- Builds complete multi-speaker scenes from a single prompt with precise emotional delivery
- Defines any voice through a text description, a character image or a reference recording
显示更多