xcLee001/SonicVale

一个开源的多角色、多情绪 AI 配音生成平台,支持小说、剧本、视频等内容的自动配音与导出。

What it solves

SonicVale (音谷) is designed to automate the process of creating multi-character, multi-emotion voiceovers for novels, scripts, and videos. It eliminates the manual effort of assigning voices and emotions to individual lines of dialogue in long-form content.

How it works

The platform integrates Large Language Models (LLMs) and Text-to-Speech (TTS) engines to transform text into audio. It uses LLMs to automatically split dialogue and assign roles, and leverages the Index-TTS-2.0 engine for emotional voice synthesis. Users can import text, manage a library of character voices, bind specific emotions to roles, and perform precise audio editing (such as adding silence or deleting segments) before exporting the final result.

Who it’s for

Content creators, authors, and video producers who need high-quality, emotionally expressive AI voiceovers for narrative-driven content with multiple characters.

Highlights

  • Automated Dialogue Splitting: Uses customizable prompts and LLMs to break down scripts into character-specific lines.
  • Multi-Emotion Support: Integrates Index-TTS-2.0 to provide nuanced emotional tones for different characters.
  • Character Management: Features a dedicated role library to maintain consistent voice bindings across multiple chapters.
  • Precision Audio Editing: Allows users to manually delete audio fragments or insert silence for better pacing.
  • Flexible Integration: Supports any LLM compatible with the OpenAI API protocol.

Related

  • Project
  • Project
  • Project
  • Project
  • Project