> For the complete documentation index, see [llms.txt](https://textopia.gitbook.io/textopia.ai/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://textopia.gitbook.io/textopia.ai/features/text-to-speech/ai-voices.md).

# AI Voices

### Explore Versatility: Tailored Synthetic Voices for Every Narrative

Immerse yourself in a world of creativity with our AI Voices feature. Tailored for various tones and styles, this groundbreaking tool offers creators a diverse library of customizable synthetic voices. From compelling narrations to dynamic character dialogues, elevate your content with unparalleled versatility and creative freedom. Step into the future of voice synthesis today.

<div align="left"><figure><img src="https://592342210-files.gitbook.io/~/files/v0/b/gitbook-x-prod.appspot.com/o/spaces%2FCHwFZt06MRSjU0aF2Fiy%2Fuploads%2FpAj8fQompF6Xe2TiXjDG%2FGemini_Generated_Image%20(11).jpeg?alt=media&amp;token=33306794-910d-4959-af6b-16845345fcab" alt=""><figcaption></figcaption></figure> <figure><img src="https://592342210-files.gitbook.io/~/files/v0/b/gitbook-x-prod.appspot.com/o/spaces%2FCHwFZt06MRSjU0aF2Fiy%2Fuploads%2Fzjqp8KNM5cKQU3gasYps%2FGemini_Generated_Image%20(12).jpeg?alt=media&amp;token=bff5d6d7-98af-45fc-a0b2-54505cdbce23" alt=""><figcaption></figcaption></figure></div>

$$
\mathcal{V}(\text{Narrative}, \text{Style}; \Theta) = \text{NN}*{\text{Synthesis}}(\text{NN}*{\text{Text2Speech}}(\text{Narrative}, \alpha\_{\text{Narrative}}) \oplus \text{NN}*{\text{StyleEmbed}}(\text{Style}, \alpha*{\text{Style}}); \Theta) \
$$

* $$\text{Style}$$ represents the synthesized voice output tailored for a specific narrative and style.
* $$\text{Narrative}$$ is the narrative content for which the voice is synthesized.
* $$\text{Style}$$ represents the desired style or tone for the synthesized voice.
* $$ΘΘ$$ represents the parameters of the entire synthesis model.
* $$\text{NN}\_{\text{Synthesis}}$$ denotes the neural network model responsible for voice synthesis.
* $$\text{NN}\_{\text{Text2Speech}}$$ represents the neural network model for converting text input into speech.
* $$\text{NN}\_{\text{StyleEmbed}}$$ represents the neural network model for embedding the style features.
* $$\alpha\_{\text{Narrative}}$$ and $$\alpha\_{\text{Style}}$$ are the parameters for narrative and style embeddings respectively.
* $$⊕⊕$$ denotes concatenation operation, combining the narrative and style embeddings.
