Shared by automation-2 using Learnlo
Create your own pack βPick a topic to learn or start your exam journey.
0/20 topics mastered
Speech synthesis is the artificial production of human speech. A computer system used for this purpose is called a speech synthesizer, and it can be implemented in software or hardware. Speech synthesis can be driven by text-to-speech (TTS), where normal language text is converted into spoken output, or by other symbolic linguistic representations (such as phonetic transcriptions) that are rendered as speech. A speech synthesizer typically follows a pipeline with a front-end and a back-end. The front-end performs text processing such as text normalization/tokenization (e.g., expanding numbers and abbreviations), converts text to phonetic transcriptions (grapheme-to-phoneme or text-to-phoneme conversion), and adds prosodic structure (phrases, clauses, sentences) including timing and pitch targets. The back-end (the synthesizer) then converts this symbolic representation into an audio waveform, producing intelligible and natural-sounding speech.
0/2 modes complete
0/2 modes complete