A new W3C markup standard for text-to-speech synthesis

Marcus Walker, J. Larson, Andrew J. Hunt · 2002

A new set of XML-based markup standards developed for the purpose of enabling voice browsing of the Internet are emerging from the Voice Browser working group, organized under the auspices of the W3C. Among the first in this series of specifications is the speech synthesis text markup standard. The Speech Synthesis Markup Language (SSML) specification is largely based on Java Speech Markup Language (JSML), but also incorporates elements and concepts from SABLE, previously published text markup standards, and from Voice eXtensible Markup Language (VoiceXML), which is itself based on JSML and SABLE. SSML also includes new elements designed to optimize the capabilities of contemporary speech synthesis engines in the task of converting text into speech. This paper summarizes the markup element design philosophy and includes descriptions of each of the speech synthesis markup elements.

Read the paper · More papers on PaperTik