NWU Institutional Repository

Part-of-speech effects on text-to-speech synthesis

dc.contributor.authorSchlunz, Georg I.
dc.contributor.authorBarnard, Etienne
dc.contributor.authorvan Huyssteen, Gerhard B.
dc.date.accessioned2018-03-07T10:27:13Z
dc.date.available2018-03-07T10:27:13Z
dc.date.issued2010
dc.description.abstractOne of the goals of text-to-speech (TTS) systems is to produce natural-sounding synthesized speech. Towards this end various natural language processing (NLP) tasks are performed to model the prosodic aspects of the TTS voice. One of the fundamental NLP tasks being used is the part-of-speech (POS) tagging of the words in the text. This paper investigates the effects of POS information on the naturalness of a hidden Markov model (HMM) based TTS voice when additional resources are not available to aid in the modeling of prosody. It is found that, when a minimal feature set is used for the HMM context labels, the addition of POS tags does improve the naturalness of the voice. However, the same effect can be accomplished by including segmental counting and positional information instead of the POS tags.en_US
dc.description.sponsorshipHuman Language Technology Competency Area, CSIR, Meraka Institute, Pretoria, South Africa Multilingual Speech Technologies, North-West University, Vanderbijlpark, South Africa Centre for Text Technology, North-West University, Potchefstroom, South Africaen_US
dc.identifier.citationGeorg Schlünz, Etienne Barnard and Gerhard van Huyssteen, “Part-of-speech effects on text-to-speech synthesis”, in Proc. Annual Symp. Pattern Recognition Association of South Africa (PRASA), pp 257-262, Stellenbosch, South Africa, 2010. [http://engineering.nwu.ac.za/multilingual-speech-technologies-must/publications]en_US
dc.identifier.urihttps://researchspace.csir.co.za/dspace/bitstream/handle/10204/4674/Schlunz_2010.pdf?sequence=1&isAllowed=y
dc.identifier.urihttp://hdl.handle.net/10394/26555
dc.language.isoenen_US
dc.publisherPattern Recognition Association of South Africa and Mechatronics International Conferenceen_US
dc.subjectSpeech effectsen_US
dc.subjectText-to-speechen_US
dc.subjectNatural Language Processing— Speech recognition and synthesisen_US
dc.titlePart-of-speech effects on text-to-speech synthesisen_US
dc.typePresentationen_US

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
Part-of-speech effects on text-to-speech synthesis.pdf
Size:
3.69 MB
Format:
Adobe Portable Document Format
Description:
Part-of-speech effects on text-to-speech synthesis

License bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
license.txt
Size:
1.61 KB
Format:
Item-specific license agreed upon to submission
Description: