Edinburgh Research Archive >
Centre for Speech Technology Research >
CSTR publications >
Please use this identifier to cite or link to this item:
|Title: ||Objective Distance Measures for Spectral Discontinuities in Concatenative Speech Synthesis|
|Authors: ||Vepa, Jithendra|
|Issue Date: ||Sep-2002|
|Citation: ||Proc. ICSLP, Denver, USA|
|Abstract: ||In unit selection based concatenative speech systems, join cost, which measures how well two units can be joined together, is one of the main criteria for selecting appropriate units from the inventory. The ideal join cost will measure
perceived discontinuity, based on easily measurable spectral properties of the units being joined, in order to ensure smooth and natural-sounding synthetic speech. In this paper we report a perceptual experiment conducted to measure the correlation between subjective human perception and various objective spectrally-based measures proposed in the literature.
Our experiments used a state-of-the art unit-selection text-to-speech system: rVoice from Rhetorical Systems Ltd.|
|Appears in Collections:||CSTR publications|
Items in ERA are protected by copyright, with all rights reserved, unless otherwise indicated.