Ir directamente a la navegación principal Ir directamente a la búsqueda Ir directamente al contenido principal

PROSODIC VS. SEGMENTAL CONTRIBUTIONS TO NATURALNESS IN A DIPHONE SYNTHESIZER

  • University of Delaware

Producción científicarevisión exhaustiva

5 Citas (Scopus)

Resumen

The relative contributions of segmental versus prosodic factors to the perceived naturalness of synthetic speech was measured by transplanting prosody between natural speech and the output of a diphone synthesizer. A small corpus was created containing matched sentence pairs wherein one member of the pair was a natural utterance and the other was a synthetic utterance generated with diphone data from the same talker. Two additional sentences were formed from each sentence pair by transplanting the prosodic structure between the natural and synthetic members of each pair. In two listening experiments subjects were asked to (a) classify each sentence as “natural” or “synthetic, or (b) rate the naturalness of each sentence. Results showed that the prosodic information was more important than segmental information in both classification and ratings of naturalness.

Idioma originalEnglish
EstadoPublished - 1998
Evento5th International Conference on Spoken Language Processing, ICSLP 1998 - Sydney
Duración: 30 nov 19984 dic 1998

Conference

Conference5th International Conference on Spoken Language Processing, ICSLP 1998
País/TerritorioAustralia
CiudadSydney
Período30/11/984/12/98

Huella

Profundice en los temas de investigación de 'PROSODIC VS. SEGMENTAL CONTRIBUTIONS TO NATURALNESS IN A DIPHONE SYNTHESIZER'. En conjunto forman una huella única.

Citar esto