ProsoDeep · Deep understanding and modelling of the hierarchical structure of Prosody
Horizon 2020 — Marie Skłodowska-Curie Actions
- Duration
- 2017-09-18 → 2019-01-17
- EU contribution
- €123,384
- Participants
- 1
- Scheme
- MSCA-IF-EF-ST
Lines connect the coordinator with its partners.
Results in brief
Deep understanding and modelling of the hierarchical structure of Prosody
Speech prosody is a multidimensional phenomenon comprising intonation, energy, and rhythm. It is the carrier of both linguistic information, e.g. sentence structure, focus and contrast, lexical stress; as well as paralinguistic information, e.g. gender, age, personality, and emotions. In recent years, prosodic research has considerably enlarged the spectrum of its properties and functions. In contrast, prosodic models able to map signals to functions or vice-versa are rare: comprehensive models of rhythm and intonation have difficulties coping with this expanding dimensionality, and machine learning techniques still have difficulties with offering structuring principles. This issue is of increasing importance to society as speech enabled applications see widespread deployment in our everyday lives. Specifically, we have seen a proliferation of virtual assistants and virtual call operators whose primary mode of communication with the user is speech. Even though these systems handle well the linguistic content of the speech signal, i.e. the spoken words, they struggle understanding information embedded in the prosody, i.e. the meaning behind these words. Thus, a model is needed that will enable these systems to disentangle and decode the various information embedded in speech prosody. The prime objective of the ProsoDeep project was to develop a state-of-the-art Deep Prosody Model that would provide a deeper understanding of the hierarchical encoding of information through the language of prosody. A secondary objective was the recording of a prosodically rich Database that can be used to analyse the interaction of multiple linguistic functions in the way they are communicated through prosody. The ProsoDeep project has achieved both of these objectives.
Data: CORDIS, © European Union
Project objective
The prime objective of the ProsoDeep project is to gain a deeper understanding of the language of prosody through the analysis of all the levels in the production hierarchy of prosody. In particular, it will exploit the benefits of the top-down and bottom-up approaches through their incorporation within a Deep Prosody Model (DPM). The DPM will facilitate the advancement of speech technologies that rely both on the synthesis of prosody, e.g. text-to-speech (TTS) systems, and its analysis, e.g. speech emotion recognition (SER). To reach this objective the project will draw on a variety of scientific fields, including signal processing, physiology of prosody production and biomechanics, linguistics, and machine learning, and will also be augmented with respiratory measurements.
Original text from CORDIS.
Participants
- INSTITUT POLYTECHNIQUE DE GRENOBLE · GRENOBLE CEDEX 1CoordinatorFrance
Links
Data: CORDIS, © European Union
