ProsoDeep · Deep understanding and modelling of the hierarchical structure of Prosody
„Хоризонт 2020“ — Действия „Мария Склодовска-Кюри“
- Период
- 2017-09-18 → 2019-01-17
- Финансиране от ЕС
- 123 384 €
- Участници
- 1
- Схема
- MSCA-IF-EF-ST
Линиите свързват координатора с партньорите.
Накратко на български
Просодията на речта изследва интонацията, ритъма и енергията, които предават емоции или разликата между въпрос и твърдение. Това помага на виртуалните асистенти да разпознават смисъла зад думите, а не само самия текст.
Кратко обяснение, генерирано от езиков модел по текста на CORDIS. Оригиналът е по-долу.
Резултати накратко
Deep understanding and modelling of the hierarchical structure of Prosody
Speech prosody is a multidimensional phenomenon comprising intonation, energy, and rhythm. It is the carrier of both linguistic information, e.g. sentence structure, focus and contrast, lexical stress; as well as paralinguistic information, e.g. gender, age, personality, and emotions. In recent years, prosodic research has considerably enlarged the spectrum of its properties and functions. In contrast, prosodic models able to map signals to functions or vice-versa are rare: comprehensive models of rhythm and intonation have difficulties coping with this expanding dimensionality, and machine learning techniques still have difficulties with offering structuring principles. This issue is of increasing importance to society as speech enabled applications see widespread deployment in our everyday lives. Specifically, we have seen a proliferation of virtual assistants and virtual call operators whose primary mode of communication with the user is speech. Even though these systems handle well the linguistic content of the speech signal, i.e. the spoken words, they struggle understanding information embedded in the prosody, i.e. the meaning behind these words. Thus, a model is needed that will enable these systems to disentangle and decode the various information embedded in speech prosody. The prime objective of the ProsoDeep project was to develop a state-of-the-art Deep Prosody Model that would provide a deeper understanding of the hierarchical encoding of information through the language of prosody. A secondary objective was the recording of a prosodically rich Database that can be used to analyse the interaction of multiple linguistic functions in the way they are communicated through prosody. The ProsoDeep project has achieved both of these objectives.
Текст от CORDIS, на английски · Данни: CORDIS, © Европейски съюз
Цел на проекта
The prime objective of the ProsoDeep project is to gain a deeper understanding of the language of prosody through the analysis of all the levels in the production hierarchy of prosody. In particular, it will exploit the benefits of the top-down and bottom-up approaches through their incorporation within a Deep Prosody Model (DPM). The DPM will facilitate the advancement of speech technologies that rely both on the synthesis of prosody, e.g. text-to-speech (TTS) systems, and its analysis, e.g. speech emotion recognition (SER). To reach this objective the project will draw on a variety of scientific fields, including signal processing, physiology of prosody production and biomechanics, linguistics, and machine learning, and will also be augmented with respiratory measurements.
Оригинален текст от CORDIS (на английски).
Участници
- INSTITUT POLYTECHNIQUE DE GRENOBLE · GRENOBLE CEDEX 1КоординаторФранция
Връзки
Данни: CORDIS, © Европейски съюз
