H2020Staff exchange2015–2019

LISTEN · Hands-free Voice-enabled Interface to Web Applications for Smart Home Environments

Horizon 2020 — Marie Skłodowska-Curie Actions

Duration
2015-06-01 → 2019-05-31
EU contribution
€414,000
Participants
4
Scheme
MSCA-RISE

Lines connect the coordinator with its partners.

Results in brief

Hands-free Voice-enabled Interface to Web Applications for Smart Home Environments

In this project, our goal has been to design a complete (software and hardware solution) voice-enabled interface specifically designed for Internet applications for the smart home environment, for controlling specific automations of the smart home, but also for providing access to the web for specific tasks, such as access to emails, social networking platforms, calendar, etc. A clear distinction compared to currently popular systems for voice-based Internet access is the Wireless Acoustic Sensor Network (WASN) approach, that supports a truly seamless operation of the voice interface. The proposed design was based on a distributed operation of several acoustic sensors, able to localise the speakers and enhance the speech capture for their particular locations, and enhance and transmit the speech signal to the speech transcription engine. LISTEN’s sensor network features low-cost hardware and stand-alone real-time operation due to a custom MEMS (micro-electro-mechanical systems) microphones design. The acoustic front-end was jointly developed and optimised with the automatic speech recognition (ASR) system. A central motivation behind forming the LISTEN consortium and objectives was to form a much-needed interdisciplinary team and research plan, so as to achieve substantial progress beyond state-of-the-art in hands-free ASR systems. The objectives of the LISTEN project have been: • Objective 1: To develop a large-vocabulary speech recognition system for the smart home, for multiple domains (e.g., command/control of the smart home functionalities, web access, web search, email / message dictation, calendar and to-do lists, social networks, etc.), and for multiple languages. Focus has been on English, Greek, Italian, and German to demonstrate the system’s easy adaptation to further languages. • Objective 2: To develop a front-end for robust speech capture in a distributed fashion, performing in real-time localisation of multiple speakers and accordingly enhancing the speech capture for these specific locations. The front-end design includes the hardware design of a low-cost distributed acoustic sensor network, where each sensor is equipped with an on-board processor for the signal localisation and enhancement to be performed in a collaborative manner. • Objective 3: To create a prototype and evaluate it in practical situations, including in the Ambient Intelligence Facility of the coordinator (FORTH). To widely disseminate the project results and maximise the potential societal and commercial impact of our activities and the developed platform.

Data: CORDIS, © European Union

Project objective

Nowadays, it is becoming increasingly affordable to enhance the home environment with several automation schemes, allowing remote control of e.g., heating/cooling, communication, lighting, media, etc. Such smart home functionalities are essential for people with disabilities and the elderly, as they not only provide assistive control of important everyday functionalities, but may prove to be life-saving in case of emergency. However, smart home functionalities may become useless for people who need those most, if they cannot be accessed via a natural, easy to use, interface.The central objective of LISTEN is to design and implement a complete system, including both the software and hardware components, enabling robust hands-free large-vocabulary voice-based access to Internet applications in smart homes. This would allow the users to have natural control (i.e., using their voice) of the smart-home web-enabled functionalities (e.g., turning on/off web-enabled “smart” appliances), but also to access specific Internet applications (e.g., web search, email dictation, access to social networks). A truly hands-free system operation of the voice interface is equally important: users will not have to turn towards a microphone or other device, or wear a headset. Therefore, LISTEN will develop (a) a robust hands-free speech capture system operating as a wireless acoustic sensor network (WASN), specifically designed for the smart home, and (b) a large-vocabulary automatic speech recognition system optimised for accessing web applications and controlling web-enabled smart home automation functionalities. LISTEN pushes the boundaries of current state-of-the-art by bridging the gap between the acoustic front-end and automatic speech recognition research communities, with the common goal of developing a smart-home-specific natural voice interface to web services.

Original text from CORDIS.

Participants

  • IDRYMA TECHNOLOGIAS KAI EREVNAS · IRAKLEIOCoordinatorGreece
  • CEDAT 85 SRL · San Vito Dei NormanniItaly
  • EML SPEECH TECHNOLOGY GMBH · KarlsruheGermany
  • RHEINISCH-WESTFAELISCHE TECHNISCHE HOCHSCHULE AACHEN · AachenGermany

Links

Data: CORDIS, © European Union