Integrating morpho-phonology in speech recognition
Sector: Commercial • Location: United Kingdom
Source: EU Funding & Tenders Portal
Automatic Speech Recognition (ASR) is considered to represent the most natural man-machine interface across the spectrum of technological space. Current commercial ASR systems rely on a ‘rich’ representation of an acoustic signal for words and their variants, resulting in major challenges in the deployment of ASR systems in areas where it could have substantial social impact. Our central goal is t
Project Information FAQ
Project Information
Want to explore the full details? View the full report
Participants
Sponsoring Agency | Obfuscated Data |
Company | Obfuscated Data |
Status
Original status | ended |
Taiyo status | Obfuscated Data |
Taiyo last update | 00-00-0000 |
Available timestamps | 00-00-0000 |
Available timestamp type | Obfuscated Data |
Contact
Contact name | Obfuscated Data |
Phone | 0000000000 |
ObfuscatedData@email.com | |
Address | Obfuscated Data, Obfuscated data, obfuscated data, Obfuscated data |
Description
Description | Automatic Speech Recognition (ASR) is considered to represent the most natural man-machine interface across the spectrum of technological space. Current commercial ASR systems rely on a ‘rich’ representation of an acoustic signal for words and their variants, resulting in major challenges in the deployment of ASR systems in areas where it could have substantial social impact. Our central goal is to translate research results from the ERC funded project MORPHON into a novel ASR system to remove such barriers. We have previously demonstrated that the use of a universal set of phonological features delivers an isolated word recognition system (FlexSR) with enhanced phoneme recognition accuracy. It is more robust under conditions of non-standard speech, dialect variation and can be easily adapted to new languages. These aspects are problematic for current ASR systems which rely on the probabilistic sequencing of whole words in their language model (LM) based on large written text corpora for training. Obtaining sufficient training data for a new LM is prohibitively expensive. Instead, MorSR will incorporate linguistic information about word-structure to reject improbable words. This reduces the search space and increases the probability of identifying correct words. A major outcome will be an innovative LM based on linguistic principles. Unlike existing approaches, it is based on speech data to capture crucial regularities that are lost in text corpora. Combined with FlexSR's key strengths in identifying subtle phonological contrasts, MorSR will not only enable improved predictions of word sequences in running speech, but also dramatically reduce the requirement for training data when adapting the system to a new language. MorSR's strengths include: (a) prediction of fine-grained possibilities of word sequences based on grammatical principles; (b) requiring considerably less training data; (c) easily adaptable to new languages; and (d) will be fast, secure and accurate. |
Original sub-sector | Obfuscated |
Original Currency | USD |
Original budget | 000000000000000 |
Procurement method | Obfuscated Data |
Budget | 000000000000000 |
Location
Region | Obfuscated |
Country | Obfuscated |
State | Obfuscated Data |
County | Obfuscated |
Location | Obfuscated Data, Obfuscated data, obfuscated data, Obfuscated data |
Source
Source reliability | High |
Data quality score | 100% |
Source | Obfuscated Data |
URL | obfuscated_data,obfuscateddata.com |
More Details
Project Type | Obfuscated Data |
Article Published Date | Obfuscated Data |
