Integrating morpho-phonology in speech recognition

Project Information

MorSR

Grant agreement ID: 838058

DOI

10.3030/838058

Project closed

EC signature date 13 February 2019

Start date 1 March 2019

End date 31 August 2020

Funded under

EXCELLENT SCIENCE - European Research Council (ERC)

Total cost

€ 149 919,00

EU contribution

€ 149 919,00

149 919,00

Coordinated by

THE CHANCELLOR, MASTERS AND SCHOLARS OF THE UNIVERSITY OF OXFORD
United Kingdom

Project description

Teaching machines to predict words in running speech

Current automatic speech recognition (ASR) technology hinges on rich acoustic representations of words and extensive training on large corpora of recorded speech to enable recognition of speech sounds in all their variance combined with probabilistic sequencing of whole words in a language model trained on large written text corpora. In contrast, building on the FlexSR recognition of phonological building-blocks, the EU-funded MorSR project will further enable systems to reject improbable words by exploiting linguistic information about word structure, such as the systematic, language-specific processes that alter the manifestations of particular speech sounds at the boundaries between words. This will improve both ASR performance and the adaptability of systems to languages where the availability of training data is reduced.

Objective

Automatic Speech Recognition (ASR) is considered to represent the most natural man-machine interface across the spectrum of technological space. Current commercial ASR systems rely on a ‘rich’ representation of an acoustic signal for words and their variants, resulting in major challenges in the deployment of ASR systems in areas where it could have substantial social impact. Our central goal is to translate research results from the ERC funded project MORPHON into a novel ASR system to remove such barriers. We have previously demonstrated that the use of a universal set of phonological features delivers an isolated word recognition system (FlexSR) with enhanced phoneme recognition accuracy. It is more robust under conditions of non-standard speech, dialect variation and can be easily adapted to new languages. These aspects are problematic for current ASR systems which rely on the probabilistic sequencing of whole words in their language model (LM) based on large written text corpora for training. Obtaining sufficient training data for a new LM is prohibitively expensive. Instead, MorSR will incorporate linguistic information about word-structure to reject improbable words. This reduces the search space and increases the probability of identifying correct words. A major outcome will be an innovative LM based on linguistic principles. Unlike existing approaches, it is based on speech data to capture crucial regularities that are lost in text corpora. Combined with FlexSR's key strengths in identifying subtle phonological contrasts, MorSR will not only enable improved predictions of word sequences in running speech, but also dramatically reduce the requirement for training data when adapting the system to a new language. MorSR's strengths include: (a) prediction of fine-grained possibilities of word sequences based on grammatical principles; (b) requiring considerably less training data; (c) easily adaptable to new languages; and (d) will be fast, secure and accurate.

Fields of science (EuroSciVoc)

CORDIS classifies projects with EuroSciVoc, a multilingual taxonomy of fields of science, through a semi-automatic process based on NLP techniques. See: The European Science Vocabulary.

This project has not yet been classified with EuroSciVoc.
Be the first one to suggest relevant scientific fields and help us improve our classification service

Keywords

Project’s keywords as indicated by the project coordinator. Not to be confused with the EuroSciVoc taxonomy (Fields of science)

Programme(s)

Multi-annual funding programmes that define the EU’s priorities for research and innovation.

H2020-EU.1.1. - EXCELLENT SCIENCE - European Research Council (ERC) MAIN PROGRAMME
See all projects funded under this programme

Topic(s)

Calls for proposals are divided into topics. A topic defines a specific subject or area for which applicants can submit proposals. The description of a topic comprises its specific scope and the expected impact of the funded project.

ERC-2018-PoC - ERC Proof of Concept Grant
See all projects funded under this topic

Funding Scheme

Funding scheme (or “Type of Action”) inside a programme with common features. It specifies: the scope of what is funded; the reimbursement rate; specific evaluation criteria to qualify for funding; and the use of simplified forms of costs like lump sums.

ERC-POC - Proof of Concept Grant

See all projects funded under this funding scheme

Call for proposal

Procedure for inviting applicants to submit project proposals, with the aim of receiving EU funding.

(opens in new window) ERC-2018-PoC

See all projects funded under this call

Host institution

THE CHANCELLOR, MASTERS AND SCHOLARS OF THE UNIVERSITY OF OXFORD

Net EU contribution

€ 149 919,00

Address

WELLINGTON SQUARE UNIVERSITY OFFICES
OX1 2JD Oxford
United Kingdom

Region

South East (England) Berkshire, Buckinghamshire and Oxfordshire Oxfordshire

Activity type

Higher or Secondary Education Establishments

Links

Contact the organisation Website

Participation in EU R&I programmes

HORIZON collaboration network

Total cost

€ 149 919,00

Beneficiaries (1)

THE CHANCELLOR, MASTERS AND SCHOLARS OF THE UNIVERSITY OF OXFORD

United Kingdom

Net EU contribution

€ 149 919,00

Project description

Teaching machines to predict words in running speech

Objective

Fields of science (EuroSciVoc)

CORDIS classifies projects with EuroSciVoc, a multilingual taxonomy of fields of science, through a semi-automatic process based on NLP techniques. See: The European Science Vocabulary.

Keywords

Project’s keywords as indicated by the project coordinator. Not to be confused with the EuroSciVoc taxonomy (Fields of science)

Programme(s)

Multi-annual funding programmes that define the EU’s priorities for research and innovation.

Topic(s)

Calls for proposals are divided into topics. A topic defines a specific subject or area for which applicants can submit proposals. The description of a topic comprises its specific scope and the expected impact of the funded project.

Funding Scheme

Funding scheme (or “Type of Action”) inside a programme with common features. It specifies: the scope of what is funded; the reimbursement rate; specific evaluation criteria to qualify for funding; and the use of simplified forms of costs like lump sums.

Call for proposal

Procedure for inviting applicants to submit project proposals, with the aim of receiving EU funding.

Host institution

Beneficiaries (1)

Share this page Share this page on social networks

Download Download the content of the page

Integrating morpho-phonology in speech recognition

Project description

Teaching machines to predict words in running speech

Objective

Fields of science (EuroSciVoc) CORDIS classifies projects with EuroSciVoc, a multilingual taxonomy of fields of science, through a semi-automatic process based on NLP techniques. See: The European Science Vocabulary.

Keywords Project’s keywords as indicated by the project coordinator. Not to be confused with the EuroSciVoc taxonomy (Fields of science)

Programme(s) Multi-annual funding programmes that define the EU’s priorities for research and innovation.

Topic(s) Calls for proposals are divided into topics. A topic defines a specific subject or area for which applicants can submit proposals. The description of a topic comprises its specific scope and the expected impact of the funded project.

Funding Scheme Funding scheme (or “Type of Action”) inside a programme with common features. It specifies: the scope of what is funded; the reimbursement rate; specific evaluation criteria to qualify for funding; and the use of simplified forms of costs like lump sums.

Call for proposal Procedure for inviting applicants to submit project proposals, with the aim of receiving EU funding.

Host institution

Beneficiaries (1)

Share this page Share this page on social networks

Download Download the content of the page

Fields of science (EuroSciVoc)

CORDIS classifies projects with EuroSciVoc, a multilingual taxonomy of fields of science, through a semi-automatic process based on NLP techniques. See: The European Science Vocabulary.

Keywords

Project’s keywords as indicated by the project coordinator. Not to be confused with the EuroSciVoc taxonomy (Fields of science)

Programme(s)

Multi-annual funding programmes that define the EU’s priorities for research and innovation.

Topic(s)

Calls for proposals are divided into topics. A topic defines a specific subject or area for which applicants can submit proposals. The description of a topic comprises its specific scope and the expected impact of the funded project.

Funding Scheme

Funding scheme (or “Type of Action”) inside a programme with common features. It specifies: the scope of what is funded; the reimbursement rate; specific evaluation criteria to qualify for funding; and the use of simplified forms of costs like lump sums.

Call for proposal

Procedure for inviting applicants to submit project proposals, with the aim of receiving EU funding.