Speech Interface at Office Workstation

Objective

The overall aims of the SPIN project were to:
-make significant advances in speech input/output algorithms
-study ergonomic aspects in relation to the integration of speech input/output facilities in the office environment
-build a demonstrator illustrating the main results of the project, consisting of a prototype of a workstation with speech facilities.
The aims of the project were to make significant advances in speech input/output algorithms, study ergonomic aspects in relation to the integration of speech input/output facilities and build a demonstrator.

The MPLPC algorithm was the coding method chosen and a simulation of refined versions and a real time breadboard implementation of the algorithm are available.
Work concentrated on algorithm studies. The results were embodied in an experimental demonstrator, MARIPA, which was a low cost recognizer based on a personal computer (PC) board. The demonstration of the first stage of the continuous speech recognizer, producing a lattice of demisyllables for continuous speech recognition, was achieved. Speech input/output assessment methodology was refined and used to improve the quality of the developed algorithm.

Text to speech synthesis systems for French, Italian and Greek were developed, with emphasis on phonetic components, prosody, development of a speech rule compiler and quality evaluations of the French and Italian diphone sets.

Automatic test software for the simulation of speaker verification was produced, running on a large speech database and a final real time automatic speaker verification system was implemented.

A 3 digital signal processing (DSP) based modular hardware for speech processing was designed and developed, and each speech algorithm implemented. An experimental final system, made up a PC and of a speech interface was built to present the results of the project.
Experiments were completed and results and guidelines delivered with regard to:
integration of speech input/output facilities in a multimedia person machine interface;
use of speech in specific office applications;
definition of the best positions for a microphone on a workstation for speech input;
definition of a measuring technique for determining the noise sensitivity of speech recognizers;
behaviour of test subjects when using a multimedia user interface.
The main results of the SPIN project were as follows:
Speech Algorithms
-Coding
Three coding methods (MPLPC, TDHS, and RELP) have been studied in detail and simulated versions of these algorithms produced. Evaluation of their intelligibility showed that the MPLPC algorithm was the best one at the chosen bit rate (9.6 Kbit/s); a simu lation of refined versions and a real-time breadboard implementation of the MPLPC algorithm are currently available.
-Speech recognition
Work was concentrated on algorithm studies. The results were embodied in an experimental demonstrator, MARIPA, which was a low-cost recogniser based on a PC board. The demonstration of the first stage of the continuous speech recogniser, producing a latt ice of demi-syllables for continuous speech recognition, was also achieved. Speech input/output assessment methodology was refined and used to improve the quality of the developed algorithm.
-Text-to-speech synthesis
Text-to-speech synthesis systems for French, Italian and Greek were developed, with emphasis on:
.phonetic components: full diphone dictionaries are available for the three languages dealt with in the project
.prosody: many rules of duration and intonation were defined
.development of a speech rule compiler
.quality evaluations of the French and Italian diphone sets.
-Speaker verification
Automatic test software for the simulation of speaker verification was produced, running on a large speech database. A final real-time automatic speaker verification system was implemented and used to control access to protected areas of the R&D laborato ries.
Hardware Implementation and Integration
A three DSP-based modular hardware for speech processing was designed and developed, and each speech algorithm (coding, speech recognition, speaker verification, text-to-speech synthesis) implemented.
An experimental final system, made up a PC and of a speech interface (built in a VME environment and connected to the PC via a serial line) was built to present the results of the project. The office application chosen was agenda planning.
Ergonomic aspects were also carefully studied. Several experiments were completed and significant results and guidelines delivered with regard to:
-integration of speech input/output facilities in a multimedia person-machine interface
-use of speech in specific office applications
-definition of the best positions for a microphone on a workstation for speech input
-definition of a measuring technique for determining the noise sensitivity of speech recognisers
-behaviour of test subjects when using a multimedia user interface.
Exploitation
The results of this project were used in project 954, IKAROS.

Fields of science (EuroSciVoc)

CORDIS classifies projects with EuroSciVoc, a multilingual taxonomy of fields of science, through a semi-automatic process based on NLP techniques. See: The European Science Vocabulary.

Programme(s)

Multi-annual funding programmes that define the EU’s priorities for research and innovation.

FP1-ESPRIT 1 - European programme (EEC) for research and development in information technologies (ESPRIT), 1984-1988

Topic(s)

Calls for proposals are divided into topics. A topic defines a specific subject or area for which applicants can submit proposals. The description of a topic comprises its specific scope and the expected impact of the funded project.

Data not available

Call for proposal

Procedure for inviting applicants to submit project proposals, with the aim of receiving EU funding.

Data not available

Funding Scheme

Funding scheme (or “Type of Action”) inside a programme with common features. It specifies: the scope of what is funded; the reimbursement rate; specific evaluation criteria to qualify for funding; and the use of simplified forms of costs like lump sums.

Data not available

Coordinator

SOCIETE ETUDES SYSTEMS AUTOMATIONS (SESA)

EU contribution

No data

Address

3 RUE DU CLOS COURTEL
35018 RENNES
France

Total cost

No data

Participants (9)

Alcatel Alsthom Recherche

France

EU contribution

No data

Address

Route de Nozay
91460 Marcoussis

Total cost

No data

CMSU-COMMUNICATION & MANAGEMENT SYSTEMS UNIT.

Greece

EU contribution

No data

Address

ZOGRAPHOU CAMPUS
15773 ATHINAI

Total cost

No data

COMMISSARIAT A L'ENERGIE ATOMIQUE

France

EU contribution

No data

Address

Rue de la Federation 31-33
75015 PARIS

Total cost

No data

Centro Studi e Laboratori Telecomunicazioni SpA

Italy

EU contribution

No data

Daimler-Benz AG

Germany

EU contribution

No data

Address

Wilhelm-Runge-Straße 11
89013 Ulm

Total cost

No data

Oros SA

France

EU contribution

No data

Address

13 chemin des Prés ZIRST
38241 Meylan

Total cost

No data

SIEMENS-NIXDORF INFORMATIONSSYSTEME AG

Germany

EU contribution

No data

Address

BERLINER STRAßE
1000 BERLIN

Total cost

No data

Scuola Normale Superiore di Pisa

Italy

EU contribution

No data

Address

Piazza dei Cavalieri 7
56126 Pisa

Total cost

No data

UNIV VAN AMSTERDAM

Netherlands

EU contribution

No data

Address

ROETERSSTRAAT
1018 AMSTERDAM

Total cost

No data

Objective

Fields of science (EuroSciVoc) CORDIS classifies projects with EuroSciVoc, a multilingual taxonomy of fields of science, through a semi-automatic process based on NLP techniques. See: The European Science Vocabulary.

Programme(s) Multi-annual funding programmes that define the EU’s priorities for research and innovation.

Topic(s) Calls for proposals are divided into topics. A topic defines a specific subject or area for which applicants can submit proposals. The description of a topic comprises its specific scope and the expected impact of the funded project.

Call for proposal Procedure for inviting applicants to submit project proposals, with the aim of receiving EU funding.

Funding Scheme Funding scheme (or “Type of Action”) inside a programme with common features. It specifies: the scope of what is funded; the reimbursement rate; specific evaluation criteria to qualify for funding; and the use of simplified forms of costs like lump sums.

Coordinator

Participants (9)

Download Download the content of the page

Fields of science (EuroSciVoc)

CORDIS classifies projects with EuroSciVoc, a multilingual taxonomy of fields of science, through a semi-automatic process based on NLP techniques. See: The European Science Vocabulary.

Programme(s)

Multi-annual funding programmes that define the EU’s priorities for research and innovation.

Topic(s)

Calls for proposals are divided into topics. A topic defines a specific subject or area for which applicants can submit proposals. The description of a topic comprises its specific scope and the expected impact of the funded project.

Call for proposal

Procedure for inviting applicants to submit project proposals, with the aim of receiving EU funding.

Funding Scheme

Funding scheme (or “Type of Action”) inside a programme with common features. It specifies: the scope of what is funded; the reimbursement rate; specific evaluation criteria to qualify for funding; and the use of simplified forms of costs like lump sums.