Project description
Open linguistic infrastructure
Human language technologies crucially depend on language resources and tools that are usable, useful and available. However, even where language resources and respective tools are available they have been developed mostly in a sporadic manner, in response to specific project needs, with relatively little regard to their long-term sustainability, IPR status, interoperability, reusability in different contexts as well as to their potential deployment in multilingual applications.The CESAR project, in close harmony with META-NET intends to address this issue by enhancing, upgrading, standardising and cross-linking a wide variety of language resources and tools and making them available, thus contributing to an open linguistic infrastructure.Partners in the CESAR consortium are key players in their respective language community with a proven track record in European language technology projects including infrastructure initiatives, such as TELRI in the past and notably CLARIN at present.The project will make available a comprehensive set of language resources and tools covering the Hungarian, Polish, Croatian, Serbian, Bulgarian and Slovak languages. Resources will include interoperable mono and multilingual speech databases, corpora, dictionaries and wordnets and relevant language technology processing tools such as tokenisers, lemmatisers, taggers and parsers. A comprehensive initial list of tools and resources identified as potential targets of project activity can be found in Section 3.2. Depending on local capacities and efficiency considerations, the resources may be either made available as web services at partners' site or contributed to the META_SHARE digital exchange facility.In addition to the technical work required, great effort will be made to ensure sustainability through mobilising the LT community, raising awareness of the fundamental role of language resources among the R&D policy makers, the media and the general public.
Human language technologies crucially depend on language resources and tools that are usable, useful and available. However, even where language resources and respective tools are available they have been developed mostly in a sporadic manner, in response to specific project needs, with relatively little regard to their long-term sustainability, IPR status, interoperability, reusability in different contexts as well as to their potential deployment in multilingual applications.The CESAR project, in close harmony with META-NET intends to address this issue by enhancing, upgrading, standardising and cross-linking a wide variety of language resources and tools and making them available, thus contributing to an open linguistic infrastructure.Partners in the CESAR consortium are key players in their respective language community with a proven track record in European language technology projects including infrastructure initiatives, such as TELRI in the past and notably CLARIN at present.The project will make available a comprehensive set of language resources and tools covering the Hungarian, Polish, Croatian, Serbian, Bulgarian and Slovak languages. Resources will include interoperable mono and multilingual speech databases, corpora, dictionaries and wordnets and relevant language technology processing tools such as tokenisers, lemmatisers, taggers and parsers. A comprehensive initial list of tools and resources identified as potential targets of project activity can be found in Section 3.2. Depending on local capacities and efficiency considerations, the resources may be either made available as web services at partners' site or contributed to the META_SHARE digital exchange facility.In addition to the technical work required, great effort will be made to ensure sustainability through mobilising the LT community, raising awareness of the fundamental role of language resources among the R&D policy makers, the media and the general public.
Programme(s)
Multi-annual funding programmes that define the EU’s priorities for research and innovation.
Multi-annual funding programmes that define the EU’s priorities for research and innovation.
Topic(s)
Calls for proposals are divided into topics. A topic defines a specific subject or area for which applicants can submit proposals. The description of a topic comprises its specific scope and the expected impact of the funded project.
Calls for proposals are divided into topics. A topic defines a specific subject or area for which applicants can submit proposals. The description of a topic comprises its specific scope and the expected impact of the funded project.
Call for proposal
Procedure for inviting applicants to submit project proposals, with the aim of receiving EU funding.
Procedure for inviting applicants to submit project proposals, with the aim of receiving EU funding.
CIP-ICT-PSP-2010-4
See other projects for this call
Funding Scheme
Funding scheme (or “Type of Action”) inside a programme with common features. It specifies: the scope of what is funded; the reimbursement rate; specific evaluation criteria to qualify for funding; and the use of simplified forms of costs like lump sums.
Funding scheme (or “Type of Action”) inside a programme with common features. It specifies: the scope of what is funded; the reimbursement rate; specific evaluation criteria to qualify for funding; and the use of simplified forms of costs like lump sums.
Coordinator
1068 Budapest
Hungary
The total costs incurred by this organisation to participate in the project, including direct and indirect costs. This amount is a subset of the overall project budget.