Objectif Science is increasingly data-driven. Scientific research funders now routinely mandate open publication of publicly-funded research data. Safely reusing such data currently requires labour-intensive curation. Provenance recording the history and derivation of the data is critical to reaping the benefits and avoiding the pitfalls of data sharing. There are hundreds of curated scientific databases in biomedicine that need fine-grained provenance; one important example is GtoPdb, a pharmacological database developed by colleagues in Edinburgh. Currently there are no reusable methodologies or practical tools that support provenance for curated databases, forcing each project to start from scratch. Research on provenance for scientific databases is still at an early stage, and prototypes have so far proven challenging to deploy or evaluate in the field. Also, most techniques to date focus on provenance within a single database, but this is only part of the problem: real solutions will have to integrate database provenance with the multiple tiers of web applications, and no-one has begun to address this challenge.I propose research on how to build support for curation into the programming language itself, building on my recent research on the Links Web programming language and on data curation. Links is a strongly-typed language that provides state-of-the-art support for language-integrated query and Web programming. I propose to build on Links and other recent language designs for heterogeneous meta-programming to develop a new language, called Skye, that can express modular, reusable curation and provenance techniques. To keep focus on the real needs of scientific databases, Skye will be evaluated in the context of GtoPdb and other scientific database projects. Bridging the gap between curation research and the practices of scientific database curators will catalyse a virtuous cycle that will increase the pace of breakthrough results from data-driven science. Champ scientifique natural sciencescomputer and information sciencesdata sciencenatural sciencescomputer and information sciencessoftwarenatural sciencescomputer and information sciencesdatabaseshumanitieshistory and archaeologyhistorymedical and health sciencesbasic medicinepharmacology and pharmacy Mots‑clés Curated databases provenance language-integrated query heterogeneous meta-programming Programme(s) H2020-EU.1.1. - EXCELLENT SCIENCE - European Research Council (ERC) Main Programme Thème(s) ERC-CoG-2015 - ERC Consolidator Grant Appel à propositions ERC-2015-CoG Voir d’autres projets de cet appel Régime de financement ERC-COG - Consolidator Grant Institution d’accueil THE UNIVERSITY OF EDINBURGH Contribution nette de l'UE € 1 995 181,00 Adresse OLD COLLEGE, SOUTH BRIDGE EH8 9YL Edinburgh Royaume-Uni Voir sur la carte Région Scotland Eastern Scotland Edinburgh Type d’activité Higher or Secondary Education Establishments Liens Contacter l’organisation Opens in new window Site web Opens in new window Participation aux programmes de R&I de l'UE Opens in new window Réseau de collaboration HORIZON Opens in new window Coût total € 1 995 181,00 Bénéficiaires (1) Trier par ordre alphabétique Trier par contribution nette de l'UE Tout développer Tout réduire THE UNIVERSITY OF EDINBURGH Royaume-Uni Contribution nette de l'UE € 1 995 181,00 Adresse OLD COLLEGE, SOUTH BRIDGE EH8 9YL Edinburgh Voir sur la carte Région Scotland Eastern Scotland Edinburgh Type d’activité Higher or Secondary Education Establishments Liens Contacter l’organisation Opens in new window Site web Opens in new window Participation aux programmes de R&I de l'UE Opens in new window Réseau de collaboration HORIZON Opens in new window Coût total € 1 995 181,00