The diverse range of organizations contributing to the global research ecosystem is believed to enhance the overall quality and resilience of its output. Mid-sized autonomous research institutes, distinct from universities, play a crucial role in this landscape. They often lead the way in new research fields and experimental methods, including those in social and organizational domains, which are vital for driving innovation. The EU-LIFE alliance was established with the goal of fostering excellence by developing and disseminating best practices among European biomedical research institutes. As directors of the 15 EU-LIFE institutes, we have spent a decade comparing and refining our processes. Now, we are eager to share the insights we've gained. To this end, we have crafted this Charter, outlining 10 principles we deem essential for research institutes to flourish and achieve ground-breaking discoveries. These principles, detailed in the Charter, encompass excellence, independence, training, internationality and inclusivity, mission focus, technological advancement, administrative innovation, cooperation, societal impact, and public engagement. Our aim is to inspire the establishment of new institutes that adhere to these principles and to raise awareness about their significance. We are convinced that they should be viewed a crucial component of any national and international innovation strategies.
Data resources are essential for the long-term preservation of scientific data and the reproducibility of science. The SIB Swiss Institute of Bioinformatics provides the life science community with a portfolio of openly accessible, high-quality databases and software platforms, which vary from expert-curated knowledgebases, such as UniProtKB/Swiss-Prot (part of the UniProt consortium) and STRING, to online platforms such as SWISS-MODEL and SwissDrugDesign. SIB's mission is to ensure that these resources are available in the long term, as long as their return on investment and their scientific impact are high. To this end, SIB provides its resources, in addition to stable financial support, with a range of high-quality, innovative services that are, to our knowledge, unique in the field. Through this first-class management framework with central services, such as user-centric consulting activities, legal support, open-science guidance, knowledge sharing and training efforts, SIB supports the promotion of excellence in resource development and operation. This review presents the ecosystem of data resources at SIB; the process used for the identification, evaluation and development of resources; and the support activities that SIB provides. A set of indicators has been put in place to select the resources and establish quality standards, reflecting their multifaceted nature and complexity. Through this paper, the reader will discover how SIB's leading tools and databases are fostered by the institute, leading them to be best-in-class resources able to tackle the burning matters that society faces from disease outbreaks and cancer to biodiversity and open science.
The SIB Swiss Institute of Bioinformatics (https://www.sib.swiss) creates, maintains and disseminates a portfolio of reliable and state-of-the-art bioinformatics services and resources for the storage, analysis and interpretation of biological data. Through Expasy (https://www.expasy.org), the Swiss Bioinformatics Resource Portal, the scientific community worldwide, freely accesses more than 160 SIB resources supporting a wide range of life science and biomedical research areas. In 2020, Expasy was redesigned through a user-centric approach, known as User-Centred Design (UCD), whose aim is to create user interfaces that are easy-to-use, efficient and targeting the intended community. This approach, widely used in other fields such as marketing, e-commerce, and design of mobile applications, is still scarcely explored in bioinformatics. In total, around 50 people were actively involved, including internal stakeholders and end-users. In addition to an optimised interface that meets users' needs and expectations, the new version of Expasy provides an up-to-date and accurate description of high-quality resources based on a standardised ontology, allowing to connect functionally-related resources.
In establishing the list of ELIXIR Core Data Resources ( https://elixir-europe.org/platforms/data/core-data-resources ) and the list of ELIXIR Deposition Databases ( https://elixir-europe.org/platforms/data/elixir-deposition-databases ), it was recognised that periodic review of the lists will be necessary, to ensure that the standards applied in making these selections are maintained, as is the relevance of the scope of the chosen resources to life science bioinformatics, as new technologies are developed and new fields of research emerge, in some cases superseding older technologies, and refocusing research priorities. This document describes the framework for Annual Indicator Monitoring and Formal Periodic Review of the ELIXIR Core Data Resources and Deposition Databases.
Supplementary information: Supplementary data are available at Bioinformatics online.
Presentation of the ELIXIR Data Platform at the ELIXIR All Hands meeting 2019 in Lisbon, Portugal.
ELIXIR ( https://www.elixir-europe.org/ ) unites Europe’s leading life science organisations in managing and safeguarding the increasing volume of data being generated by publicly funded research. It coordinates, integrates and sustains bioinformatics resources across its member states and enables users in academia and industry to access services that are vital for their research. There are currently 23 Nodes in ELIXIR, and we work together using a ‘Hub and Nodes’ model. ELIXIR's activities are coordinated across five 'Platforms': Data, Tools, Interoperability, Compute and Training. The goal of the ELIXIR Data Platform ( https://www.elixir-europe.org/platforms/data ) is to drive the use, re-use and value of life science data. It aims to do this by providing users with robust, long-term sustainable data resources within a coordinated, scalable and connected data ecosystem. This presentation will outline the initiatives currently underway in the ELIXIR Data Platform. Topics will include the ELIXIR Core Data Resources, selected on the basis of a set of Indicators that demonstrate their fundamental importance to the wider life-science community, and the related set of Deposition Databases for the long-term preservation of biological data. Our work on Literature-Data Integration and Scalable Curation for biocurators, which builds on our text mining work with EuropePMC, will be summarised. Our commitment to Long Term Sustainability of life science data resources, including our contribution to the Global Biodata Coalition, will also be covered. Work on all these topics will continue through the new ELIXIR Scientific Programme set for 2019-2023. Lastly, the Data Platform is currently engaged in seven Implementation Studies, involving fifteen ELIXIR Nodes working with 40 Data Resources across Europe. These studies are due to be completed mid-2019, and the tasks they are engaged in will be introduced.
The process for selecting ELIXIR Core Data Resources has been developed by the Data Platform as part of EXCELERATE deliverable 3.1 and was recently published on ELIXIR’s F1000Research channel (Link to article: Identifying ELIXIR Core Data Resources ). At the Heads of Nodes retreat, 14-15 June, 2016, it was agreed that the process needs to be tested before full adoption. This document describes the process to be tested, culminating in a review of the process by the ELIXIR SAB in March 2017. In this initial test process and evaluation we expect to select up to 20 Core Data Resources. Based on the experience from this first selection round the HoN committee could recommend a refinement of selection criteria & objectives for future. Should the process prove to be sufficiently robust, the candidate resources identified will become the first set of Core Data Resources, and the process will be repeated in October 2017 for new candidate Core Data Resources.
The process for selecting ELIXIR Core Data Resources has been developed by the Data Platform and was published on ELIXIR’s F1000Research channel (Link to article: Identifying ELIXIR Core Data Resources ). On the basis of that work, the first round of selection of Core Data Resources ran and the initial list was published on 25th July 2017. This document has been developed to support the second round of Core Data Resource selection - the results to be published during June of 2018. A further round is planned for 2019.
The core mission of ELIXIR is to build a stable and sustainable infrastructure for biological information across Europe. At the heart of this are the data resources, tools and services that ELIXIR offers to the life-sciences community, providing stable and sustainable access to biological data. ELIXIR aims to ensure that these resources are available long-term and that the life-cycles of these resources are managed such that they support the scientific needs of the life-sciences, including biological research. ELIXIR Core Data Resources are defined as a set of European data resources that are of fundamental importance to the wider life-science community and the long-term preservation of biological data. They are complete collections of generic value to life-science, are considered an authority in their field with respect to one or more characteristics, and show high levels of scientific quality and service. Thus, ELIXIR Core Data Resources are of wide applicability and usage. This paper describes the structures, governance and processes that support the identification and evaluation of ELIXIR Core Data Resources. It identifies key indicators which reflect the essence of the definition of an ELIXIR Core Data Resource and support the promotion of excellence in resource development and operation. It describes the specific indicators in more detail and explains their application within ELIXIR’s sustainability strategy and science policy actions, and in capacity building, life-cycle management and technical actions. The identification process is currently being implemented and tested for the first time. The findings and outcome will be evaluated by the ELIXIR Scientific Advisory Board in March 2017. Establishing the portfolio of ELIXIR Core Data Resources and ELIXIR Services is a key priority for ELIXIR and publicly marks the transition towards a cohesive infrastructure.
On November 18-19, 2016, the Human Frontier Science Program Organization (HFSPO) hosted a meeting of senior managers of key data resources and leaders of several major funding organizations to discuss the challenges associated with sustaining biological and biomedical (i.e., life sciences) data resources and associated infrastructure. A strong consensus emerged from the group that core data resources for the life sciences should be supported through a coordinated international effort(s) that better ensure long-term sustainability and that appropriately align funding with scientific impact. Ideally, funding for such data resources should allow for access at no charge, as is presently the usual (and preferred) mechanism. Below, the rationale for this vision is described, and some important considerations for developing a new international funding model to support core data resources for the life sciences are presented.
Millions of life scientists across the world rely on bioinformatics data resources for their research projects. Data resources can be very expensive, especially those with a high added value as the expert-curated knowledgebases. Despite the increasing need for such highly accurate and reliable sources of scientific information, most of them do not have secured funding over the near future and often depend on short-term grants that are much shorter than their planning horizon. Additionally, they are often evaluated as research projects rather than as research infrastructure components. In this work, twelve funding models for data resources are described and applied on the case study of the Universal Protein Resource (UniProt), a key resource for protein sequences and functional information knowledge. We show that most of the models present inconsistencies with open access or equity policies, and that while some models do not allow to cover the total costs, they could potentially be used as a complementary income source. We propose the Infrastructure Model as a sustainable and equitable model for all core data resources in the life sciences. With this model, funding agencies would set aside a fixed percentage of their research grant volumes, which would subsequently be redistributed to core data resources according to well-defined selection criteria. This model, compatible with the principles of open science, is in agreement with several international initiatives such as the Human Frontiers Science Program Organisation (HFSPO) and the OECD Global Science Forum (GSF) project. Here, we have estimated that less than 1% of the total amount dedicated to research grants in the life sciences would be sufficient to cover the costs of the core data resources worldwide, including both knowledgebases and deposition databases.
Switzerland has been a pioneer in the field of bioinformatics since the early 1980s. As time passed, the need for one entity to gather and represent bioinformatics on a national scale was felt and, in 1998, the SIB Swiss Institute of Bioinformatics was created. Hence, 2018 marks the Institute's 20th anniversary. Today, the Institute federates 65 research and service groups across the country-whose activity domains range from genomics, proteomics, medicine and health to structural biology, systems biology, phylogeny and evolution-and a group whose sole task is dedicated to training. The Institute hosts 12 competence centres that provide bioinformatics and biocuration expertise to life scientists across the country. SIB sensed early on that the wealth of data produced by modern technologies in medicine and the growing self-awareness of patients was about to revolutionize the way medical data are considered. In 2012, it created a Clinical Bioinformatics group to address the issue of personalized health, thus working towards a more global approach to patient management, and more targeted and effective therapies. In this respect, SIB has a major role in the Swiss Personalized Health Network to make patient-related data available to research throughout the country. The uniqueness of the Institute's governance structure has also inspired the structure of other European life science organizations, notably ELIXIR.
The SIB Swiss Institute of Bioinformatics (www.isb-sib.ch) provides world-class bioinformatics databases, software tools, services and training to the international life science community in academia and industry. These solutions allow life scientists to turn the exponentially growing amount of data into knowledge. Here, we provide an overview of SIB's resources and competence areas, with a strong focus on curated databases and SIB's most popular and widely used resources. In particular, SIB's Bioinformatics resource portal ExPASy features over 150 resources, including UniProtKB/Swiss-Prot, ENZYME, PROSITE, neXtProt, STRING, UniCarbKB, SugarBindDB, SwissRegulon, EPD, arrayMap, Bgee, SWISS-MODEL Repository, OMA, OrthoDB and other databases, which are briefly described in this article.
The core mission of ELIXIR is to build a stable and sustainable infrastructure for biological information across Europe. At the heart of this are the data resources, tools and services that ELIXIR offers to the life-sciences community, providing stable and sustainable access to biological data. ELIXIR aims to ensure that these resources are available long-term and that the life-cycles of these resources are managed such that they support the scientific needs of the life-sciences, including biological research. ELIXIR Core Data Resources are defined as a set of European data resources that are of fundamental importance to the wider life-science community and the long-term preservation of biological data. They are complete collections of generic value to life-science, are considered an authority in their field with respect to one or more characteristics, and show high levels of scientific quality and service. Thus, ELIXIR Core Data Resources are of wide applicability and usage. This paper describes the structures, governance and processes that support the identification and evaluation of ELIXIR Core Data Resources. It identifies key indicators which reflect the essence of the definition of an ELIXIR Core Data Resource and support the promotion of excellence in resource development and operation. It describes the specific indicators in more detail and explains their application within ELIXIR’s sustainability strategy and science policy actions, and in capacity building, life-cycle management and technical actions. The identification process is currently being implemented and tested for the first time. The findings and outcome will be evaluated by the ELIXIR Scientific Advisory Board in March 2017. Establishing the portfolio of ELIXIR Core Data Resources and ELIXIR Services is a key priority for ELIXIR and publicly marks the transition towards a cohesive infrastructure.