This article builds on critical theories of access and accessibility to explore the extent to which archival websites provide for the specific functional requirements of disabled visitors, and the methods that archival practitioners can adopt to further disabled inclusion in archival spaces. We take a mixed methods approach with an emancipatory ethos, combining accessibility testing of archival websites for the technical attributes that contribute to accessibility for those identifying as visually, sensorily, cognitively and physically disabled, and interviews with archival practitioners on their approaches to include and source support for inclusion in their respective institutions. Together, these methods construct an understanding of the current state of disabled inclusion in digital archival collections in the United Kingdom. The article demonstrates that while access is one of the main principles of archival theory, the nature of that access is often considered in its broadest sense, without considering the individual requirements of those engaging with archives. Defaulting to non-disabled visitors further marginalizes the disabled users with specific access requirements, thus rendering them invisible in the development of physical and virtual reading rooms. We encourage a mindset shift towards foregrounding the needs of disabled users, offering a framework around which other practitioners may shape their initiatives to work towards a more holistic, imperfect accessibility policy.
This article discusses the development of a Digital Humanities (DH) Association for the UK and Ireland. It explores processes for evidence gathering, methods for building community through consultation, approaches to defining values and purpose through collaboration, and how to practise openness that is both radical and responsible. It begins by outlining the landscape of DH in the UK and Ireland, highlighting differences and similarities between the two countries. Next, it addresses four key areas of focus in the planning for the new Association: community, consultation, and inclusivity; the importance of advocacy for DH and the role of a DH Association in national policy-making; the centrality of training and the development of career pathways in and from DH; and how to go about implementing a values-led organization. Finally, it reflects on the value of international collaboration in the field of DH, both between Ireland and the UK and among international subject associations and infrastructures.
Background Stakeholder input into Artificial Intelligence (AI) and Machine Learning (ML) is critical for creating AI systems that are both innovative and accountable. This paper examines READ-COOP (https://readcoop.eu), a platform cooperative that develops and hosts its own AI and ML Automated Text Recognition (ATR) tools (https://transkribus.org). This case study demonstrates an alternative cooperative governance model for technology innovation and creating responsible AI infrastructure. Methods We employ Reflection-In-Action and a qualitative Member questionnaire to document the development of READ-COOP from European Commission funded research project to independent cooperative. We assess the cooperative’s structure, management, and community engagement from 2019 to 2024, including membership dynamics, use, governance, and operational efficacy. Results As of October 2024, READ-COOP has 227 Members from 30 countries, and 235,000 registered User accounts. Transkribus has processed 90 million digital images of historical texts, demonstrating effective AI utilization in the cultural heritage sector, winning the European Union’s Horizon Impact Award 2020. The cooperative approach facilitates democratic decision-making, leading to sustainable growth, and significant stakeholder involvement. Qualitative feedback indicates high levels of satisfaction with the cooperative’s governance and the perceived integrity and utility of the AI infrastructure, while supporting exceptional engagement with technology innovation. Conclusions READ-COOP demonstrates that a cooperative business model has potential to sustain and support innovation in AI and ML infrastructures while promoting democratic participation and equitable ownership in particular contexts, in this case the cultural heritage sector. We suggest that cooperative frameworks are particularly suitable for AI infrastructures initially funded through public grants, providing a sustainable transition from public development to long-term, sustainable community-ownership, where focussed tool providers have adequate community support. We recommend wider application and exploration of cooperative models for innovation in AI and ML technologies for responsible creation, governance, and use, although we recognise READ-COOP’s unique context, community, and success may be an outlier.
Anonymised dataset from automated emails relating to the Transkribus Scholarship (Nov 2020 - Mar 2022), which encourages and supports Handwritten Text Recognition (HTR) work from students, workshop leads and ECRs: https://readcoop.eu/transkribus/scholarship/
PurposeThis paper focuses on image-to-text manuscript processing through Handwritten Text Recognition (HTR), a Machine Learning (ML) approach enabled by Artificial Intelligence (AI). With HTR now achieving high levels of accuracy, we consider its potential impact on our near-future information environment and knowledge of the past.Design/methodology/approachIn undertaking a more constructivist analysis, we identified gaps in the current literature through a Grounded Theory Method (GTM). This guided an iterative process of concept mapping through writing sprints in workshop settings. We identified, explored and confirmed themes through group discussion and a further interrogation of relevant literature, until reaching saturation.FindingsCatalogued as part of our GTM, 120 published texts underpin this paper. We found that HTR facilitates accurate transcription and dataset cleaning, while facilitating access to a variety of historical material. HTR contributes to a virtuous cycle of dataset production and can inform the development of online cataloguing. However, current limitations include dependency on digitisation pipelines, potential archival history omission and entrenchment of bias. We also cite near-future HTR considerations. These include encouraging open access, integrating advanced AI processes and metadata extraction; legal and moral issues surrounding copyright and data ethics; crediting individuals’ transcription contributions and HTR’s environmental costs.Originality/valueOur research produces a set of best practice recommendations for researchers, data providers and memory institutions, surrounding HTR use. This forms an initial, though not comprehensive, blueprint for directing future HTR research. In pursuing this, the narrative that HTR’s speed and efficiency will simply transform scholarship in archives is deconstructed.
Chapter Twenty-seven INFORMATIONAL ABUNDANCE AND MATERIAL ABSENCE IN THE DIGITISED EARLY MODERN PRESS: THE CASE FOR CONTEXTUAL DIGITISATION was published in The Edinburgh History of the British and Irish Press, Volume 1 on page 586.
This chapter addresses two key questions: what can the concept of the black box add to our understanding of library catalogues as data in digital library user studies? And how might data-driven approaches help us to increase the transparency of these black boxes and render them critically addressable? Libraries are complex systems comprising a complex interrelationship of staff, space, users and technical infrastructure. However, digital library user studies have not applied the same attention to the creation of large-scale datasets as they have to the ethical and methodological implications of reporting on them. This chapter positions the study of library catalogue data in relation to black box theory and the collections as data imperative. It argues that collaborations between data science, the critical digital humanities, and library and information science can help us to be more transparent in how we reuse catalogue data, and to redefine how this data is created, processed, and documented in the first place.
The UK-Ireland Digital Humanities Network is an AHRC/IRC-funded project to undertake research and consultation towards the implementation of a permanent Digital Humanities association for the UK and Ireland. This is the fourth discussion paper produced by the Network, in consultation with the wider Digital Humanities community in the two countries and beyond. It summarises the findings of the fourth workshop organised by the Network, and offers recommendations based on these findings.
Handwritten Text Recognition (HTR) technology is now a mature machine learning tool, becoming integrated in the digitisation processes of libraries and archives, speeding up the transcription of primary sources and facilitating full text searching and analysis of historic texts at scale. However, research into how HTR is changing our information environment is scant. This paper presents a systematic literature review regarding how researchers are using one particular HTR platform, Transkribus, to indicate the domains where HTR is applied, the approach taken, and how the technology is understood. 381 papers from 2015 to 2020 were gathered from Google Scholar, Scopus, and Web of Science, then grouped and coded into categories using quantitative and qualitative approaches. Published research that mentions Transkribus is international and rapidly growing. Transkribus features primarily in archival and library science publications, while a long tail of broad and eclectic disciplines, including history, computer science, citizen science, law and education, demonstrate the wider applicability of the tool. The most common paper categories were humanities applications (67%), technological (25%), users (5%) and tutorials (3%). This paper presents the first overarching review of HTR as featured in published research, while also elucidating how HTR is affecting the information environment.
This chapter addresses two key questions: what can the concept of the black box add to our understanding of library catalogues as data in digital library user studies?And how might datadriven approaches help us to increase the transparency of these black boxes and render them critically addressable?Libraries are complex systems comprising a complex interrelationship of staff, space, users and technical infrastructure.However, digital library user studies have not applied the same attention to the creation of large-scale datasets as they have to the ethical and methodological implications of reporting on them.This chapter positions the study of library catalogue data in relation to black box theory and the collections as data imperative.It argues that collaborations between data science, the critical digital humanities, and library and information science can help us to be more transparent in how we reuse catalogue data, and to redefine how this data is created, processed, and documented in the first place.
Purpose To date, there has been little research into users of the Legal Deposit Libraries (Non-Print Works) Regulations 2013. This paper addresses that gap by presenting key findings from the AHRC-funded Digital Library Futures project. Its purpose is to present a “user-centric” perspective on the potential future impact of the digital collections that are being created under electronic legal deposit regulations. Design/methodology/approach The study utilises a mixed methods case study of two academic legal deposit libraries in the United Kingdom: The Bodleian Libraries, University of Oxford; and Cambridge University Library. It combines surveys of users, web log analysis and expert interviews with librarians and cognate professionals. Findings User perspectives on NPLD were not fully considered in the planning and implementation of the 2013 regulations. The authors present findings from their user survey to show how contemporary tensions between user behaviour and access protocols risk limiting the instrumental value of NPLD collections, which have high perceived legacy value. Originality/value This is the first study to address the user context for UK Non-Print Legal Deposit. Its value lies in presenting a research-led user assessment of NPLD and in proposing “user-centric” analysis as an addition to the existing “four pillars” of legal deposit research.
The UK-Ireland Digital Humanities Network is an AHRC/IRC-funded project to undertake research and consultation towards the implementation of a permanent Digital Humanities association for the UK and Ireland. This report arises from an event on 'Next generation careers in and from Digital Humanities', organised for the Network by University College Cork, NUI Galway and the University of London. Co-located with the 10th Oxford Digital Humanities Summer School, it was held online on 15 July 2021.
To date, there has been relatively little discussion of how the UK doctoral funding landscape shapes digital humanities pedagogy for postgraduate research students. This article sets out to address this relative lack, by introducing the inter- and multi-disciplinary context in which many students in the UK work. We examine the phenomenon of students who are not necessarily interested in becoming DH practitioners, but have identified DH as a knowledge gap in their own disciplinary practice. Such a realisation changes the nature of the learner within DH communities of practice, requiring a different form of learning. This study therefore explores learning within a community of practice, the inter- and multi-disciplinary space in which digital humanities practitioners operate. First, drawing on the diverse disciplinary landscape, it highlights an individual's learning journey through self-determined learning (heutagogy). Second, it outlines an idea of digital humanities pedagogy for postgraduate research based on current frameworks of digital literacies and broader researcher development in the UK, framing research activity as learning. Third, it presents the DEAR model for learning and teaching design, which is based on four principles: Diversity; Employability; Application; and Reflection. Finally, it provides an evaluation of the DEAR model in the context of one UK Doctoral Training Partnership (DTP). It contributes to understanding of pedagogical practices for doctoral-level DH training and provides a set of recommendations for instructors to adopt and adapt these pedagogical principles in their own programmes.
This is the second discussion paper produced by the UK-Ireland Digital Humanities Network in consultation with the wider Digital Humanities (DH) Community in the two countries and beyond. It summarises the findings of the second workshop organised by the network, and offers recommendations based on these findings.