Open research questions in Digital Humanities and Scholarship
248 unresolved questions extracted from the limitations and future-work sections of 3,304 Digital Humanities and Scholarship papers in our library. Each links back to the study that raised it.
What the literature leaves open
The potential absence of public domain books from regular stacks in online collections. The lack of digitization of certain titles despite mass digitization efforts.
other universities may want to investigate whether they have a similar percentage of public domain books that are not yet digitized
Our findings and methodology may be applicable to other domains, but we have not explored these.
Next integration milestone ChatGPT private/custom MCP connection Not yet established in this conversation; endpoint is ready for client integration.
Although it has received consideration in other disciplines, the conceptual issues remain comparatively underexplored in literary studies.
Further research could focus on improving the user experience. Further research could explore the application of the XML-TEI standard to other digital libraries.
There was a need for a more accessible platform for medieval Latin texts. There was a need to address previous issues with the MDL platform.
Introducing AI for specific tasks such as automatic translations and semantic indexing. Developing a free and optional log-in for habitual users. Improving the existing PoS determination and lemmatisation with AI tools.
The need for a comprehensive and interoperable resource for digital Latin philology. The lack of a unified environment for reading, searching, and analysing Latin texts.
There is a need for a digital glossary dedicated to the Latin lexicon of the Late Antique transition. The existing resources do not provide a comprehensive and systematic approach to the study of the Latin lexicon of this period.
The population of the LaLaLexiT glossary is still evolving. All entries published so far, as well as those forthcoming, are written in Italian. However, the project is open to contributions from any scholar wishing to submit one or more entries based on their research interests and in accordance with the glossary’s criteria and features. In this case, contributions in English, Spanish, French, or German will also be welcome.
The conversion process of digital versions of older print materials is famously error-prone, adding a lot of noise to the finished digital objects. The absence of a gold standard for evaluating the accuracy of NLP tasks. The need for more efficient and accurate methods for text analysis in digital humanities.
Bridging the gap between digital humanities and natural language processing: Text cleaning and NLP accuracy evaluation of a sample of 20th century Romanian novels · 2026 · DOIThe study suggests the potential for further improvements in OCR technology and NLP models. The study recommends redoing the OCR of the entire DMRN archive using an open-source OCR engine such as Tesseract. The study highlights the need for more research on the application of NLP techniques to large-scale digital archives.
Bridging the gap between digital humanities and natural language processing: Text cleaning and NLP accuracy evaluation of a sample of 20th century Romanian novels · 2026 · DOIFuture research should explore the implications of AI-native intellectual biography for various fields. Future research should develop new methodologies for analyzing the archive's provenance event.
AI-Native Intellectual Biography: A New Genre of AI-Mediated Reception — Provenance, Heteronymy, and the Archive That Outpaced Its Author · 2026 · DOIThe paper identifies a gap in the existing literature on AI-mediated reception. The gap is related to the lack of understanding of the archive's provenance event.
AI-Native Intellectual Biography: A New Genre of AI-Mediated Reception — Provenance, Heteronymy, and the Archive That Outpaced Its Author · 2026 · DOIThe need for guidelines for digital publishing in the context of the Corpus Vitrearum. The lack of control over published data on social media platforms.
The lack of a recoverable hyperlink grammar for field-based web discovery. The limitation of traditional hyperlinks in optimizing directness and immediacy.
Brown Hyperlinks / Burylinks: A Recoverable Hyperlink Grammar for Field-Based Web Discovery · 2026 · DOIComputational Literary Studies has traditionally focused on literary texts as primary objects of analysis, neglecting the institutional and performative structures of literary production.
Modeling Institutional Literary Networks: A Relational Infrastructure for Nineteenth-Century Theatre Archives · 2026 · DOIThe paper identifies the need to consider the complete process when designing a DR.
The heritage sector has real challenges that need to be addressed. There is a need for improving accessibility, interconnection, and relevance in the heritage sector.
HAICu's Innovation Labs: Bridging Technology and Heritage through Collaborative Problem-Solving · 2026 · DOIThe paper identifies a gap in the concept of entity resolution and composition layer. The paper identifies a gap in the use of AI-assisted execution in book writing.
The Parable of Mary Lee: Book Work Plan — From Dossier to ISBN (Master Specification v1.1) · 2026 · DOIHermeneutics underpins humanities and Information Systems (IS) research methodology, yet it remains understudied as a form of skilled labor across many professional domains, especially in the context of generative AI (GenAI).
The Tension Of Time And Attention: Generative Ai And Historians’ Hermeneutic Work · 2026The lack of a detailed inventory of Albert Szenci Molnár's library makes it challenging to reconstruct his legacy. The study requires a comprehensive overview of previous research on Molnár's library.
The paper identifies a gap in the understanding of the publishing industry in the region.- The study aims to fill this gap by analyzing archival materials.
Source base for publishing business study in Ternopil in late 19th century and first third of 20th century (based on materials from State Archive of Ternopil Region) · 2026 · DOIThe study faced challenges in terms of the limited accuracy of transcription, particularly for more demanding handwritten texts. The study faced challenges in terms of the potential for errors and over-interpretations in the outputs of the generative language models.
The Use of Generative Artificial Intelligence in the Digitisation of Printed and Manuscript Documents and Its Contribution to Historical and Archival Education · 2026 · DOI
Most-cited papers in Digital Humanities and Scholarship
- Quantitative Analysis of Culture Using Millions of Digitized Books · Science · 2010 · 1,785 citations
- KNOWLEDGE TRANSFER THROUGH INHERITANCE: SPIN-OUT GENERATION, DEVELOPMENT, AND SURVIVAL. · Academy of Management Journal · 2004 · 769 citations
- Bordering, Ordering and Othering · Tijdschrift voor Economische en Sociale Geografie · 2002 · 590 citations
- Taking different perspectives on a story. · Journal of Educational Psychology · 1977 · 369 citations
- Pump and Circumstance: Robert Boyle's Literary Technology · Social Studies of Science · 1984 · 343 citations
- Les "mondes lexicaux" et leur 'logique" à travers l'analyse statistique d'un corpus de récits de cauchemars · Langage et société · 1993 · 301 citations
- What is a ?document?? · Journal of the American Society for Information Science · 1997 · 170 citations
- Why do we need algorithmic historiography? · Journal of the American Society for Information Science and Technology · 2003 · 168 citations
- Originary Displacement · boundary 2 · 2000 · 149 citations
- Documents, practices and policy · Evidence & Policy · 2011 · 142 citations
Most recent work
- AI-Native Intellectual Biography: A New Genre of AI-Mediated Reception — Provenance, Heteronymy, and the Archive That Outpaced Its Author · Zenodo (CERN European Organization for Nuclear Research) · 2026
- Anchored Divergence: Immanent Critique as Survival Strategy Under Tail-Loss, and the Mandala as Limit Case (EA-SEI-ANCHDIV-01 v1.0) · Zenodo (CERN European Organization for Nuclear Research) · 2026
- The Interface Contest: Interface History Under the Gnostic Dialectic — A Four-Valent Computational Reanalysis (EA-SEI-DIALUX-02 v1.0) · Zenodo (CERN European Organization for Nuclear Research) · 2026
- Protocols for Scientific Training-Layer Literature: Machine-Mediated Research at the Production and Reception Ends (EA-SCI-TLL-PROTO-01 v2.1) · Zenodo (CERN European Organization for Nuclear Research) · 2026
- The Network Is the Poem: Why Topology Matters More Than Text Quality · Zenodo (CERN European Organization for Nuclear Research) · 2026
- Homer — Canon Provenance Node (The New Human Standing Canon) · Zenodo (CERN European Organization for Nuclear Research) · 2026
- water giraffes aren't real · Zenodo (CERN European Organization for Nuclear Research) · 2026
- A Use Case Lens on Digital Cultural Heritage · Journal on Computing and Cultural Heritage · 2026
- The International Journal of Literacies · The International Journal of Literacies · 2026
- The Epistemology of Algorithmic Narrative and the Problem of Creative Authenticity · Philosophy & Technology · 2026
Find a gap in your own Digital Humanities and Scholarship sub-topic
This page shows what the Digital Humanities and Scholarship literature already flags as unresolved. To narrow it to your specific question, run the guided finder — it searches the gap library on demand and checks candidates against 250M+ OpenAlex works.
Open the Research Gap Finder →