Digital Scholarship in the Humanities
Journal Data
| Type | JOURNAL |
|---|---|
| Publisher | Oxford University Press (GB) |
| ISSN | 2055-7671 / 2055-768X |
| Scopus | 21100465194 |
| Wikidata | Q42590839 |
| OpenAlex | S2734814886 |
| MAG | 2734814886 |
| Website | http://dsh.oxfordjournals.org |
| Total publications | 127 |
| Coverage period | 2014 - 2025 |
| Country | GB |
| Language | EN |
| Indexing | Scopus indexed |
| Cited by | 10 |
| Impact factor | 0.5 |
| SJR | 0.422 (Q1) |
| SNIP | 1.523 |
| CiteScore | 3 |
| h-index | 2 |
| 2-year mean citedness | 0.0351 |
| Female authorship share | 44.5% |
Digital Scholarship in the Humanities focuses on the application of computational methods to humanities research. While some anthropologists engage with digital methods, the journal's primary scope is methodological and technological within the broader humanities, making its direct relevance to social anthropology occasional
Linguistics and Language · Information Systems · Computer Science Applications · Advanced Text Analysis Techniques · Aesthetic Perception and Analysis · Authorship Attribution and Profiling · Computational and Text Analysis Methods · Digital and Traditional Archives Management · Digital Games and Media
Crossing boundaries through corpora: Innovative corpus approaches within and beyond linguistics. S. Buschfeld, P. Ronan, T. Neumaier, A. Weilinghoff, & L. Westermayer (eds)
Driven by breakthroughs in computational capabilities and the expansion of data-intensive paradigms, corpus-based methodologies are experiencing both disciplinary diffusion and substantial innovation in research design and application. However, new perspectives inevitably bring new questions, given the empirical turn of contemporary linguistic inquiry (cf Chomsky 1956), how can we integrate theory and corpora, and how do we bridge the gap between…
Growing and pruning the archive: An Agent-Based Model to Build Letter Correspondence Networks
The selection and availability bias concerning digital archives can be a problem for empirical studies. In the context of historical correspondence networks, we still have limited information about how missing data can affect and hinder the variables of our interest. Along these lines, we introduce an agent-based model to reconstruct historical communication using letters. We use this model to simulate the letter-sending process and its subsequen…
Research on the authorship identification of The Tale of Genji based on quantitative analysis
Numerous scholarly contributions have enriched our understanding of The Tale of Genji; however, the question of its authorship remains a subject of debate, with no definitive resolution to date. This study addresses this controversy by employing a rigorous quantitative statistical approach. We conducted a comprehensive analysis of various linguistic indicators within The Tale of Genji, encompassing aspects such as distribution of part-of-speech, …
Aigc empowers the sustainable development of traditional Chinese paper-cut
This research aims to explore the application of artificial intelligence-generated content (AIGC) technology in traditional Chinese paper-cut design and promote the protection and inheritance of traditional Chinese paper-cut culture. First, the paper-cut works of paper-cut artist are analyzed to extract the characteristic factors of her design styles. Second, based on the characteristics of the paper-cut style, a dedicated dataset for model train…
Challenges and solutions for the digital edition and geocoding of an 18th-century encyclopedia: The Diccionario histórico-geográfico de las Indias Occidentales
This article presents the methodological and technical challenges encountered in the creation of a digital edition of Antonio de Alcedo’s Diccionario histórico-geográfico de las Indias Occidentales (1786–1789), combined with the extraction and geocoding of its geographic content. Our work sought not only to recover the dictionary’s spatial information as structured data for a historical gazetteer, but also to preserve the text as a coherent docum…
Classification of Saikaku’s works using topic modelling
This study quantitatively analysed the thematic structure and temporal transitions in Saikaku Ihara’s ukiyo-zōshi using latent Dirichlet allocation (LDA), dynamic topic modelling (DTM), and Word2Vec. Previous studies have discussed the themes and genres of Saikaku’s works; however, comprehensive quantitative analysis has been lacking. Applying LDA to twenty-four works, this study examined topic transitions. The results confirmed previous research…
Integrating AI in historical research: A New Method to Analyzing National Narratives
This paper presents a human-machine collaboration method to integrating artificial intelligence (AI) into historical research, specifically in the analysis of national narratives. Using President Ronald Reagan’s portrayal of America as a case study, the paper demonstrates how AI tools like natural language processing, BERTopic., and generative large language models, when guided by historians, can streamline the analysis of national narratives, si…
Plume de nom: An experimental approach to handwriting analysis using OCR technology
The following article presents a case study using neural-network-based Optical Character Recognition (OCR) technology to revisit an unresolved problem regarding several manuscripts ascribed to the eminent Humanist Poggio Bracciolini. Granted OCR software is designed for the automatic transcription of text in a digital image, the experiment uses a side product produced by the OCR engine, the confidence score, to aid the comparison between differen…
Register variations in homepage bios of humanities scholars at Chinese universities and those in English-speaking countries
The scholar’s homepage has become an indispensable genre in scholarly life. Previous studies have explored techniques for extracting specific details from scholars’ homepage biographies, such as their publication history and educational background. However, there is a lack of research investigating the linguistic properties of homepage bios. Register refers to the linguistic variations that configure the contextual factors constraining language u…
Historical evolution of Chinese music aesthetics in the digital age
The research aims to analyze the evolution of characteristics of Chinese music within the context of the digitization of the music sphere. The study encompasses an analysis of changes in composition, performance, and dissemination of Chinese music influenced by digital technologies, as well as the transformation of traditional musical genres. Key aspects of the transformation of China’s musical landscape are examined, including the evolution of p…
Sibling-texts keyword analysis: Exploring Topic and Register Keywords
This article introduces a novel method for refining approaches to distant reading by proposing a procedure that categorizes prominent units (keywords) into two types: those pertinent to the topic and those associated with the genre/register of a text. This differentiation holds significant potential for more accurate modeling of topics and further applications in various domains of digital humanities. For example, register-related keywords may as…
Literary Simulation and the Digital Humanities: Reading, Editing, Writing. Manuel Portela
What if a book could evolve and change alongside its readers? Manuel Portela’s Literary Simulation and the Digital Humanities: Reading, Editing, Writing transforms this speculative question into a daring reality. By leveraging digital technology, Portela expands literature beyond static boundaries, transforming the acts of reading, editing, and writing into dynamic, interactive processes that invite active participation. This vision aligns seamle…
The most important affairs of the state are sacrifice and war: Digital visualization of the scattered early Zhou epics in the Book of Songs
The aim of this article is to discuss the characteristics of the culture of the Zhou Dynasty of China and try to find out some similarities at the early development of human civilization. For the research purpose, the author selected five epic poems about the early Zhou ethnicity scattered in the Book of Songs: Shengmin, Gongliu, Mian, Huangyi, and Daming, and established digital maps of the early Zhou epic, including a family relationship map an…
Artificial intelligence and machine learning in the preservation and innovation of intangible cultural heritage: Ethical Considerations and Design Frameworks
This study explores the application of artificial intelligence (AI) and machine learning in the preservation and innovation of intangible cultural heritage. Through analysis of implementations at major cultural institutions including the Palace Museum and Dunhuang Academy, the research demonstrates AI-driven preservation strategies achieving 90–95 per cent success rates in pattern recognition and 88 per cent accuracy in oral history preservation.…
Approaches to Digital Humanities Pedagogy: A systematic literature review within educational practice
Over the years, Digital Humanities Pedagogy has swiftly evolved, with a growing emphasis on the intersection between Computing and Humanities. However, in the associated literature, a lack of clarity has been found regarding the adoption of diverse approaches to DH pedagogy in different educational settings. This study tackles this research gap by analyzing the current state of knowledge in the field through a Systematic Literature Review. The ai…
Exploring the crisis motif in contemporary German-speaking, English-speaking, and Croatian literature—a digital and interdisciplinary approach
This article presents preliminary findings of the research project UIP-2020-02-3695 Analysis of Systems in Crisis and of New Consciousness in 21st Ct. Literature that is being financially supported by the Croatian Science Foundation. At the level of scientific contribution, the research consolidates theoretical insights into the phenomenon of permanent crisis of the modern society that is being projected through literature. This article provides …
Analysis of modern strategies for using artificial intelligence technologies in the creation of fantasy content
This study explores the integration of artificial intelligence (AI) in the creation of fantasy narratives, examining the new possibilities AI offers for authors and readers and its impact on the genre’s evolution. The study used literature review, textual analysis, comparative analysis, data analysis, experimental methods, ethical and legal examination, and data processing methods using Natural Language Processing software. Through analysis of fi…
Who wrote ‘The World of Saddam Hussein?’ A supervised machine learning approach
The 2003 novel ‘The World of Saddam Hussein’ (TWOSH), published under the pseudonym Mahdy Haidar, has long concealed the true identity of its author. This study compiles the works of eight authors previously speculated by literary critics as potential authors. Using a supervised machine learning approach with a large set of classification algorithms, I employed two distinct feature sets: (1) top frequent words and (2) top free-occurring function …
Cultural gene decoding: Digital Protection of Intangible Paper-Cut and Construction of Gene Bank
With globalization and modernization accelerating, the protection of intangible cultural heritage faces significant challenges. Chinese paper-cutting, a cultural treasure, has become a key focus due to its unique value and exquisite craftsmanship. However, inheritance is threatened by aging practitioners, skill loss, and market decline. This highlights the urgent need for digital preservation and the creation of a cultural gene bank. This paper p…
Sentiment analysis of Chinese ancient poetry based on multidimensional knowledge attention
Poetry was a unique literary genre in ancient China as an important way to express sentiments. Chinese ancient poetry not only has simple words, strict meters, and rich semantic relationships, but also widely use rhetorical techniques such as simile and personification, as well as metaphorical means such as allusion and imagery, which makes it difficult to understand their implicit sentiments quickly and accurately. Therefore, this article attemp…
Research on image design of Fujian paper-cut pattern based on Kansei engineering and WOA-BP neural network
In order to design Fujian paper-cut patterns that meet the perceptual needs of consumers and better inherit and develop them in modern society, a Fujian paper-cut pattern image design method based on perceptual engineering and the Whale Optimization Algorithm optimized BP neural network (WOA-BP) neural network is proposed. First, based on the theory of Kansei engineering, six representative paper-cut pattern samples and their main modeling featur…
Two sides of the same coin? Cross-linguistic sentiment comparison and thematic discovery of reader’s reception of Wolf Totem
Despite growing attention to the global dissemination of cross-cultural literature, rigorous research on cross-linguistic sentiment analysis remains limited, especially for works like Wolf Totem, which hold both worldwide reach and cultural depth. This study conducted a cross-language sentiment and thematic analysis of online book reviews by Chinese and English readers, utilizing advanced linguistic techniques such as BERT, MAXQDA, and ProtAnt. T…
A note on applying the Syuzhet program to film dialogue
The Syuzhet program is tested in two particular cases, and produces highly questionable results
Stylometry at the service of history of science: The Renaissance of Copernicus
A recent study, following a subtle but unusual for history of science argumentation method starting from the premises established by stylometry, discovered a drastic stylistic contrast between Copernicus’s early opus Commentariolus and his mature writings. The finding challenged the long-established view that Copernicus became a humanistically minded scholar early in his life and composed Commentariolus between 1509 and 1514. The present study ve…
Code review in digital humanities
Software and computational methods offer tremendous possibilities for digital humanities research, both accelerating existing work and opening up entirely new questions. However, software also has the potential to introduce new kinds of errors into the research workflow. How do we know that the software developed for a digital humanities project is error free and does what we think it does? Code review is a widespread technique to improve softwar…
Digital Humanities in the Anthropocene
This keynote address for the 2014 Digital Humanities conference is a practitioner's talk, and-though the abstract belies it-an optimistic one. I take as given the evidence that human beings are irrevocably altering the conditions for life on Earth and that, despite certain unpredictabilities, we live at the cusp of a mass extinction. What is the place of digital humanities (DH) practice in the new social and geological era of the Anthropocene? Wh…
Defining digital humanities and examining its relationship with linguistics through the lens of Digital Scholarship in the Humanities
Digital humanities (DH) is an emerging interdisciplinary academic field that has gained prominence in recent decades. This study explores the evolution of topics, research impact, and attractiveness of DH through the lens of the journal Digital Scholarship in the Humanities (DSH), a leading platform for DH research, from 1986 to 2023 (in three phases: 1986–2003, 2004–2014, and 2015–2023). The study also examines the role of linguistic research in…
Gecem Project Database: A digital humanities solution to analyse complex historical realities in early modern China and Europe
The GECEM Project Database stands out as a new Digital Humanities solution to accurately order and analyse the new historical Big Data gathered in Chinese and European historical archives. Traditional challenges such as capture, storage, analysis, data curation, searching, sharing, transfer, visualization, querying, updating, and information privacy are being tackled and solved within the design of this new multi-relational database. The implemen…
Disentangling semantic and prosodic features of English poetry
The distinction between genre and form is still contested in literary studies. While scholars associated with the New Formalism are criticized for perceiving everything as a form, digital humanists tend to argue that everything is a genre. In this research, we employed machine learning models to classify 36,635 English poems in the Chadwyck-Healey Literature Collections into twenty-seven categories, focusing on their semantic features (lexicons) …
Digital humanities approach to analyzing the roles and military power of Supreme Commanders and Grand Coordinators in the Ming Dynasty: A computational analysis of Ming Shilu
This article represents a digital humanities research endeavor that attempts to explore the roles and military power of Supreme Commanders and Grand Coordinators in the Ming Dynasty, employing computational analysis of the Ming Shilu. By leveraging a semi-supervised text classification framework to identify military paragraphs without needing prior annotation and generating heat maps based on location entities identified in the text, we discerned…
Digital Humanities in Poland from the Perspective of the Historical Linguist of the Polish Language: Achievements, Needs, Demands
The article presents the achievements of digital humanities in Poland, draws attention to the needs related to the development of digitization, and points to possible future undertakings aimed to popularize the accomplishments in this field. Apart from digitization, issues such as methods of sharing historic and linguistic sources in the digital form are covered. These sources include historic and scientific dictionaries of the Polish language, r…
Representing stories as interdependent dynamics of character activities and plots: A two-mode network relational event model
Recent advances in data science and machine learning have enhanced our ability to analyze and understand the structure of social interactions in fictional stories by using formal and quantitative approaches. However, an objective assessment of these aspects of fictional stories remains a relatively new and technically difficult field. In this brief report, we introduce our study in which we modeled story dynamics from a novel perspective. By impl…
Technical and methodological foundations of digital indexing of medieval and early modern court books
Ever since the beginnings of the modern historiography, the court books have posed a challenge for editors in Poland, both due to their number and variety. They constitute one of the richest sources enabling a variety of historical research. The publication of the sources’ content can be shared owing to new approaches stemming from constant development of IT tools and their application in the humanities. The solution proposed in our article is a …
Dynamic evolution of sentiments in Never Let Me Go: Insights from multifractal theory and its implications for literary analysis
The moods, feelings, and attitudes represented in a novel will resonate in the reader by activating similar sentiments. It is generally accepted that sentiment analysis can capture aspects of such moods, feelings, and attitudes and can be used to summarize a novel’s plot in a story arc. With the availability of a number of algorithms to automatically extract sentiment-based story arcs, new approaches for their utilization becomes pertinent. We pr…
The integration of heterogeneous information from diverse disciplines regarding persons and goods
This article presents a relational database capable of integrating data from a variety of types of written sources as well as material remains. In response to historical research questions, information from such diverse sources as documentary, bioanthropological, isotopic, and DNA analyses has been assessed, homogenized, and situated in time and space. Multidisciplinary ontologies offer complementary and integrated perspectives regarding persons …
Standards and quantification of coin iconography: Possibilities and challenges
The use of digital technologies and big data in the humanities and social sciences provided many opportunities for cultural heritage management and research, enabling data sharing and interdisciplinary collaborations. These developments increased the need for standardized data formats. General and domain-specific standards for describing and classifying cultural data, based on linked data principles, are developed to support increasingly numerous…
Existe correlación entre importancia y centralidad? Evaluación de personajes con redes sociales en obras teatrales de la Edad de Plata
The objective of this study is to answer the question: Do the central nodes of a social network of characters correspond with the protagonists of a theater play? To answer this question we evaluate different measures of centrality along with other textual quantitative values in relation to other manually annotated metadata on a corpus of twenty five dramatic plays of the Spanish theatre of the Silver Age (1868-1936). The found results show that c…
Of narrative time and space: Geography meets history via linguistics
The article explores issues of narrative time and space. It embraces a conception of geography, of space, as place involving relations among people, with ‘their own stories to tell’. And as story, as narrative, geography can be captured by a ‘story grammar’: Who, What, When, Where, Why, and How (the 5 Ws + H). When and Where, time and space, are the fundamental axes of narrative, different cultures differently grounding narrative in time (the ‘on…
Finding our way home: A theory and pedagogy of anti-colonial Q-mapping
This brief-report details my experiences teaching critical theory at the intersection of race/ class/ gender/ sexuality/ nationality through analogue mapping exercises from my tutorial room at Murdoch University in Western Australia to a group mapping project at the 2018 Digital Humanities Summer Institute (DHSI) at the University of Victoria, Canada (UVic). Given the proximity of both countries to the history of the British Empire and current af…
Tracking causal relations in the news: Data, tools, and models for the analysis of argumentative statements in online media
Online debates and debate spheres challenge our assumptions about democracy, politics, journalism, trust, and truth in ways that make them a necessary object of study. In the present article, we argue that the study of online arguments can benefit from an interdisciplinary approach that combines computational methods for text analysis with conceptual models of opinion dynamics. The article thereby seeks to make a conceptual and methodological con…
Gecem Project Database: A digital humanities solution to analyse complex historical realities in early modern China and Europe
The GECEM Project Database stands out as a new Digital Humanities solution to accurately order and analyse the new historical Big Data gathered in Chinese and European historical archives. Traditional challenges such as capture, storage, analysis, data curation, searching, sharing, transfer, visualization, querying, updating, and information privacy are being tackled and solved within the design of this new multi-relational database. The implemen…
Gender relations in Spanish theatre during the Silver Age: A quantitative comparison of works in the Spanish Drama Corpus
One of the many changes witnessed by Spanish society at the beginning of the 20th century was the early reshaping of the role of women, including in the realm of theatre. During the first three decades of the new century, Spanish theatre was thriving, favouring the emergence of new gender roles: there were new female playwrights, professional actresses, stage designers, costume designers, theatre company directors, etc. Against this background, i…
Modelling Chinese contemporary calligraphy: The WRITE data model
This article presents the WRITE data model and dataset, a comprehensive collection of Chinese contemporary calligraphic data, utilizing Linked Open Data (LOD) principles. Calligraphy plays a pivotal role in Chinese culture, reflecting national identity and cultural transformations. The objective of this study is to enhance understanding and provide new tools for exploring Chinese contemporary calligraphy through LOD. The WRITE data model comprise…
Rhetoric behind the digital screen: Wayfinding across the splinternet of AI—a rhetorical quartet of an affective writer
Drawing on and recontextualizing Elizabeth Forbes’ theory regarding the quartile facets of a developing writer—(1) maker, (2) artist, (3) creator, and (4) performer, and Karen Lunsford’s approach of wayfinding—(1) worlds apart, (2) literacy in the wild, (3) ecologies and networks, and (4) transfer in writing studies, this study, inspired by the notion of affective economy and splinternet, offers a rhetorical quartet pertaining to the literary car…
Can machine translation of literary texts fool stylometry
This article uses standard authorship-attribution stylometry to tell machine translations made with DeepL and Google Translate from human translations. This is done using a Burrows-like distance measure procedure of cluster analysis, later visualized through network analysis. Using a corpus of French literary classics translated into English by humans and machines as illustration, this article shows that, in most cases, translations of each text …
Lexical richness viewed through lexical diversity, density, and sophistication
Recognizing vocabulary as the essential language foundation and determining its richness continues to be a challenge in quantitative linguistics. This issue has also been widely applied in various practical contexts, including language acquisition, language change, psychology, and cognitive science. However, vocabulary is complex, and lexical richness has been typically measured using three distinct indicators: lexical diversity, density, and sop…
Revolutionizing the stage: Exploring the Multidimensional Landscape of Digital Theater
This study explores the transformative effects of digital technologies on the traditional stage and examines the multifaceted nature of digital theater (DT). It clarifies the concept of DT by identifying its essential attributes and establishing an analytical framework for its classification, based on the interplay of three critical dimensions: artistic creation, digital technology, and audience experience. This article uses this analytical frame…
ChatGPT as speechwriter for the French presidents
Generative artificial intelligence proposes several large language models (LLMs) to automatically generate a message in response to users’ requests. Such scientific breakthroughs promote new writing assistants but with some fears. The main focus of this study is to analyze the written style of one LLM called ChatGPT by comparing its generated messages with those of the recent French presidents. To achieve this we compare end-of-the-year addresses…
Documentary’s expanded fields: New media and the twenty-first-century documentary. Jihoon Kim
What constitutes a documentary? Film theorists have argued the question for decades, but Jihoon Kim takes an expansive definition to reorganize and shape the various media forms that have been grouped under the heading “documentary.” In his new book “The Documentary's Expanded Fields: New Media and the Twenty-First-Century Documentary,” Kim moves discussions beyond the traditional documentary film into a consideration of gallery installations, ac…
The ongoing birth of the narrator: Empirical Evidence for the Emergence of the Author–narrator Distinction in Literary Criticism
This article explores the historical evolution of the distinction between author and narrator in German-language literary criticism, an area largely unexplored by quantitative methods. While narratologists often distinguish between a fictional narrator and the author, the practical adoption of this distinction by readers remains under-examined. We hypothesize a semantic shift in the term ‘narrator’ from referring to the actual author to a fictive…
Beyond the surface: Stylometric analysis of GPT-4o’s capacity for literary style imitation
This study aims to explore the ability of GPT-4o to imitate the literary style of renowned authors. Ernest Hemingway and Mary Shelley were selected due to their contrasting literary styles and their overall impact on world literature. Using three distinct prompting strategies—zero-shot generation, zero-shot imitation, and in-context learning—we generated forty-five stylistic imitations and analyzed them alongside the authors’ original texts. To e…
Detecting authorship between generative AI models and humans: A Burrows’s Delta approach
The distinction between artificial intelligence (AI)- and human-generated texts has become increasingly significant with the emergence of ChatGPT and other generative AI models, which have garnered millions of users. In this study, we assess Burrows’s Delta, a well-established algorithm in authorship attribution, as a potential AI detector in argumentative essays. Our results prove that Burrows’s Delta is an effective tool for detecting AI-genera…
From coin to data: The Impact of Object Detection on Digital Numismatics
In this work, we investigate the application of advanced object detection techniques to digital numismatics, focusing on the analysis of historical coins. Leveraging models such as Contrastive Language-Image Pre-training (CLIP), we develop a flexible framework for identifying and classifying specific coin features using both image and textual descriptions. By examining two distinct datasets, modern Russian coins featuring intricate ‘Saint George …
Synergizing structure and semantics: A Knowledge Graph-Transformer Framework for Narrator Disambiguation in Hadith Networks
Historical transmission chains (isnads) are fundamental to verifying authenticity in Hadith literature, yet narrator identity resolution is a persistent challenge due to onomastic ambiguity and complex naming conventions. While traditional methods lack scalability and modern language models overlook crucial network structures, this study bridges the gap by synergizing structural and semantic information. We introduce a novel hybrid framework that…
On audiences’ feelings and needs of Hero: A Digital-Intelligent Humanities Perspective
Audiences’ reviews are critical to the reception study of movies. This article takes the reviews of Hero as an example, and analyzes the reception effect from a digital-intelligent humanities perspective with transformer-based models, especially bidirectional encoder representations from transformers (BERT)-based sentiment analysis and BERTopic modeling, which are generally regarded as the state-of-the-art deep learning models. The results of sen…