Papers for

digital humanities teams

Papers whose findings have a practical use for this group, as judged from the abstract. Open a paper to read what it means in practice.

Computational method finds biblical references in karen blixen stories

Retrieving Biblical Intertextual References in Karen Blixen's Seven Gothic Tales

Abstract: Identifying intertextual references is central to literary scholarship, but computationally difficult when source material is transformed through paraphrase, allusion, historical language, and translation. We investigate this problem through biblical intertextuality in Karen Blixen's Seven Gothic Tales. Drawing on the commentary to a critical edition, we construct a benchmark of 189 annotated references and evaluate retrieval against all 31,170 verses of historically plausible Danish Old and New Testament translations. We compare TF-IDF and BM25 with multilingual and Danish sentence encoders, examine the effect of linguistic normalization, and fine-tune a Danish encoder using hard negatives and five-fold cross-validation. We analyze performance across automatically derived lexical-overlap strata representing quotations, paraphrases, and allusions. Linguistically normalized BM25 provides a strong zero-shot baseline, attaining an overall R@10 of 0.365 and retrieving every quotation within its ten highest-ranked verses. The best zero-shot dense model achieves a comparable overall score of 0.360 while performing better on allusions. Fine-tuning DFM-large raises its overall R@10 from 0.265 to 0.508 and more than doubles its performance on allusions, from 0.138 to 0.339. However, evaluation against editorial annotations alone understates the model's scholarly usefulness: a literary scholar judged seven of 30 selected rank-one predictions counted as false positives to be meaningful additional references. These findings show both the potential and the epistemic limits of computational intertextual retrieval. Rather than treating scholarly annotations as exhaustive or model outputs as discoveries, we propose retrieval models as heuristic co-readers that recover documented references and generate candidates for expert-led close reading.

Mon 28 SeptComputation and Language
The gist
Finding hidden references to the Bible in literature can be hard, especially when the text doesn't quote directly but hints or paraphrases. The authors worked on spotting these references in Karen Blixen's Seven Gothic Tales by comparing story passages to old Danish Bible translations. They tested different methods and improved one by training it with examples, which helped spot many more subtle references. Their approach serves as a helpful tool to suggest possible connections for scholars to explore further, rather than replacing expert reading.
Open → 2609.35765v1

Modeling chinese character evolution using manifold learning and neural ODEs

Tracing the Evolution of Oracle Bone Characters Across Three Millennia

Abstract: Of the approximately 4,500 Oracle Bone Inscription (OBI) characters discovered from the Shang dynasty, only about 1,600 have been deciphered. Many computational approaches compare OBI with glyphs from one historical period at a time. However, during the evolution of Chinese characters, significant structural or semantic changes often occur in uncertain dynasties. A single-period reference may be insufficient when relevant forms change substantially between observed eras. Therefore, we propose the \textbf{Manifold-based Script Evolution Framework (MSEF)}, a framework that models the evolution series (OBI, Bronze, Seal, Clerical, Regular) of Chinese characters as the continual evolution of a manifold space. MSEF represents each character as an era-specific manifold point and learns continuous inter-era transition rules via Neural Ordinary Differential Equations. Both manifold space and transition dynamics can be trained end-to-end through character evolution pairs across any two eras.

Mon 28 SeptComputation and Language
The gist
Many ancient Chinese characters from oracle bones are hard to understand because their forms changed a lot over thousands of years. The authors created a new method that represents each character’s shape in different historical eras as a point in a special space, tracking how those points change continuously over time. They use a kind of neural network that can model smooth transformations to learn how characters evolved from one era to the next. This helps link ancient characters with their later versions, improving understanding of their development.
Open → 2609.35674v1

Agents improve alignment of classical texts and translations with fewer defects

When Do Agents Help? Embedding, LLM and Agentic Alignment of Classical Texts and Their Translations

Abstract: Classical texts aligned with their translations support machine translation, retrieval and computational research, but evidence comparing alignment workflows is scattered. This study compares seven systems on 452 texts in Pali, Sanskrit, Mishnaic Hebrew and Tibetan, comprising 9,833 human-aligned units: four embedding pipelines, a direct LLM call, an autonomous agent, and the agent revised by an independent auditor. Generative workflows recover 93-94% of reference correspondences, against at most 77% for embeddings. A ceiling analysis shows that sentence boundaries make some references unrepresentable by the embedding pipelines. Reference recovery is similar across generative workflows: the agent's advantage is 0.5 percentage points (95% CI -0.02 to 1.17), and auditing adds no established benefit. Agents nevertheless produce structurally valid output for all 452 texts, against 437 for direct calls. A blinded three-LLM panel assesses every generative mismatch against the source and human reference. Most mismatches are labelled defensible editorial variation; consensus major-error labels cover only 0.06-0.14% of units. The panel labels significantly fewer residual defects for agents than direct calls (0.7% versus 1.4%), suggesting that reference recovery alone understates alignment quality. On ten long Pali discourses taken as published online, agents and audited agents raise recovery from the direct call's 71% to 84% and 92%. Identical reference-located chunks bring all three to 93%. Agents thus improve structural reliability and reduce judged defects on short passages, while their large recovery advantage on long documents disappears after chunking. In this setting, independent auditing offers little measurable additional benefit on prepared passages.

Sun 27 SeptComputation and Language
The gist
Matching ancient texts with their translations helps computers translate and study these works, but it is hard to do accurately. The authors compare several computer methods to line up passages in original texts with their translations. They find that AI agents can recover more matching parts than standard embedding techniques and make fewer mistakes. However, the advantage of agents is smaller when long texts are split into chunks. Having a second check by another AI adds little benefit. Overall, agents produce more reliable and better-aligned outputs.
Open → 2609.33691v1

Tracing how researchers track changing ideas in physics fields

Tracing individual knowledge trajectories in a changing field: the case of general relativity and gravitation

Abstract: Historians have reconstructed the twentieth-century transformation of general relativity and gravitation (GRG) at the field level and through individual careers, but connecting these scales requires a way to compare researchers with the changing field over time. We develop such a comparison, setting a researcher's publications and references against GRG field literature from the same, earlier, and later two-year periods. Building on Own Vocabulary and Embedding Density Estimation from our earlier two-case study (arXiv:2501.00391), we extend the analysis to the fifty most-published authors in a NASA/ADS corpus of about 180,000 GRG records (1911 to 2000) and add two citation-based measures, Referenced Vocabulary and Citation Identity. The four measures compare an author's written language, cited literature, semantic neighbourhood, and cited-authority configuration with the surrounding field. The earlier cases suggested that closer field-vocabulary alignment accompanies a denser semantic neighbourhood. Across the fifty authors this holds only partially. Written and cited vocabularies tend to move together, usually resembling later GRG literature as the field turned towards astrophysical and cosmological research. Semantic neighbourhoods more often lie where the field's publications were concentrated in earlier periods, while co-citation patterns follow no single temporal direction, and the two citation measures frequently place the same researcher differently despite drawing on identical reference lists. Individual trajectories can thus combine vocabulary tied to later field states with older semantic or citation structures, and these divergent cases mark patterns for closer historical investigation. The approach transfers to other fields with defensible corpus boundaries and adequate coverage of texts, references, and disambiguated author identities.

Wed 16 SeptComputation and Language
The gist
Understanding how scientists' work connects to the bigger field over time is hard because the field itself changes. The authors develop new ways to compare what a researcher writes and cites with the wider field of physics literature at different times. They studied 50 top authors in general relativity and gravitation and found that authors’ word choices and citations don’t always move in sync with the field’s evolution. This method helps spot unique patterns in researchers’ work and could be applied to other fields with enough text and citation data.
Open → 2609.18697v1

Luke is confirmed as author of both gospel and Acts books

Is Luke the Author of a Gospel and the Acts of the Apostles?

Abstract: According to Christian tradition, Luke is credited with authoring a Gospel and the Acts of the Apostles, even if his name does not appear in either book, both originally written in Koine Greek. Several biblical scholars assume that both texts were written by a common author, while others deduce the presence of two authors. Different studies have been found to support either finding, some based on qualitative evaluation, while a few others consider the occurrence frequency differences between the two books. To propose an enhanced quantitative analysis, this study is grounded on two recent authorship attribution models. The Burrows' Delta, applied with eleven different feature sizes, demonstrates common authorship. An author verification model confirms this finding. The following experiments consider several stylistic representations, feature sizes, and distance functions to confirm that Luke is the true author of both books.

Tue 15 SeptComputation and LanguageArtificial IntelligenceDigital Libraries
The gist
There has been debate about whether Luke wrote both a Gospel and the Acts of the Apostles because his name is not directly on these texts. Some scholars think there was one author, others think two different ones wrote them. The authors used computer-based methods to analyze writing style and vocabulary frequency in both books. Their analysis strongly supports that the same person, likely Luke, wrote both.
Open → 2609.17762v1