PulseAugur
EN
LIVE 19:38:54

Computational linguistics studies analyze lexical transmission in religious texts · 4 sources tracked

Two new computational linguistics studies analyze lexical transmission and stylistic features within religious texts. The first paper examines Bengali and Sanskrit devotional literature from the 8th to 19th centuries, using TF-IDF and cosine similarity to quantify vocabulary overlap between Buddhist, Shakta, and Vaishnava traditions. The second study applies stylometric analysis to English translations of the Buddhist Pali Canon, comparing Sutta, Vinaya, and Abhidhamma texts using Zipf distributions, TTR, and vocabulary overlap metrics. AI

IMPACT These studies demonstrate novel applications of computational linguistics and stylometry for analyzing historical and religious texts, potentially informing future research in digital humanities and computational social science.

RANK_REASON The cluster contains two academic papers published on arXiv detailing computational linguistic analysis of religious texts. [lever_c_demoted from research: ic=5 ai=0.4]

Read on arXiv cs.IR (Information Retrieval) →

AI-generated summary · Google Gemini · from 5 sources. How we write summaries →

Computational linguistics studies analyze lexical transmission in religious texts · 4 sources tracked

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
The cluster contains two academic papers published on arXiv detailing computational linguistic analysis of religious texts. [lever_c_demoted from research: ic=5 ai=0.4]
Source corroboration
5 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
paper, other
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
Standard
On-topic for AI-industry coverage; kept in the public index.
Story freshness
94 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

Full methodology in our editorial standards.

COVERAGE [5]

  1. arXiv cs.CL TIER_1 English(EN) · Joy Bose ·

    From Vajrayana Tara to Bengali Baul: A Computational Study of Lexical Transmission Across Buddhist, Shakta, and Vaishnava Traditions in Bengal

    arXiv:2606.26803v1 Announce Type: new Abstract: We present a computational corpus study of vocabulary relationships across eight tradition layers of Bengali and Sanskrit devotional literature spanning the 8th to 19th centuries, encompassing Buddhist Vajrayana, Shakta Tantra, Vais…

  2. arXiv cs.IR (Information Retrieval) TIER_1 English(EN) · Joy Bose ·

    From Vajrayana Tara to Bengali Baul: A Computational Study of Lexical Transmission Across Buddhist, Shakta, and Vaishnava Traditions in Bengal

    We present a computational corpus study of vocabulary relationships across eight tradition layers of Bengali and Sanskrit devotional literature spanning the 8th to 19th centuries, encompassing Buddhist Vajrayana, Shakta Tantra, Vaishnava, and Baul traditions. Using a corpus of 75…

  3. arXiv cs.CL TIER_1 English(EN) · Joy Bose ·

    Three Buddhist Vocabularies: Computational Stylometry of the English Pali Canon across Sutta, Vinaya, and Abhidhamma

    arXiv:2606.25372v1 Announce Type: new Abstract: We present a computational stylometric analysis of the Tipitaka across all three Pitakas in English translation, extending earlier work on the Sutta Pitaka alone. The corpus spans 134,831 segments from Bhikkhu Sujato's Sutta Pitaka …

  4. arXiv cs.IR (Information Retrieval) TIER_1 English(EN) · Joy Bose ·

    Three Buddhist Vocabularies: Computational Stylometry of the English Pali Canon across Sutta, Vinaya, and Abhidhamma

    We present a computational stylometric analysis of the Tipitaka across all three Pitakas in English translation, extending earlier work on the Sutta Pitaka alone. The corpus spans 134,831 segments from Bhikkhu Sujato's Sutta Pitaka (114,591 segments, CC0), Bhikkhu Brahmali's Vina…

  5. Hugging Face Daily Papers TIER_1 English(EN) ·

    Three Buddhist Vocabularies: Computational Stylometry of the English Pali Canon across Sutta, Vinaya, and Abhidhamma

    We present a computational stylometric analysis of the Tipitaka across all three Pitakas in English translation, extending earlier work on the Sutta Pitaka alone. The corpus spans 134,831 segments from Bhikkhu Sujato's Sutta Pitaka (114,591 segments, CC0), Bhikkhu Brahmali's Vina…