11/05/2018
Ancient Questions and Modern Answers:
Using Statistics to Describe Biblical Hebrew
Etienne van de Bijl, Master’s Student, Business Analytics,
Cody Kingham, Master’s Student, Biblical Studies and Digital Humanities
Our interdisciplinary research project, entitled “A Probabilistic Approach to Linguistic Variation and Change in Biblical Hebrew,” examines the differences in syntactic tendencies between various books in the ancient Hebrew Bible. Our method combines probabilistic methods from the business analytics world with the classical problem of textual origin in biblical studies. Doing this research has required us to bind and share knowledge between two very different disciplines. In week-to-week meetings, this often meant walking each other through new concepts, while seeking to use our expertise to address our research question. As a result, we each take new experiences back to our own disciplines. Here are some of the valuable insights we’ve gained.
From the Angle of Data Science
Roughly speaking, data science is all about gaining knowledge from data. Many applications of data science areas include marketing, manufacturing, telecommunications and many more! Therefore, applying data science to analyse biblical texts seems suitable. Two high-level goals in data science are predicting future events and/or describe objects of interest. Often, provided data sets are not consistent or contain a lot of missing values. As a data scientist, one has to deal with these kind of problems in an adequate manner. The data provided by the Eep Talstra Centre for Bible and Computer (Hebrew data) is well maintained, which makes it great to work with. The cooperation between a Biblical expert such as Cody and a data scientist as Etienne is fundamental for this kind of interdisciplinary research. On the one hand, linguistic knowledge and, as data scientist would refer, “Business knowledge” is required to do meaningful analysis on these texts, while on the other hand, methods and algorithms are required to capture and compare syntax.
From the Angle of Biblical Studies
The field of biblical studies is perhaps one of the oldest areas of academic interests in the western world. As such, every topic in this field is connected to many additional concerns that simultaneously touch on philosophy, theology, anthropology, history, and more. To bring in new methods, such as computational statistics, is to introduce something completely new to very old and relevant questions. And that is exciting! In our work, we are addressing the question of the Hebrew Bible’s textual origins, i.e. when was it written, or where did it come from. We do this by comparing the syntactic tendencies of texts, to see if there are differences that we can see with statistical models. So far we have found some very interesting things. And we are looking forward to presenting our final results.