Related papers: BERT in Plutarch's Shadows
The ability to transmit and receive complex information via language is unique to humans and is the basis of traditions, culture and versatile social interactions. Through the disruptive introduction of transformer based large language…
The Pythagorean Theorem is one of the oldest, more famous and more useful theorems of Mathematics, and possibly the one that has had the most impact in the evolution of this and other sciences. In this article, we look at it from different…
It will be shown in this article that an ontological approach for some problems related to the interpretation of Quantum Mechanics could emerge from a re-evaluation of the main paradox of early Greek thought: the paradox of Being and…
Language is an outcome of our complex and dynamic human-interactions and the technique of natural language processing (NLP) is hence built on human linguistic activities. Bidirectional Encoder Representations from Transformers (BERT) has…
In this paper we concentrate on the nature of the liar paradox as a cognitive entity; a consistently testable configuration of properties. We elaborate further on a quantum mechanical model [Aerts, Broekaert, Smets 1999] that has been…
As a celebration of the \emph{Tractatus} 100th anniversary it might be worth revisiting its relation to the later writings. From the former to the latter, David Pears recalls that ``everyone is aware of the holistic character of…
Prior work on scientific question answering has largely emphasized chatbot-style systems, with limited exploration of fine-tuning foundation models for domain-specific reasoning. In this study, we developed a chatbot for the University of…
This article is the first part of a study of the so-called 'mathematical part' of Plato's Theaetetus (147d-148b). The subject of this 'mathematical part' is the irrationality, one of the most important topics in early Greek mathematics. As…
The 2020s have been witnessing a very significant advance in the development of generative artificial intelligence tools, including text generation systems based on large language models. These tools have been increasingly used to generate…
In the past several decades, many authorship attribution studies have used computational methods to determine the authors of disputed texts. Disputed authorship is a common problem in Classics, since little information about ancient…
Though state-of-the-art sentence representation models can perform tasks requiring significant knowledge of grammar, it is an open question how best to evaluate their grammatical knowledge. We explore five experimental methods inspired by…
Transformer based pre-trained models such as BERT and its variants, which are trained on large corpora, have demonstrated tremendous success for natural language processing (NLP) tasks. Most of academic works are based on the English…
We explore the technical details and historical evolution of Charles Peirce's articulation of a truth table in 1893, against the background of his investigation into the truth-functional analysis of propositions involving implication. In…
With a growing number of BERTology work analyzing different components of pre-trained language models, we extend this line of research through an in-depth analysis of discourse information in pre-trained and fine-tuned language models. We…
In this work, we employ quantitative methods from the realm of statistics and machine learning to develop novel methodologies for author attribution and textual analysis. In particular, we develop techniques and software suitable for…
Four centuries before modern statistical linguistics was born, Leon Battista Alberti (1404--1472) compared the frequency of vowels in Latin poems and orations, making the first quantified observation of a stylistic difference ever. Using a…
The question of the origins of logic as a formal discipline is of special interest to the historian of physics since it represents a turning inward to examine the very nature of reasoning and the relationship between thought and reality. In…
We introduce AnnualBERT, a series of language models designed specifically to capture the temporal evolution of scientific text. Deviating from the prevailing paradigms of subword tokenizations and "one model to rule them all", AnnualBERT…
We report our models for detecting age, language variety, and gender from social media data in the context of the Arabic author profiling and deception detection shared task (APDA). We build simple models based on pre-trained bidirectional…
In previous work, it has been shown that BERT can adequately align cross-lingual sentences on the word level. Here we investigate whether BERT can also operate as a char-level aligner. The languages examined are English, Fake-English,…