Related papers: BERT in Plutarch's Shadows
This article proposes a fresh and direct reading of foundational texts of philosophy and aims at bringing back the inflamed debates that are contemporaneous with the birth of Greek axiomatics, and indeed at understanding what is timeless in…
This project introduces BrAIcht, an AI conversational agent that creates dialogues in the distinctive style of the famous German playwright Bertolt Brecht. BrAIcht is fine-tuned using German LeoLM, a large language model with 7 billion…
Transformer-based language models, such as BERT and its variants, have achieved state-of-the-art performance in several downstream natural language processing (NLP) tasks on generic benchmark datasets (e.g., GLUE, SQUAD, RACE). However,…
These are notes and slides from a Pecha-Kucha talk given on March 6, 2013. The presentation tinkered with the question whether calculus on graphs could have emerged by the time of Archimedes, if the concept of a function would have been…
Annotated parallel text in Latin and English of the paper of Adam Adamandy Kocha\'nski "Solutio Theorematum Ab illustri Viro in Actis hujus Anni Mense Januario, pag. 28. propositorum", Acta Eruditorum, Lipsiae 1682, pp. 230-236, in which he…
The Sorites paradox is the name of a class of paradoxes that arise when vague predicates are considered. Vague predicates lack sharp boundaries in extension and is therefore not clear exactly when such predicates apply. Several approaches…
Ambiguities of natural language do not preclude us from using it and context helps in getting ideas across. They, nonetheless, pose a key challenge to the development of competent machines to understand natural language and use it as humans…
Isaac Newton, in popular imagination the Ur-scientist, was an outstanding humanist scholar. His researches on, among others, ancient philosophy, are thorough and appear to be connected to and fit within his larger philosophical and…
Large language models like GPT-4 have achieved remarkable proficiency in a broad spectrum of language-based tasks, some of which are traditionally associated with hallmarks of human intelligence. This has prompted ongoing disagreements…
This work describes experiments which probe the hidden representations of several BERT-style models for morphological content. The goal is to examine the extent to which discrete linguistic structure, in the form of morphological features…
This study investigates the internal mechanisms of BERT, a transformer-based large language model, with a focus on its ability to cluster narrative content and authorial style across its layers. Using a dataset of narratives developed via…
Recent works have demonstrated that multilingual BERT (mBERT) learns rich cross-lingual representations, that allow for transfer across languages. We study the word-level translation information embedded in mBERT and present two simple…
In 1615 Paolo A. Foscarini, a Carmelite monk lived in a monastery of south Italy near Cosenza (Calabria), published a Trattato which, at variance to what was common at the time, has not been written in Latin, but in volgare, the ancient…
Plutarchus, circa 100 AD, in his early book on "astrophysics" --in which he exposed, in a sense, a general theory of gravitation-- wrote the noticeable passage: <<The Moon gets the guarantee of not falling down just from its motion and from…
Much information available to applied researchers is contained within written language or spoken text. Deep language models such as BERT have achieved unprecedented success in many applications of computational linguistics. However, much…
The Greek fictional narratives often termed love novels or romances, ranging from the first century CE to the middle of the 15th century, have long been considered as similar in many ways, not least in the use of particular literary motifs.…
The inference of politically-charged information from text data is a popular research topic in Natural Language Processing (NLP) at both text- and author-level. In recent years, studies of this kind have been implemented with the aid of…
Obtaining large-scale annotated data for NLP tasks in the scientific domain is challenging and expensive. We release SciBERT, a pretrained language model based on BERT (Devlin et al., 2018) to address the lack of high-quality, large-scale…
As premodern texts are passed down over centuries, errors inevitably accrue. These errors can be challenging to identify, as some have survived undetected for so long precisely because they are so elusive. While prior work has evaluated…
Can large language models be trained to produce philosophical texts that are difficult to distinguish from texts produced by human philosophers? To address this question, we fine-tuned OpenAI's GPT-3 with the works of philosopher Daniel C.…