中文
相关论文

相关论文: Scientific evaluation of Charles Dickens

200 篇论文

The Gutenberg Literary English Corpus (GLEC) provides a rich source of textual data for research in digital humanities, computational linguistics or neurocognitive poetics. However, so far only a small subcorpus, the Gutenberg English…

计算与语言 · 计算机科学 2020-10-22 Arthur M. Jacobs , Annette Kinder

Corpus-based statistical analysis plays a significant role in linguistic research, and ample evidence has shown that different languages exhibit some common laws. Studies have found that letters in some alphabetic writing languages have…

计算与语言 · 计算机科学 2020-06-03 Qinghua Chen , Yan Wang , Mengmeng Wang , Xiaomeng Li

As the use of AI tools by students has become more prevalent, instructors have started using AI detection tools like GPTZero and QuillBot to detect AI written text. However, the reliability of these detectors remains uncertain. In our…

人工智能 · 计算机科学 2025-07-01 Selin Dik , Osman Erdem , Mehmet Dik

It has often been said, correctly, that a monkey forever randomly typing on a keyboard would eventually produce the complete works of William Shakespeare. Almost just as often it has been pointed out that this "eventually" is well beyond…

历史与综述 · 数学 2025-12-17 Ioannis Kontoyiannis

We investigated gender bias in letters of recommendation as a possible cause of the under-representation of women in Experimental Particle Physics (EPP), where about 15% of faculty are female -- well below the 60% level in psychology and…

物理与社会 · 物理学 2022-02-18 R. H. Bernstein , M. W. Macy , C. J. Cameron , S. Williams-Ceci , W. M. Williams , S. J. Ceci

We randomly deploy questions constructed with and without use of the LLM tool and gauge the ability of the students to correctly answer, as well as their ability to correctly perceive the difference between human-authored and LLM-authored…

计算机与社会 · 计算机科学 2025-03-26 Gavin Witsken , Igor Crk , Eren Gultepe

In this note we present the worst-character rule, an efficient variation of the bad-character heuristic for the exact string matching problem, firstly introduced in the well-known Boyer-Moore algorithm. Our proposed rule selects a position…

数据结构与算法 · 计算机科学 2010-12-08 Domenico Cantone , Simone Faro

We present in this paper a numerical investigation of literary texts by various well-known English writers, covering the first half of the twentieth century, based upon the results obtained through corpus analysis of the texts. A fractal…

其他凝聚态物理 · 物理学 2009-11-11 L. L. Goncalves , L. B. Goncalves

Raymond Smullyan came up with a puzzle that George Boolos called The Hardest Logic Puzzle Ever.[1] The puzzle has truthful, lying, and random gods who answer yes or no questions with words that we don't know the meaning of. The challenge is…

综合数学 · 数学 2026-05-06 Daniel Vallstrom

A small group of postdocs, graduate students, and undergraduates inadvertently formed a longitudinal study contrasting expected productivity levels with actual productivity levels. Over the last nine months, our group self-reported 559…

科普物理 · 物理学 2021-04-01 Kaley Brauer

We consider the task of predicting how literary a text is, with a gold standard from human ratings. Aside from a standard bigram baseline, we apply rich syntactic tree fragments, mined from the training set, and a series of hand-picked…

计算与语言 · 计算机科学 2017-04-12 Andreas van Cranenburgh , Rens Bod

Based on data from a large-scale experiment with human subjects, we conclude that the logarithm of probability to guess a word in context (unpredictability) depends linearly on the word length. This result holds both for poetry and prose,…

信息论 · 计算机科学 2007-07-16 Dmitrii Manin

Hidden structural patterns in written texts have been subject of considerable research in the last decades. In particular, mapping a text into a time series of sentence lengths is a natural way to investigate text structure. Typically,…

计算与语言 · 计算机科学 2018-05-07 Denner S. Vieira , Sergio Picoli , Renio S. Mendes

Understanding the complexity of human language requires an appropriate analysis of the statistical distribution of words in texts. We consider the information retrieval problem of detecting and ranking the relevant words of a text by means…

计算与语言 · 计算机科学 2008-06-07 Juan P. Herrera , Pedro A. Pury

Which statistical features distinguish a meaningful text (possibly written in an unknown system) from a meaningless set of symbols? Here we answer this question by comparing features of the first half of a text to its second half. This…

计算与语言 · 计算机科学 2023-07-18 Weibing Deng , R. Xie , S. Deng , Armen E. Allahverdyan

The mathematical distinction between prose and verse may be detected in writings that are not apparently lineated, for example in T. S. Eliot's "Burnt Norton", and Jim Crace's "Quarantine". In this paper we offer comments on appropriate…

计算与语言 · 计算机科学 2007-05-23 John Constable , Hideaki Aoyama

The meteoric rise in text generation capability has been accompanied by parallel growth in interest in machine-generated text detection: the capability to identify whether a given text was generated using a model or written by a person.…

计算与语言 · 计算机科学 2026-04-24 Kevin Stowe , Svetlana Afanaseva , Rodolfo Raimundo , Yitao Sun , Kailash Patil

The $T$-test is probably the most popular statistical test; it is routinely recommended by the textbooks. The applicability of the test relies upon the validity of normal or Student's approximation to the distribution of Student's statistic…

统计理论 · 数学 2021-01-01 S. Y. Novak

Creative writing has long been considered a uniquely human endeavor, requiring voice and style that machines could not replicate. This assumption is challenged by Generative AI that can emulate thousands of author styles in seconds with…

人工智能 · 计算机科学 2026-01-27 Tuhin Chakrabarty , Paramveer S. Dhillon

We correct a common (but mistaken) attribution of the evaluation of the probability integral, usually attributed to Poisson, Gauss, or Laplace.

历史与综述 · 数学 2019-10-22 Fausto Di Biase