Related papers: Extended Comment on Language Trees and Zipping
The mapping of lexical meanings to wordforms is a major feature of natural languages. While usage pressures might assign short words to frequent meanings (Zipf's law of abbreviation), the need for a productive and open-ended vocabulary,…
This paper attempts to classify various blinding strategies used in particle physics. It argues that the blinding technique is not used consistently throughout searches for new physics. More importantly, the blinding technique, in its…
We present a simple new method for proving that languages are not regular. We prove the correctness of the method, illustrate the ease of using the method on well-known examples of nonregular languages, and prove two additional theorems on…
Many challenges in natural language processing require generating text, including language translation, dialogue generation, and speech recognition. For all of these problems, text generation becomes more difficult as the text becomes…
Prompt sensitivity, referring to the phenomenon where paraphrasing (i.e., repeating something written or spoken using different words) leads to significant changes in large language model (LLM) performance, has been widely accepted as a…
Disagreements help drive science. How does one identify and track them in scholarly literature? We ask the research question will searching review articles (RA) will be more time efficient for this purpose than searching non-review ones…
A reproducibility crisis has been reported in science, but the extent to which it affects AI research is not yet fully understood. Therefore, we performed a systematic replication study including 30 highly cited AI studies relying on…
In recent years, criticism of the methodology of particle physics beyond the Standard Model has increased, diagnosing too much reliance on aesthetic criteria for theory development and evaluation. Faced with several decades of experimental…
This is an addendum to the Reply Comment [Phys. Rev. Lett. 102, 139602 (2009), arXiv:0811.0518] to Comment [Phys. Rev. Lett. 102, 139601 (2009), arXiv:0810.4791] on Letter [Phys. Rev. Lett. 100, 116101 (2008), arXiv:0804.1898].
Has the style of scientific communication changed due to the growing use of large language models in the writing process? We address this question in the domain of Natural Language Processing by leveraging two data resources we create: a…
This paper presents a fundamental algorithm for parsing natural language sentences into dependency trees. Unlike phrase-structure (constituency) parsers, this algorithm operates one word at a time, attaching each word as soon as it can be…
The author of the comment~[arXiv:2302.04190] criticizes our published results in Phys. Rev. Lett. \textbf{125}, 064301 (2020) about the Tennis Racket Effect (TRE). The TRE is a geometric effect which occurs in the free rotation of any…
Peer review is essential for scientific progress but faces growing challenges due to increasing submission volumes and reviewer fatigue. Existing automated review approaches struggle with factual accuracy, rating consistency, and analytical…
The machine learning publication process is broken, of that there can be no doubt. Many of these flaws are attributed to the current workflow: LaTeX to PDF to reviewers to camera ready PDF. This has understandably resulted in the desire for…
This paper presents novel prompting techniques to improve the performance of automatic summarization systems for scientific articles. Scientific article summarization is highly challenging due to the length and complexity of these…
This is the 3rd version with major updates and revisions to the 2nd version of the same title. In this article we have collected and corrected some common mistakes made by Chinese students in writing astronomy and physics literature in…
Providing natural language explanations for recommendations is particularly useful from the perspective of a non-expert user. Although several methods for providing such explanations have recently been proposed, we argue that an important…
In this paper, we classify scientific articles in the domain of natural language processing (NLP) and machine learning (ML), as core subfields of artificial intelligence (AI), into whether (i) they extend the current state-of-the-art by the…
Purpose The purpose of this paper is to explore which structures of academic articles referees would pay more attention to, what specific content referees focus on, and whether the distribution of PRC is related to the citations.…
Large language models (LLMs) are increasingly used in academic peer review, yet their reliability, alignment with human judgment, and robustness to adversarial attacks remain poorly understood. We present a systematic benchmark of…