Related papers: Not satisfied about the analysis
In the recently posted arXiv:2312.04495v3 [1], its authors used a computer code written by me supplied to them by the authors of Ref. [2], and claimed that the results that they obtained invalidate the results and conclusions presented in…
We present several results, including some remarks on the Hopf Lemma.
After the rejection of their comment [arXiv:cond-mat/0609399v1] to our Phys. Rev. Lett. {\bf 97}, 100601 (2006), the Authors informed us that an extended version of their comment is going to be published in a different journal under the…
This article is a response to recent proposals by Pearl and others for a new approach to personalised treatment decisions, in contrast to the traditional one based on statistical decision theory. We argue that this approach is dangerously…
This work presents our efforts to reproduce the results of the human evaluation experiment presented in the paper of Vamvas and Sennrich (2022), which evaluated an automatic system detecting over- and undertranslations (translations…
The paper was withdrawn by the author. It contained various errors.
This paper has been withdrawn by the authors due to a gap in the proof of the main result (in 5.3).
Understanding how humans revise their beliefs in light of new information is crucial for developing AI systems which can effectively model, and thus align with, human reasoning. While theoretical belief revision frameworks rely on a set of…
Large language models offer a tempting solution to address the peer review crisis. This position paper argues that today's AI systems should not be used to produce paper reviews. We ground this position in an empirical comparison of human-…
In this paper, we focus on online reviews and employ artificial intelligence tools, taken from the cognitive computing field, to help understanding the relationships between the textual part of the review and the assigned numerical score.…
Text Style Transfer (TST) evaluation is, in practice, inconsistent. Therefore, we conduct a meta-analysis on human and automated TST evaluation and experimentation that thoroughly examines existing literature in the field. The meta-analysis…
Unfortunately, the article "A Comparative Study to Benchmark Cross-project Defect Prediction Approaches" has a problem in the statistical analysis which was pointed out almost immediately after the pre-print of the article appeared online.…
Recent studies comparing AI-generated and human-authored literary texts have produced conflicting results: some suggest AI already surpasses human quality, while others argue it still falls short. We start from the hypothesis that such…
This is an update of my problem list.
In [2] the author claims to provide a counterexample to a result in a recent paper [1]. In this note, we prove that the details of his example is false and this example is compatible with our result in [1] and so is not a countreexample.
This paper has been withdrawn by the authors.
Since state-of-the-art approaches to offensive language detection rely on supervised learning, it is crucial to quickly adapt them to the continuously evolving scenario of social media. While several approaches have been proposed to tackle…
This paper has been withdrawn by the author due to a crucial error in the submission action.
Finding claims that researchers have made considerable progress in artificial intelligence over the last several decades is easy. However, our everyday interactions with cognitive systems (e.g., Siri, Alexa, DALL-E) quickly move from…
arXiv admin note: This version removed by arXiv administrators as the submitter did not have the right to agree to the license at the time of submission