English
Related papers

Related papers: Why Did You Not Compare With That? Identifying Pap…

200 papers

Keyphrase boundary classification (KBC) is the task of detecting keyphrases in scientific articles and labelling them with respect to predefined types. Although important in practice, this task is so far underexplored, partly due to the…

Computation and Language · Computer Science 2017-04-27 Isabelle Augenstein , Anders Søgaard

In this thesis, we look at the problem of assigning each identifier of a document to a namespace. At the moment, there does not exist a special dataset where all identifiers are grouped to namespaces, and therefore we need to create such a…

Information Retrieval · Computer Science 2016-01-14 Alexey Grigorev

Natural language processing (NLP) techniques have been widely applied in the requirements engineering (RE) field to support tasks such as classification and ambiguity detection. Although RE research is rooted in empirical investigation, it…

Computation and Language · Computer Science 2025-07-30 Meet Bhatt , Nic Boilard , Muhammad Rehan Chaudhary , Cole Thompson , Jacob Idoko , Aakash Sorathiya , Gouri Ginde

Long-context understanding has emerged as a critical capability for large language models (LLMs). However, evaluating this ability remains challenging. We present SCALAR, a benchmark designed to assess citation-grounded long-context…

Computation and Language · Computer Science 2026-01-23 Renxi Wang , Honglin Mu , Liqun Ma , Lizhi Lin , Yunlong Feng , Timothy Baldwin , Xudong Han , Haonan Li

This paper presents a pipeline to detect and explain anomalous reviews in online platforms. The pipeline is made up of three modules and allows the detection of reviews that do not generate value for users due to either worthless or…

Computation and Language · Computer Science 2024-02-29 David Novoa-Paradela , Oscar Fontenla-Romero , Bertha Guijarro-Berdiñas

We present a very simple, unsupervised method for the pairwise matching of documents from heterogeneous collections. We demonstrate our method with the Concept-Project matching task, which is a binary classification task involving pairs of…

Computation and Language · Computer Science 2019-04-30 Mark-Christoph Müller

We present our submission to the AXOLOTL-24 shared task. The shared task comprises two subtasks: identifying new senses that words gain with time (when comparing newer and older time periods) and producing the definitions for the identified…

Computation and Language · Computer Science 2024-07-08 Aleksei Dorkin , Kairit Sirts

Binary relevance is a simple approach to solve multi-label learning problems where an independent binary classifier is built per each label. A common challenge with this in real-world applications is that the label space can be very large,…

Information Retrieval · Computer Science 2019-05-29 Dora Jambor , Peng Yu

Causality is essential in scientific research, enabling researchers to interpret true relationships between variables. These causal relationships are often represented by causal graphs, which are directed acyclic graphs. With the recent…

Computation and Language · Computer Science 2025-02-19 Ivaxi Sheth , Bahare Fatemi , Mario Fritz

While increasingly complex approaches to question answering (QA) have been proposed, the true gain of these systems, particularly with respect to their expensive training requirements, can be inflated when they are not compared to adequate…

Information Retrieval · Computer Science 2018-07-06 Vikas Yadav , Rebecca Sharp , Mihai Surdeanu

Artificial neural networks are being proposed as models of parts of the brain. The networks are compared to recordings of biological neurons, and good performance in reproducing neural responses is considered to support the model's…

Neurons and Cognition · Quantitative Biology 2023-09-01 Yena Han , Tomaso Poggio , Brian Cheung

Three publicly-available LLM specifically designed for legal tasks have been implemented and shown that classification accuracy can benefit from training over legal corpora, but why and how? Here we use two publicly-available legal…

Machine Learning · Computer Science 2025-01-30 Richard K. Belew

The research process includes many decisions, e.g., how to entitle and where to publish the paper. In this paper, we introduce a general framework for investigating the effects of such decisions. The main difficulty in investigating the…

Digital Libraries · Computer Science 2022-08-23 Ryoma Sato , Makoto Yamada , Hisashi Kashima

Citation information in scholarly data is an important source of insight into the reception of publications and the scholarly discourse. Outcomes of citation analyses and the applicability of citation based machine learning approaches…

Digital Libraries · Computer Science 2022-01-12 Tarek Saier , Michael Färber , Tornike Tsereteli

Understanding how co-authors distribute credit is critical for accurately assessing scholarly collaboration. In this study, we uncover the implicit structures within scientific teamwork by systematically analyzing author contributions…

Digital Libraries · Computer Science 2026-02-26 Itai Assraf , Michael Fire

We present a baseline for the CLPsych 2025 A.1 task: classifying self-states in mental health data taken from Reddit. We use few-shot learning with a 4-bit quantized Gemma 2 9B model and a data preprocessing step which first identifies…

Computation and Language · Computer Science 2025-04-22 Laerdon Kim

In this paper, we show that citation counts work better than a random baseline (by a margin of 10%) in distinguishing excellent research, while Mendeley reader counts don't work better than the baseline. Specifically, we study the potential…

Digital Libraries · Computer Science 2018-02-15 Drahomira Herrmannova , Robert M. Patton , Petr Knoth , Christopher G. Stahl

Clinical information extraction (e.g., 2010 i2b2/VA challenge) usually presents tasks of concept recognition, assertion classification, and relation extraction. Jointly modeling the multi-stage tasks in the clinical domain is an…

Computation and Language · Computer Science 2026-03-10 Fei Cheng , Ribeka Tanaka , Sadao Kurohashi

The main objective of this paper is to empirically test whether the identification of highly-cited documents through Google Scholar is feasible and reliable. To this end, we carried out a longitudinal analysis (1950 to 2013), running a…

Digital Libraries · Computer Science 2018-04-30 Alberto Martín-Martín , Enrique Orduna-Malea , Anne-Wil Harzing , Emilio Delgado López-Cózar

Fact checking is an essential challenge when combating fake news. Identifying documents that agree or disagree with a particular statement (claim) is a core task in this process. In this context, stance detection aims at identifying the…

Computation and Language · Computer Science 2021-05-18 Arjun Roy , Pavlos Fafalios , Asif Ekbal , Xiaofei Zhu , Stefan Dietze
‹ Prev 1 8 9 10 Next ›