English
Related papers

Related papers: Deep Author Name Disambiguation using DBLP Data

200 papers

We introduce a novel method for converting text data into abstract image representations, which allows image-based processing techniques (e.g. image classification networks) to be applied to text-based comparison problems. We apply the…

Computation and Language · Computer Science 2020-02-07 Stephen M. Petrie , T'Mir D. Julius

Deep Research Agents increasingly automate survey generation, yet whether they match human experts at retrieving essential papers and organizing them into expert-like taxonomies remains unclear. Existing benchmarks emphasize writing quality…

Author similarity and detection is an integral first step in detecting state-led disinformation campaigns in an automated fashion. Current detection techniques require an analyst or subject matter expert to hand-curate accounts. Stylometric…

Social and Information Networks · Computer Science 2019-12-10 A. Kingsland , D. Fortin , E. Cary , S. Smith , K. Pazdernik , R. Perko

The issue of gender bias in scientific publications has been a topic of ongoing debate. One aspect of this debate concerns whether women receive equal credit for their contributions compared to men. Conventional wisdom suggests that women…

Digital Libraries · Computer Science 2025-08-05 Keigo Kusumegi , Daniel E. Acuña , Yukie Sano

Authorship attribution refers to the task of automatically determining the author based on a given sample of text. It is a problem with a long history and has a wide range of application. Building author profiles using language models is…

Computation and Language · Computer Science 2016-02-25 Zhenhao Ge , Yufang Sun

Following the work of Krumov et al. [Eur. Phys. J. B 84, 535 (2011)] we revisit the question whether the usage of large citation datasets allows for the quantitative assessment of social (by means of coauthorship of publications) influence…

Physics and Society · Physics 2014-10-03 David F. Klosik , Stefan Bornholdt , Marc-Thorsten Hütt

Globalization and the world wide web has resulted in academia and science being an international and multicultural community forged by researchers and scientists with different ethnicities. How ethnicity shapes the evolution of membership,…

Digital Libraries · Computer Science 2014-11-06 Zhaohui Wu , Dayu Yuan , Pucktada Treeratpituk , C. Lee Giles

Term suggestion or recommendation modules can help users to formulate their queries by mapping their personal vocabularies onto the specialized vocabulary of a digital library. While we examined actual user queries of the social sciences…

Information Retrieval · Computer Science 2013-12-02 Philipp Schaer , Philipp Mayr , Thomas Lüke

Domain-specific named entity recognition (NER) on Computer Science (CS) scholarly articles is an information extraction task that is arguably more challenging for the various annotation aims that can beset the task and has been less studied…

Computation and Language · Computer Science 2022-11-15 Jennifer D'Souza , Sören Auer

This paper introduces a suite of approaches and measures to study the impact of co-authorship teams based on the number of publications and their citations on a local and global scale. In particular, we present a novel weighted graph…

Other Condensed Matter · Physics 2007-05-23 Katy Börner , Luca Dall'Asta , Weimao Ke , Alessandro Vespignani

Authorship has entangled style and content inside. Authors frequently write about the same topics in the same style, so when different authors write about the exact same topic the easiest way out to distinguish them is by understanding the…

Computation and Language · Computer Science 2024-11-28 Javier Huertas-Tato , Adrián Girón-Jiménez , Alejandro Martín , David Camacho

Automatic processing of bibliographic data becomes very important in digital libraries, data science and machine learning due to its importance in keeping pace with the significant increase of published papers every year from one side and…

Digital Libraries · Computer Science 2021-06-24 Zeyd Boukhers , Philipp Mayr , Silvio Peroni

AA is the process of attributing an unidentified document to its true author from a predefined group of known candidates, each possessing multiple samples. The nature of AA necessitates accommodating emerging new authors, as each individual…

Information Retrieval · Computer Science 2024-08-20 Mostafa Rahgouy , Hamed Babaei Giglou , Mehnaz Tabassum , Dongji Feng , Amit Das , Taher Rahgooy , Gerry Dozier , Cheryl D. Seals

Different entities with the same name can be difficult to distinguish. Handling confusing entity mentions is a crucial skill for language models (LMs). For example, given the question "Where was Michael Jordan educated?" and a set of…

Computation and Language · Computer Science 2024-08-12 Yoonsang Lee , Xi Ye , Eunsol Choi

Authorship analysis plays an important role in diverse domains, including forensic linguistics, academia, cybersecurity, and digital content authentication. This paper presents a systematic literature review on two key sub-tasks of…

Computation and Language · Computer Science 2025-05-22 Nudrat Habib , Tosin Adewumi , Marcus Liwicki , Elisa Barney

Entity Disambiguation aims to link mentions of ambiguous entities to a knowledge base (e.g., Wikipedia). Modeling topical coherence is crucial for this task based on the assumption that information from the same semantic context tends to…

Computation and Language · Computer Science 2015-04-30 Hongzhao Huang , Larry Heck , Heng Ji

We introduce ParaNames, a multilingual parallel name resource consisting of 118 million names spanning across 400 languages. Names are provided for 13.6 million entities which are mapped to standardized entity types (PER/LOC/ORG). Using…

Computation and Language · Computer Science 2022-07-13 Jonne Sälevä , Constantine Lignos

Predicting the emergence of future research collaborations between authors in academic social networks (SNs) is a very effective example that demonstrates the link prediction problem. This problem refers to predicting the potential…

Social and Information Networks · Computer Science 2025-09-23 Doaa Hassan , Mohammad Al Hasan

The rapid advancement of large language models (LLMs) has enabled powerful authorship inference capabilities, raising growing concerns about unintended deanonymization risks in textual data such as news articles. In this work, we introduce…

Computation and Language · Computer Science 2026-02-27 Boyang Zhang , Yang Zhang

Modern performance on several natural language processing (NLP) tasks has been enhanced thanks to the Transformer-based pre-trained language model BERT. We employ this concept to investigate a local publication database. Research papers are…

Computation and Language · Computer Science 2023-06-16 Zineddine Bettouche , Andreas Fischer