中文
相关论文

相关论文: Narrative Fingerprints: Multi-Scale Author Identif…

200 篇论文

A new method for visualizing the relatedness of scientific areas is developed that is based on measuring the overlap of researchers between areas. It is found that closely related areas have a high propensity to share a larger number of…

数字图书馆 · 计算机科学 2013-03-11 F. G. Serpa , Adam M. Graves , Artjay Javier

The use of statistical methods to analyze large databases of text has been useful to unveil patterns of human behavior and establish historical links between cultures and languages. In this study, we identify literary movements by treating…

物理与社会 · 物理学 2013-02-19 Diego R. Amancio , Osvaldo N. Oliveira , Luciano da F. Costa

Scientific novelty drives advances at the research frontier, yet it is also associated with heightened uncertainty and potential resistance from incumbent paradigms, leading to complex patterns of scientific impact. Prior studies have…

数字图书馆 · 计算机科学 2026-04-15 Yi Zhao , Yang Chenggang , Yuzhuo Wang , Tong Bao , Zhang Heng , Chengzhi Zhang

We provide a general framework to model the growth of networks consisting of different coupled layers. Our aim is to estimate the impact of one such layer on the dynamics of the others. As an application, we study a scientometric network,…

物理与社会 · 物理学 2020-09-16 Vahan Nanumyan , Christoph Gote , Frank Schweitzer

Books are typically segmented into chapters and sections, representing coherent subnarratives and topics. We investigate the task of predicting chapter boundaries, as a proxy for the general task of segmenting long texts. We build a Project…

计算与语言 · 计算机科学 2020-11-10 Charuta Pethe , Allen Kim , Steven Skiena

The recognition of individual contributions is central to the scientific reward system, yet coauthored papers often obscure who did what. Traditional proxies like author order assume a simplistic decline in contribution, while emerging…

社会与信息网络 · 计算机科学 2025-04-22 Lulin Yang , Jiaxin Pei , Lingfei Wu

The expectation that scientific productivity follows regular patterns over a career underpins many scholarly evaluations. However, recent studies of individual productivity patterns reveal a puzzle: the average number of papers published…

应用统计 · 统计学 2026-02-13 Sam Zhang , Nicholas LaBerge , Samuel F. Way , Daniel B. Larremore , Aaron Clauset

Recent advances in text-to-speech (TTS) have been driven by large, multi-domain speech corpora, yet the expressive potential of audiobook data remains underexamined. We argue that human-narrated audiobooks, particularly fictional works,…

音频与语音处理 · 电气工程与系统科学 2026-04-22 Gaspard Michel , Elena V. Epure , Christophe Cerisara

This paper investigates the relationship between scientists' cognitive profile and their ability to generate innovative ideas and gain scientific recognition. We propose a novel author-level metric based on the semantic representation of…

综合经济学 · 经济学 2023-12-19 Pierre Pelletier , Kevin Wirtz

The rapid adoption of generative AI tools is reshaping how scholars produce and communicate knowledge, raising questions about who benefits and who is left behind. We analyze over 230,000 Scopus-indexed computer science articles between…

计算机与社会 · 计算机科学 2025-09-11 Farhan Kamrul Khan , Hazem Ibrahim , Nouar Aldahoul , Talal Rahwan , Yasir Zaki

The proliferation of AI-generated text has intensified the need for reliable authorship verification, yet current output-based methods are increasingly unreliable. We observe that the ordinary typing interface captures rich cognitive…

密码学与安全 · 计算机科学 2026-05-26 David Condrey

The risk of misusing text-to-image generative models for malicious uses, especially due to the open-source development of such models, has become a serious concern. As a risk mitigation strategy, attributing generative models with neural…

计算机视觉与模式识别 · 计算机科学 2025-07-25 Murthy L , Subarna Tripathi

Amidst the rising capabilities of generative AI to mimic specific human styles, this study investigates the ability of state-of-the-art large language models (LLMs), including GPT-4o, Gemini 1.5 Pro, and Claude Sonnet 3.5, to emulate the…

计算与语言 · 计算机科学 2026-04-14 Nasser A Alsadhan

Recent studies comparing AI-generated and human-authored literary texts have produced conflicting results: some suggest AI already surpasses human quality, while others argue it still falls short. We start from the hypothesis that such…

计算与语言 · 计算机科学 2025-06-05 Guillermo Marco , Julio Gonzalo , Víctor Fresno

Authorship identification tasks, which rely heavily on linguistic styles, have always been an important part of Natural Language Understanding (NLU) research. While other tasks based on linguistic style understanding benefit from deep…

计算与语言 · 计算机科学 2020-10-01 Weicheng Ma , Ruibo Liu , Lili Wang , Soroush Vosoughi

Books have historically been the primary mechanism through which narratives are transmitted. We have developed a collection of resources for the large-scale analysis of novels, including: (1) an open source end-to-end NLP analysis pipeline…

计算与语言 · 计算机科学 2023-11-08 Charuta Pethe , Allen Kim , Rajesh Prabhakar , Tanzir Pial , Steven Skiena

Language models have demonstrated remarkable capabilities on standard benchmarks, yet they struggle increasingly from mode collapse, the inability to generate diverse and novel outputs. Our work introduces NoveltyBench, a benchmark…

计算与语言 · 计算机科学 2025-08-12 Yiming Zhang , Harshita Diddee , Susan Holm , Hanchen Liu , Xinyue Liu , Vinay Samuel , Barry Wang , Daphne Ippolito

Concepts and methods of complex networks can be used to analyse texts at their different complexity levels. Examples of natural language processing (NLP) tasks studied via topological analysis of networks are keyword identification,…

计算与语言 · 计算机科学 2017-02-07 Vanessa Queiroz Marinho , Graeme Hirst , Diego Raphael Amancio

Recent studies have evaluated the creativity/novelty of large language models (LLMs) primarily from a semantic perspective, using benchmarks from cognitive science. However, accessing the novelty in scholarly publications is a largely…

计算与语言 · 计算机科学 2024-09-26 Ethan Lin , Zhiyuan Peng , Yi Fang

Textual domain is a crucial property within the Natural Language Processing (NLP) community due to its effects on downstream model performance. The concept itself is, however, loosely defined and, in practice, refers to any non-typological…