English
Related papers

Related papers: PR2: A Language Independent Unsupervised Tool for …

200 papers

Serendipity-oriented recommender systems aim to counteract over-specialization in user preferences. However, evaluating a user's serendipitous response towards a recommended item can be challenging because of its emotional nature. In this…

Information Retrieval · Computer Science 2024-12-18 Yu Tokutake , Kazushi Okamoto

Personality recognition from text is typically cast as hard-label classification, which obscures the graded, prototype-like nature of human personality judgments. We present ProtoMBTI, a cognitively aligned framework for MBTI inference that…

Computation and Language · Computer Science 2025-12-30 Haoyuan Li , Yuanbo Tong , Yuchen Li , Zirui Wang , Chunhou Liu , Jiamou Liu

Large language models have advanced web agents, yet current agents lack personalization capabilities. Since users rarely specify every detail of their intent, practical web agents must be able to interpret ambiguous queries by inferring…

Computation and Language · Computer Science 2026-05-28 Serin Kim , Sangam Lee , Dongha Lee

A growing body of research examines personality traits in Large Language Models (LLMs), particularly in human-agent collaboration. Prior work has frequently applied the Big Five inventory to assess LLM behavior analogous to human…

Human-Computer Interaction · Computer Science 2026-03-20 Kim Zierahn , Cristina Cachero , Anna Korhonen , Nuria Oliver

Large language models (LLMs) make it possible to generate synthetic behavioural data at scale, offering an ethical and low-cost alternative to human experiments. Whether such data can faithfully capture psychological differences driven by…

Computation and Language · Computer Science 2025-11-27 Manuel Pratelli , Marinella Petrocchi

Through prompting, large-scale pre-trained models have become more expressive and powerful, gaining significant attention in recent years. Though these big models have zero-shot capabilities, in general, labeled data are still required to…

Machine Learning · Computer Science 2023-05-02 Korawat Tanwisuth , Shujian Zhang , Huangjie Zheng , Pengcheng He , Mingyuan Zhou

Self-supervised learning offers an efficient way of extracting rich representations from various types of unlabeled data while avoiding the cost of annotating large-scale datasets. This is achievable by designing a pretext task to form…

Machine Learning · Computer Science 2023-10-11 Pouya Mehralian , Bagher BabaAli , Ashena Gorgan Mohammadi

We introduce Stanza, an open-source Python natural language processing toolkit supporting 66 human languages. Compared to existing widely used toolkits, Stanza features a language-agnostic fully neural pipeline for text analysis, including…

Computation and Language · Computer Science 2020-04-24 Peng Qi , Yuhao Zhang , Yuhui Zhang , Jason Bolton , Christopher D. Manning

Personality computing and affective computing, where the recognition of personality traits is essential, have gained increasing interest and attention in many research areas recently. We propose a novel approach to recognize the Big Five…

Computer Vision and Pattern Recognition · Computer Science 2019-11-04 Süleyman Aslan , Uğur Güdükbay

AI-powered recruitment tools are increasingly adopted in personnel selection, yet they struggle to capture the requisition (req)-specific personal competencies (PCs) that distinguish successful candidates beyond job categories. We propose a…

Computation and Language · Computer Science 2026-04-02 Wanxin Li , Denver McNeney , Nivedita Prabhu , Charlene Zhang , Renee Barr , Matthew Kitching , Khanh Dao Duc , Anthony S. Boyce

As language models achieve increasingly human-like capabilities in conversational text generation, a critical question emerges: to what extent can these systems simulate the characteristics of specific individuals? To evaluate this, we…

Computation and Language · Computer Science 2025-06-04 Quan Shi , Carlos E. Jimenez , Stephen Dong , Brian Seo , Caden Yao , Adam Kelch , Karthik Narasimhan

Natural language processing techniques are increasingly applied to identify social trends and predict behavior based on large text collections. Existing methods typically rely on surface lexical and syntactic information. Yet, research in…

Computation and Language · Computer Science 2016-09-29 Ekaterina Shutova , Patricia Lichtenstein

Human annotation for syntactic parsing is expensive, and large resources are available only for a fraction of languages. A question we ask is whether one can leverage abundant unlabeled texts to improve syntactic parsers, beyond just using…

Computation and Language · Computer Science 2019-02-22 Caio Corro , Ivan Titov

Questionnaires are a common method for detecting the personality of Large Language Models (LLMs). However, their reliability is often compromised by two main issues: hallucinations (where LLMs produce inaccurate or irrelevant responses) and…

Computation and Language · Computer Science 2024-10-14 Baohua Zhan , Yongyi Huang , Wenyao Cui , Huaping Zhang , Jianyun Shang

Media houses reporting on public figures, often come with their own biases stemming from their respective worldviews. A characterization of these underlying patterns helps us in better understanding and interpreting news stories. For this,…

Computation and Language · Computer Science 2023-09-13 Sharath Srivatsa , Srinath Srinivasa

Trained on various human-authored corpora, Large Language Models (LLMs) have demonstrated a certain capability of reflecting specific human-like traits (e.g., personality or values) by prompting, benefiting applications like personalized…

Computation and Language · Computer Science 2025-12-01 Yuzhuo Bai , Shitong Duan , Muhua Huang , Jing Yao , Zhenghao Liu , Peng Zhang , Tun Lu , Xiaoyuan Yi , Maosong Sun , Xing Xie

Recent privacy research on large language models (LLMs) has shown that they achieve near-human-level performance at inferring personal data from online texts. With ever-increasing model capabilities, existing text anonymization methods are…

Artificial Intelligence · Computer Science 2025-02-04 Robin Staab , Mark Vero , Mislav Balunović , Martin Vechev

Significant work has been done on learning regular expressions from a set of data values. Depending on the domain, this approach can be very successful. However, significant time is required to learn these expressions and the resulting…

Databases · Computer Science 2024-03-18 Michael J. Mior

Efficiently selecting relevant content from vast candidate pools is a critical challenge in modern recommender systems. Traditional methods, such as item-to-item collaborative filtering (CF) and two-tower models, often fall short in…

Information Retrieval · Computer Science 2026-01-26 Shaoqing Wang , Yingcai Ma , Kairui Fu , Ziyang Wang , Dunxian Huang , Yuliang Yan , Jian Wu

Imbuing Large Language Models (LLMs) with specific personas is prevalent for tailoring interaction styles, yet the impact on underlying cognitive capabilities remains unexplored. We employ the Neuron-based Personality Trait Induction (NPTI)…

Computation and Language · Computer Science 2026-05-13 Jiaqi Chen , Ming Wang , Tingna Xie , Shi Feng , Yongkang Liu