English
Related papers

Related papers: Effect of Monetary Reward on Users' Individual Str…

200 papers

Reinforcement learning (RL) can align language models with non-differentiable reward signals, such as human preferences. However, a major challenge arises from the sparsity of these reward signals - typically, there is only a single reward…

Computation and Language · Computer Science 2024-02-20 Meng Cao , Lei Shu , Lei Yu , Yun Zhu , Nevan Wichers , Yinxiao Liu , Lei Meng

Generative Artificial Intelligence is reshaping online communication by enabling large-scale production of Machine-Generated Text (MGT) at low cost. While its presence is rapidly growing across the Web, little is known about how MGT…

Social and Information Networks · Computer Science 2025-10-09 Lucio La Cava , Luca Maria Aiello , Andrea Tagarelli

Reward models (RMs) play a crucial role in aligning large language models (LLMs) with human preferences and enhancing reasoning quality. Traditionally, RMs are trained to rank candidate outputs based on their correctness and coherence.…

Machine Learning · Computer Science 2025-02-21 Yuhui Xu , Hanze Dong , Lei Wang , Caiming Xiong , Junnan Li

We study the effect of persistence of engagement on learning in a stochastic multi-armed bandit setting. In advertising and recommendation systems, repetition effect includes a wear-in period, where the user's propensity to reward the…

Machine Learning · Computer Science 2020-06-19 Priyank Agrawal , Theja Tulabandhula

Generative Artificial Intelligence (AI) tools are increasingly deployed across social media platforms, yet their implications for user behavior and experience remain understudied, particularly regarding two critical dimensions: (1) how AI…

Human-Computer Interaction · Computer Science 2025-06-18 Anders Giovanni Møller , Daniel M. Romero , David Jurgens , Luca Maria Aiello

Popularity systems, like Twitter retweets, Reddit upvotes, and Pinterest pins have the potential to guide people toward posts that others liked. That, however, creates a feedback loop that reduces their informativeness: items marked as more…

Human-Computer Interaction · Computer Science 2018-09-05 Maria Glenski , Greg Stoddard , Paul Resnick , Tim Weninger

We study the impact of content moderation policies in online communities. In our theoretical model, a platform chooses a content moderation policy and individuals choose whether or not to participate in the community according to the…

Data Structures and Algorithms · Computer Science 2023-10-17 Cynthia Dwork , Chris Hays , Jon Kleinberg , Manish Raghavan

We present a strategic analysis of a trust model that has recently been proposed for promoting cooperative behaviour in user-centric networks. The mechanism for cooperation is based on a combination of reputation and virtual currency…

Computer Science and Game Theory · Computer Science 2013-03-05 Marta Kwiatkowska , David Parker , Aistis Simaitis

Alignment of Large Language Models (LLMs) aims to align outputs with human preferences, and personalized alignment further adapts models to individual users. This relies on personalized reward models that capture user-specific preferences…

Computation and Language · Computer Science 2026-04-21 Hongru Cai , Yongqi Li , Tiezheng Yu , Fengbin Zhu , Wenjie Wang , Fuli Feng , Wenjie Li

Content creators compete for user attention. Their reach crucially depends on algorithmic choices made by developers on online platforms. To maximize exposure, many creators adapt strategically, as evidenced by examples like the sprawling…

Computer Science and Game Theory · Computer Science 2023-07-07 Jiri Hron , Karl Krauth , Michael I. Jordan , Niki Kilbertus , Sarah Dean

Reward function is essential in reinforcement learning (RL), serving as the guiding signal to incentivize agents to solve given tasks, however, is also notoriously difficult to design. In many cases, only imperfect rewards are available,…

Machine Learning · Computer Science 2023-02-06 Jianxiong Li , Xiao Hu , Haoran Xu , Jingjing Liu , Xianyuan Zhan , Qing-Shan Jia , Ya-Qin Zhang

People employ efficient planning strategies. But how are these strategies acquired? Previous research suggests that people can discover new planning strategies through learning from reinforcements, a process known as metacognitive…

Artificial Intelligence · Computer Science 2025-05-30 Ruiqi He , Falk Lieder

Positive feedback via likes and awards is central to online governance, yet which attributes of users' posts elicit rewards -- and how these vary across authors and communities -- remains unclear. To examine this, we combine…

Human-Computer Interaction · Computer Science 2026-02-03 Agam Goyal , Charlotte Lambert , Eshwar Chandrasekharan

Peer recommendation is a crowdsourcing task that leverages the opinions of many to identify interesting content online, such as news, images, or videos. Peer recommendation applications often use social signals, e.g., the number of prior…

Physics and Society · Physics 2016-01-28 Tad Hogg , Kristina Lerman

Content creators compete for exposure on recommendation platforms, and such strategic behavior leads to a dynamic shift over the content distribution. However, how the creators' competition impacts user welfare and how the relevance-driven…

Computer Science and Game Theory · Computer Science 2023-05-04 Fan Yao , Chuanhao Li , Denis Nekipelov , Hongning Wang , Haifeng Xu

We model social media as collections of users producing and consuming content. Users value consuming content, but doing so uses up their scarce attention, and hence they prefer content produced by more able users. Users also value receiving…

General Economics · Economics 2021-04-05 Apostolos Filippas , John Horton

The rapid development of the Internet has profoundly changed human life. Humans are increasingly expressing themselves and interacting with others on social media platforms. However, although artificial intelligence technology has been…

Computation and Language · Computer Science 2024-07-11 Haochen Xue , Chong Zhang , Chengzhi Liu , Fangyu Wu , Xiaobo Jin

We design profit-maximizing mechanisms to sell an excludable and non-rival good with positive and/or negative network effects. Buyers have heterogeneous private values that depend on how many others also consume the good. In optimum, an…

Theoretical Economics · Economics 2025-07-31 Vincent Meisner , Pascal Pillath

The extraordinary capabilities of large language models (LLMs) such as ChatGPT and GPT-4 are in part unleashed by aligning them with reward models that are trained on human preferences, which are often represented as rankings of responses…

Machine Learning · Computer Science 2025-10-31 Ziang Song , Tianle Cai , Jason D. Lee , Weijie J. Su

In prior methods, it was observed that the application of Convolutional Neural Networks agent in Deep Reinforcement Learning to financial data resulted in an enhanced reward. In this study, a specific permutation was applied to the feature…

Computational Finance · Quantitative Finance 2024-02-07 Sina Montazeri , Akram Mirzaeinia , Amir Mirzaeinia