English
Related papers

Related papers: A simplicity bubble problem and zemblanity in digi…

200 papers

While preference-based recommendation algorithms effectively enhance user engagement by recommending personalized content, they often result in the creation of ``filter bubbles''. These bubbles restrict the range of information users…

Human-Computer Interaction · Computer Science 2024-04-09 Mengyan Wang , Yuxuan Hu , Shiqing Wu , Weihua Li , Quan Bai , Verica Rupar

Recommender systems are widely applied in digital platforms such as news websites to personalize services based on user preferences. In news websites most of users are anonymous and the only available data is sequences of items in anonymous…

Information Retrieval · Computer Science 2021-12-20 Alireza Gharahighehi , Celine Vens

Traditionally categorical data analysis (e.g. generalized linear models) works with simple, flat datasets akin to a single table in a database with no notion of missing data or conflicting versions. In contrast, modern data analysis must…

Databases · Computer Science 2017-08-11 Jason Morton

This article explores the concept of technological singularity and the factors that could accelerate or hinder its arrival. The butterfly effect is used as a framework to understand how seemingly small changes in complex systems can have…

Computers and Society · Computer Science 2025-03-11 Hooman Shababi

Model multiplicity refers to the existence of multiple machine learning models that describe the data equally well but may produce different predictions on individual samples. In medicine, these models can admit conflicting predictions for…

As machine learning algorithms are deployed on sensitive data in critical decision making processes, it is becoming increasingly important that they are also private and fair. In this paper, we show that, when the data has a long-tailed…

Machine Learning · Computer Science 2022-12-27 Amartya Sanyal , Yaxi Hu , Fanny Yang

The unprecedented demand for large amount of data has catalyzed the trend of combining human insights with machine learning techniques, which facilitate the use of crowdsourcing to enlist label information both effectively and efficiently.…

Machine Learning · Statistics 2018-06-26 Yao Zhou , Jingrui He

Recently it has been demonstrated that causal entropic forces can lead to the emergence of complex phenomena associated with human cognitive niche such as tool use and social cooperation. Here I show that even more fundamental traits…

Information Theory · Computer Science 2016-04-20 Fouad Khan

There is a growing interest in societal concerns in machine learning systems, especially in fairness. Multicalibration gives a comprehensive methodology to address group fairness. In this work, we address the multicalibration error and…

Machine Learning · Computer Science 2021-06-08 Eliran Shabat , Lee Cohen , Yishay Mansour

Adaptive dynamical systems arise in a multitude of contexts, e.g., optimization, control, communications, signal processing, and machine learning. A precise characterization of their fundamental limitations is therefore of paramount…

Information Theory · Computer Science 2010-10-13 Maxim Raginsky

Despite extensive research spanning several decades, class imbalance is still considered a profound difficulty for both machine learning and deep learning models. While data oversampling is the foremost technique to address this issue,…

Machine Learning · Computer Science 2025-02-12 Sukumar Kishanthan , Asela Hevapathige

Collaborative filtering based algorithms, including Recurrent Neural Networks (RNN), tend towards predicting a perpetuation of past observed behavior. In a recommendation context, this can lead to an overly narrow set of suggestions lacking…

Information Retrieval · Computer Science 2019-07-04 Zachary A. Pardos , Weijie Jiang

In this paper the theory of semi-bounded rationality is proposed as an extension of the theory of bounded rationality. In particular, it is proposed that a decision making process involves two components and these are the correlation…

Artificial Intelligence · Computer Science 2013-05-28 Tshilidzi Marwala

Bottleneck problems are an important class of optimization problems that have recently gained increasing attention in the domain of machine learning and information theory. They are widely used in generative models, fair machine learning…

Machine Learning · Computer Science 2022-07-12 Behrooz Razeghi , Flavio P. Calmon , Deniz Gunduz , Slava Voloshynovskiy

This paper presents an interdisciplinary approach to analyze the emergence and impact of filter bubbles in social phenomena, especially in both digital and offline environments, by applying the concepts of quantum field theory. Filter…

Physics and Society · Physics 2024-04-30 Yasuko Kawahata

Distributional data tells us that a man can swallow candy, but not that a man can swallow a paintball, since this is never attested. However both are physically plausible events. This paper introduces the task of semantic plausibility:…

Computation and Language · Computer Science 2018-04-11 Su Wang , Greg Durrett , Katrin Erk

Traditional machine learning relies on explicit models and domain assumptions, limiting flexibility and interpretability. We introduce a model-free framework using surprisal (information theoretic uncertainty) to directly analyze and…

The flourishing of fake news is favored by recommendation algorithms of online social networks which, based on previous users activity, provide content adapted to their preferences and so create filter bubbles. We introduce an analytically…

Physics and Society · Physics 2020-10-28 Giordano De Marzo , Andrea Zaccaria , Claudio Castellano

Many learning machines such as normal mixtures and layered neural networks are not regular but singular statistical models, because the map from a parameter to a probability distribution is not one-to-one. The conventional statistical…

Statistics Theory · Mathematics 2015-06-03 Koshi Yamada , Sumio Watanabe

Machine learning has gained widespread attention as a powerful tool to identify structure in complex, high-dimensional data. However, these techniques are ostensibly inapplicable for experimental systems where data is scarce or expensive to…

‹ Prev 1 4 5 6 7 8 10 Next ›