English
Related papers

Related papers: Inferring individual attributes from search engine…

200 papers

Searching health information on web has become an integral part of today's world, and many people turn to the Web for healthcare information and healthcare assessment. Our pilot study investigates users' preferences for the type of search…

Information Retrieval · Computer Science 2014-10-30 Shanu Sushmita , Si-Chi Chin

Inference of online social network users' attributes and interests has been an active research topic. Accurate identification of users' attributes and interests is crucial for improving the performance of personalization and recommender…

Social and Information Networks · Computer Science 2015-04-21 Quanzeng You , Sumit Bhatia , Jiebo Luo

Personalization is being applied to great extend in many systems. This paper presents a multi-dimensional user data model and its application in web search. Online and Offline activities of the user are tracked for creating the user model.…

Information Retrieval · Computer Science 2013-06-20 Nithin K. Anil , Sharath Basil Kurian , Aby Abahai T , Surekha Mariam Varghese

Attribute inference - the process of analyzing publicly available data in order to uncover hidden information - has become a major threat to privacy, given the recent technological leap in machine learning. One way to tackle this threat is…

Artificial Intelligence · Computer Science 2023-04-25 Marcin Waniek , Navya Suri , Abdullah Zameek , Bedoor AlShebli , Talal Rahwan

In this work we analyze the problem of, given the probability distribution of a population, questioning an unknown individual that is representative of the distribution so that our uncertainty about certain characteristics is significantly…

Computational Complexity · Computer Science 2026-01-22 David Pantoja , Ismael Rodriguez , Fernando Rubio , Clara Segura

In many data exploration tasks it is meaningful to identify groups of attribute interactions that are specific to a variable of interest. For instance, in a dataset where the attributes are medical markers and the variable of interest…

Machine Learning · Statistics 2017-03-17 Andreas Henelius , Antti Ukkonen , Kai Puolamäki

Given only data generated by a standard confounding graph with unobserved confounder, the Average Treatment Effect (ATE) is not identifiable. To estimate the ATE, a practitioner must then either (a) collect deconfounded data;(b) run a…

Machine Learning · Statistics 2021-03-09 Kyra Gan , Andrew A. Li , Zachary C. Lipton , Sridhar Tayur

The Internet of Things (IoT) promises to improve user utility by tuning applications to user behavior, but revealing the characteristics of a user's behavior presents a significant privacy risk. Our previous work has established the…

Cryptography and Security · Computer Science 2020-07-14 Nazanin Takbiri , Minting Chen , Dennis L. Goeckel , Amir Houmansadr , Hossein Pishro-Nik

Computational social scientists often harness the Web as a "societal observatory" where data about human social behavior is collected. This data enables novel investigations of psychological, anthropological and sociological research…

Computers and Society · Computer Science 2016-03-15 Fariba Karimi , Claudia Wagner , Florian Lemmerich , Mohsen Jadidi , Markus Strohmaier

Objective: To enable privacy-preserving learning of high quality generative and discriminative machine learning models from distributed electronic health records. Methods and Results: We describe general and scalable strategy to build…

Cryptography and Security · Computer Science 2018-06-19 Marina Blanton , Ah Reum Kang , Subhadeep Karan , Jaroslaw Zola

The increased popularity and ubiquitous availability of online social networks and globalised Internet access have affected the way in which people share content. The information that users willingly disclose on these platforms can be used…

Social and Information Networks · Computer Science 2016-07-12 Maria Han Veiga , Carsten Eickhoff

Major search engines deploy personalized Web results to enhance users' experience, by showing them data supposed to be relevant to their interests. Even if this process may bring benefits to users while browsing, it also raises concerns on…

Information Retrieval · Computer Science 2015-08-18 Van Tien Hoang , Angelo Spognardi , Francesco Tiezzi , Marinella Petrocchi , Rocco De Nicola

A new statistical based model approach to characterize a user's behavior in an Internet access link is presented. The real patterns of Internet traffic in a heterogeneous Campus Network are studied. We find three clearly different patterns…

Adaptation and Self-Organizing Systems · Physics 2007-05-23 Carmen Pellicer-Lostao , Daniel Morato , Ricardo Lopez-Ruiz

Suppose you find the same username on different online services, what is the probability that these usernames refer to the same physical person? This work addresses what appears to be a fairly simple question, which has many implications…

Cryptography and Security · Computer Science 2015-03-18 Daniele Perito , Claude Castelluccia , Mohamed Ali Kaafar , Pere Manils

Assessing the diversity of a dataset of information associated with people is crucial before using such data for downstream applications. For a given dataset, this often involves computing the imbalance or disparity in the empirical…

Computers and Society · Computer Science 2021-07-16 Vijay Keswani , L. Elisa Celis

Understanding how people interact with the web is key for a variety of applications, e.g., from the design of effective web pages to the definition of successful online marketing campaigns. Browsing behavior has been traditionally…

Computers and Society · Computer Science 2021-05-05 Luca Vassio , Idilio Drago , Marco Mellia , Zied Ben Houidi , Mohamed Lamine Lamali

Advances in algorithmic fairness have largely omitted sexual orientation and gender identity. We explore queer concerns in privacy, censorship, language, online safety, health, and employment to study the positive and negative effects of…

Computers and Society · Computer Science 2021-04-29 Nenad Tomasev , Kevin R. McKee , Jackie Kay , Shakir Mohamed

We propose a new methodology for selecting and ranking covariates associated with a variable of interest in a context of high-dimensional data under dependence but few observations. The methodology successively intertwines the clustering of…

People increasingly turn to the Internet when they have a medical condition. The data they create during this process is a valuable source for medical research and for future health services. However, utilizing these data could come at a…

Computer Science and Game Theory · Computer Science 2020-03-24 Gilie Gefen , Omer Ben-Porat , Moshe Tennenholtz , Elad Yom-Tov

Online users generate tremendous amounts of data. To better serve users, it is required to share the user-related data among researchers, advertisers and application developers. Publishing such data would raise more concerns on user…

Cryptography and Security · Computer Science 2018-06-27 Ghazaleh Beigi