中文
相关论文

相关论文: Seed-Driven Geo-Social Data Extraction -- Full Ver…

200 篇论文

Optimization tasks over relational data, such as clustering, often suffer from the prohibitive cost of join operations, which are necessary to access the full dataset. While geometric data structures like BBD trees yield fast approximation…

数据库 · 计算机科学 2026-03-13 Aryan Esmailpour , Stavros Sintos

The increase and rapid growth of data produced by scientific instruments, the Internet of Things (IoT), and social media is causing data transfer performance and resource consumption to garner much attention in the research community. The…

性能 · 计算机科学 2023-09-29 Hasibul Jamil , Lavone Rodolph , Jacob Goldverg , Tevfik Kosar

Web-based services often run randomized experiments to improve their products. A popular way to run these experiments is to use geographical regions as units of experimentation, since this does not require tracking of individual users or…

社会与信息网络 · 计算机科学 2019-02-19 David Rolnick , Kevin Aydin , Jean Pouget-Abadie , Shahab Kamali , Vahab Mirrokni , Amir Najmi

Local clustering aims to identify a cluster within a given graph that includes a designated seed node or a significant portion of a group of seed nodes. This cluster should be well-characterized, i.e., it has a high number of internal edges…

社会与信息网络 · 计算机科学 2023-01-19 Adil Chhabra , Marcelo Fonseca Faraj , Christian Schulz

We investigate the novel problem of voting-based opinion maximization in a social network: Find a given number of seed nodes for a target campaigner, in the presence of other competing campaigns, so as to maximize a voting-based score for…

社会与信息网络 · 计算机科学 2022-09-15 Arkaprava Saha , Xiangyu Ke , Arijit Khan , Laks V. S. Lakshmanan

We introduce a new and increasingly relevant setting for distributed optimization in machine learning, where the data defining the optimization are distributed (unevenly) over an extremely large number of \nodes, but the goal remains to…

机器学习 · 计算机科学 2015-11-12 Jakub Konečný , Brendan McMahan , Daniel Ramage

We propose a simple and efficient clustering method for high-dimensional data with a large number of clusters. Our algorithm achieves high-performance by evaluating distances of datapoints with a subset of the cluster centres. Our…

机器学习 · 计算机科学 2022-03-30 Georgios Exarchakis , Omar Oubari , Gregor Lenz

Current dataset collection methods typically scrape large amounts of data from the web. While this technique is extremely scalable, data collected in this way tends to reinforce stereotypical biases, can contain personally identifiable…

计算机视觉与模式识别 · 计算机科学 2025-09-15 Vikram V. Ramaswamy , Sing Yu Lin , Dora Zhao , Aaron B. Adcock , Laurens van der Maaten , Deepti Ghadiyaram , Olga Russakovsky

Climate change is posing new challenges to crop-related concerns including food insecurity, supply stability and economic planning. As one of the central challenges, crop yield prediction has become a pressing task in the machine learning…

机器学习 · 计算机科学 2022-01-25 Joshua Fan , Junwen Bai , Zhiyun Li , Ariel Ortiz-Bobea , Carla P. Gomes

Social networks are commonly used for marketing purposes. For example, free samples of a product can be given to a few influential social network users (or "seed nodes"), with the hope that they will convince their friends to buy it. One…

社会与信息网络 · 计算机科学 2019-01-17 Siyu Lei , Silviu Maniu , Luyi Mo , Reynold Cheng , Pierre Senellart

The problem of selecting an optimal seed set to maximise influence in networks has been a subject of intense research in recent years. However, despite numerous works addressing this area, it remains a topic that requires further…

社会与信息网络 · 计算机科学 2025-04-16 Michał Czuba , Piotr Bródka

A new strategy for global geometry optimization of clusters is presented. Important features are a restriction of search space to favorable nearest-neighbor distance ranges, a suitable cluster growth representation with diminished…

chem-ph · 物理学 2009-10-28 Bernd Hartke

Real-world networks often come with side information that can help to improve the performance of network analysis tasks such as clustering. Despite a large number of empirical and theoretical studies conducted on network clustering methods…

机器学习 · 统计学 2022-07-29 Guillaume Braun , Hemant Tyagi , Christophe Biernacki

Methods for ranking the importance of nodes in a network have a rich history in machine learning and across domains that analyze structured data. Recent work has evaluated these methods though the seed set expansion problem: given a subset…

社会与信息网络 · 计算机科学 2017-05-04 Isabel Kloumann , Johan Ugander , Jon Kleinberg

A distribution system can flexibly adjust its substation-level power output by aggregating its local distributed energy resources (DERs). Due to DER and network constraints, characterizing the exact feasible power output region is…

最优化与控制 · 数学 2023-10-10 Qi Li , Jianzhe Liu , Bai Cui , Wenzhan Song , Jin Ye

We consider the problem of data collection from a continental-scale network of energy harvesting sensors, applied to tracking mobile assets in rural environments. Our application constraints favour a highly asymmetric solution, with heavily…

网络与互联网体系结构 · 计算机科学 2016-03-09 Kai Li , Chau Yuen , Branislav Kusy , Raja Jurdak , Aleksandar Ignjatovic , Salil S. Kanhere , Sanjay Jha

Group Search Optimizer(GSO) is one of the best algorithms, is very new in the field of Evolutionary Computing. It is very robust and efficient algorithm, which is inspired by animal searching behaviour. The paper describes an application of…

神经与进化计算 · 计算机科学 2013-08-20 G. Kishore Kumar , V. K. Jayaraman

As a widely observable social effect, influence diffusion refers to a process where innovations, trends, awareness, etc. spread across the network via the social impact among individuals. Motivated by such social effect, the concept of…

社会与信息网络 · 计算机科学 2020-12-24 Liang Ma

Centroid based clustering methods such as k-means, k-medoids and k-centers are heavily applied as a go-to tool in exploratory data analysis. In many cases, those methods are used to obtain representative centroids of the data manifold for…

机器学习 · 计算机科学 2022-06-16 Ahmed Imtiaz Humayun , Randall Balestriero , Anastasios Kyrillidis , Richard Baraniuk

Traditional clustering algorithms often struggle with high-dimensional and non-uniformly distributed data, where low-density boundary samples are easily disturbed by neighboring clusters, leading to unstable and distorted clustering…

机器学习 · 计算机科学 2025-10-28 Qi Li , Jun Wang