中文
相关论文

相关论文: Data-Driven Investigative Journalism For Connectas…

200 篇论文

We introduce a set of image transformations that can be used as corruptions to evaluate the robustness of models as well as data augmentation mechanisms for training neural networks. The primary distinction of the proposed transformations…

计算机视觉与模式识别 · 计算机科学 2022-05-02 Oğuzhan Fatih Kar , Teresa Yeo , Andrei Atanov , Amir Zamir

Corruption has a significant impact on economic growth, democracy, and inequality. It has severe consequences at the human level are incalculable. Public procurement, where public resources are used to purchase goods or services from the…

应用统计 · 统计学 2022-12-16 Andrea Falcón-Cortés , Andrés Aldana , Hernán Larralde

Political corruption is inherently an affiliation process linking agents to corruption cases; yet it is often studied via one-mode projections that connect co-offenders within the same scandal, implying a loss of information that…

Text classification models, especially neural networks based models, have reached very high accuracy on many popular benchmark datasets. Yet, such models when deployed in real world applications, tend to perform badly. The primary reason is…

计算与语言 · 计算机科学 2020-02-04 Utkarsh Desai , Srikanth Tamilselvam , Jassimran Kaur , Senthil Mani , Shreya Khare

Training deep neural models in the presence of corrupted supervision is challenging as the corrupted data points may significantly impact the generalization performance. To alleviate this problem, we present an efficient robust algorithm…

机器学习 · 计算机科学 2021-02-16 Boyang Liu , Mengying Sun , Ding Wang , Pang-Ning Tan , Jiayu Zhou

We analyze expenditure patterns of discretionary funds by Brazilian congress members. This analysis is based on a large dataset containing over $7$ million expenses made publicly available by the Brazilian government. This dataset has, up…

计算机与社会 · 计算机科学 2018-12-05 Hsiang Hsu , Flavio P. Calmon , José Cândido Silveira Santos Filho , Andre P. Calmon , Salman Salamatian

Deep learning models often face challenges when handling real-world image corruptions. In response, researchers have developed image corruption datasets to evaluate the performance of deep neural networks in handling such corruptions.…

计算机视觉与模式识别 · 计算机科学 2023-06-13 Harshitha Machiraju , Michael H. Herzog , Pascal Frossard

In supervised learning one wishes to identify a pattern present in a joint distribution $P$, of instances, label pairs, by providing a function $f$ from instances to labels that has low risk $\mathbb{E}_{P}\ell(y,f(x))$. To do so, the…

机器学习 · 统计学 2015-07-07 Brendan van Rooyen , Robert C. Williamson

Deep learning (DL) models are widely used in real-world applications but remain vulnerable to distribution shifts, especially due to weather and lighting changes. Collecting diverse real-world data for testing the robustness of DL models is…

计算机视觉与模式识别 · 计算机科学 2025-05-09 Shashank Agnihotri , David Schader , Nico Sharei , Mehmet Ege Kaçar , Margret Keuper

Data collection in economically constrained countries often necessitates using approximate and biased measurements due to the low-cost of the sensors used. This leads to potentially invalid predictions and poor policies or decision making.…

机器学习 · 计算机科学 2019-12-02 Michael T. Smith , Joel Ssematimba , Mauricio A. Alvarez , Engineer Bainomugisha

Corruption has been an important issue as it becomes obstacle to achieve the better and more efficient economic governmental system. The paper defines corruption in two ways, as state capture and administrative corruption to grasp the…

适应与自组织系统 · 物理学 2007-05-23 Hokky Situngkir

Complex networked systems can be modeled and represented as graphs, with nodes representing the agents and the links describing the dynamic coupling between them. The fundamental objective of network identification for dynamic systems is to…

系统与控制 · 电气工程与系统科学 2020-06-09 Venkat Ram Subramanian , Andrew Lamperski , Murti V. Salapaka

A novel correction algorithm is proposed for multi-class classification problems with corrupted training data. The algorithm is non-intrusive, in the sense that it post-processes a trained classification model by adding a correction…

机器学习 · 计算机科学 2020-02-13 Jun Hou , Tong Qin , Kailiang Wu , Dongbin Xiu

Machine Learning models increasingly face data integrity challenges due to the use of large-scale training datasets drawn from the Internet. We study what model developers can do if they detect that some data was manipulated or incorrect.…

机器学习 · 计算机科学 2024-10-18 Shashwat Goel , Ameya Prabhu , Philip Torr , Ponnurangam Kumaraguru , Amartya Sanyal

The rise of financial crime that has been observed in recent years has created an increasing concern around the topic and many people, organizations and governments are more and more frequently trying to combat it. Despite the increase of…

Machine learning systems are increasingly used to support public sector decision-making across a variety of sectors. Given concerns around accountability in these domains, and amidst accusations of intentional or unintentional bias, there…

计算机与社会 · 计算机科学 2018-11-06 Michael Veale

Crime linkage is the process of analyzing criminal behavior data to determine whether a pair or group of crime cases are connected or belong to a series of offenses. This domain has been extensively studied by researchers in sociology,…

机器学习 · 计算机科学 2024-11-05 Vinicius Lima , Umit Karabiyik

Missing covariates in regression or classification problems can prohibit the direct use of advanced tools for further analysis. Recent research has realized an increasing trend towards the usage of modern Machine Learning algorithms for…

机器学习 · 统计学 2022-03-23 Burim Ramosaj , Justus Tulowietzki , Markus Pauly

Real data are rarely pure. Hence the past half-century has seen great interest in robust estimation algorithms that perform well even when part of the data is corrupt. However, their vast majority approach optimal accuracy only when given a…

机器学习 · 计算机科学 2022-02-14 Ayush Jain , Alon Orlitsky , Vaishakh Ravindrakumar

It is needed to ensure the integrity of systems that process sensitive information and control many aspects of everyday life. We examine the use of machine learning algorithms to detect malware using the system calls generated by…