中文
相关论文

相关论文: Characterising harmful data sources when construct…

200 篇论文

A growing ecosystem of large, open-source foundation models has reduced the labeled data and technical expertise necessary to apply machine learning to many new problems. Yet foundation models pose a clear dual-use risk, indiscriminately…

机器学习 · 计算机科学 2023-08-10 Peter Henderson , Eric Mitchell , Christopher D. Manning , Dan Jurafsky , Chelsea Finn

Simulation models are widely used in practice to facilitate decision-making in a complex, dynamic and stochastic environment. But they are computationally expensive to execute and optimize, due to lack of analytical tractability. Simulation…

最优化与控制 · 数学 2021-06-14 L. Jeff Hong , Xiaowei Zhang

The estimation of unknown values of parameters (or hidden variables, control variables) that characterise a physical system often relies on the comparison of measured data with synthetic data produced by some numerical simulator of the…

机器学习 · 计算机科学 2019-01-28 Xi Chen , Mike Hobson

Metaheuristic search algorithms look for solutions that either maximise or minimise a set of objectives, such as cost or performance. However most real-world optimisation problems consist of nonlinear problems with complex constraints and…

神经与进化计算 · 计算机科学 2022-06-29 Manjinder Singh , Alexander E. I. Brownlee , David Cairns

In computational social science (CSS), researchers analyze documents to explain social and political phenomena. In most scenarios, CSS researchers first obtain labels for documents and then explain labels using interpretable regression…

统计方法学 · 统计学 2024-01-17 Naoki Egami , Musashi Hinck , Brandon M. Stewart , Hanying Wei

Fast inference of numerical model parameters from data is an important prerequisite to generate predictive models for a wide range of applications. Use of sampling-based approaches such as Markov chain Monte Carlo may become intractable…

机器学习 · 计算机科学 2022-08-10 Yu Wang , Fang Liu , Daniele E. Schiavazzi

Estimating the probability of failure for complex real-world systems using high-fidelity computational models is often prohibitively expensive, especially when the probability is small. Exploiting low-fidelity models can make this process…

In order to optimally design materials, it is crucial to understand the structure-property relations in the material by analyzing the effect of microstructure parameters on the macroscopic properties. In computational homogenization, the…

计算工程、金融与科学 · 计算机科学 2022-08-24 Theron Guo , Ondřej Rokoš , Karen Veroy

In multi-objective design tasks, the computational cost increases rapidly when high-fidelity simulations are used to evaluate objective functions. Surrogate models help mitigate this cost by approximating the simulation output, simplifying…

应用统计 · 统计学 2025-06-03 Omer F. Erdem , David P. Broughton , Josef Svoboda , Chengkun Huang , Majdi I. Radaideh

When data is generated by multiple sources, conventional training methods update models assuming equal reliability for each source and do not consider their individual data quality. However, in many applications, sources have varied levels…

机器学习 · 计算机科学 2025-02-17 Alexander Capstick , Francesca Palermo , Tianyu Cui , Payam Barnaghi

High-fidelity numerical simulations of partial differential equations (PDEs) given a restricted computational budget can significantly limit the number of parameter configurations considered and/or time window evaluated for modeling a given…

机器学习 · 计算机科学 2023-09-04 Paolo Conti , Mengwu Guo , Andrea Manzoni , Attilio Frangi , Steven L. Brunton , J. Nathan Kutz

Machine learning models are increasingly being used in important decision-making software such as approving bank loans, recommending criminal sentencing, hiring employees, and so on. It is important to ensure the fairness of these models so…

机器学习 · 计算机科学 2020-09-23 Sumon Biswas , Hridesh Rajan

Model merging techniques aim to integrate the abilities of multiple models into a single model. Most model merging techniques have hyperparameters, and their setting affects the performance of the merged model. Because several existing…

When evaluating the effectiveness of a treatment, policy, or intervention, the desired measure of effectiveness may be expensive to collect, not routinely available, or may take a long time to occur. In these cases, it is sometimes possible…

统计方法学 · 统计学 2022-11-10 Denis Agniel , Layla Parast , Boris Hejblum

The present paper proposes a Bayesian framework for inverse problems that seamlessly integrates optimization and inversion to enable rapid surrogate modeling, accurate parameter inference, and rigorous uncertainty quantification. Bayesian…

计算工程、金融与科学 · 计算机科学 2026-02-05 Mihaela Chiappetta , Massimo Carraturo , Alexander Raßloff , Markus Kästner , Ferdinando Auricchio

We study the interplay between surrogate methods for structured prediction and techniques from multitask learning designed to leverage relationships between surrogate outputs. We propose an efficient algorithm based on trace norm…

机器学习 · 计算机科学 2019-03-05 Giulia Luise , Dimitris Stamos , Massimiliano Pontil , Carlo Ciliberto

The performance of machine learning surrogates is critically dependent on data quality and quantity. This presents a major challenge, as high-fidelity (HF) data is often scarce and computationally expensive to acquire, while low-fidelity…

机器学习 · 计算机科学 2026-02-03 Jice Zeng , David Barajas-Solano , Hui Chen

A growing literature on human-AI decision-making investigates strategies for combining human judgment with statistical models to improve decision-making. Research in this area often evaluates proposed improvements to models, interfaces, or…

计算机与社会 · 计算机科学 2023-05-29 Luke Guerdan , Amanda Coston , Zhiwei Steven Wu , Kenneth Holstein

One method to solve expensive black-box optimization problems is to use a surrogate model that approximates the objective based on previous observed evaluations. The surrogate, which is cheaper to evaluate, is optimized instead to find an…

最优化与控制 · 数学 2021-05-28 Rickard Karlsson , Laurens Bliek , Sicco Verwer , Mathijs de Weerdt

Solving inverse problems in cardiovascular modeling is particularly challenging due to the high computational cost of running high-fidelity simulations. In this work, we focus on Bayesian parameter estimation and explore different methods…

机器学习 · 统计学 2025-12-22 Chloe H. Choi , Andrea Zanoni , Daniele E. Schiavazzi , Alison L. Marsden