English
Related papers

Related papers: Estimating Open Access Mandate Effectiveness: The …

200 papers

Label error is a ubiquitous problem in annotated data. Large amounts of label error substantially degrades the quality of deep learning models. Existing methods to tackle the label error problem largely focus on the classification task, and…

Rubrics provide a flexible way to train LLMs on open-ended long-form answers where verifiable rewards are not applicable and human preferences provide coarse signals. Prior work shows that reinforcement learning with rubric-based rewards…

Computation and Language · Computer Science 2025-10-13 MohammadHossein Rezaei , Robert Vacareanu , Zihao Wang , Clinton Wang , Bing Liu , Yunzhong He , Afra Feyza Akyürek

The quantification and inference of predictive importance for exposure covariates have recently gained significant attention in the context of interpretable machine learning. Contemporary scientific investigations often involve data…

Methodology · Statistics 2024-12-31 Zitao Wang , Nian Si , Zijian Guo , Molei Liu

This study investigates the determinants for the uptake of institutional and subject repository Open Access (OA) in the university landscape of Germany and considers three factors: the disciplinary profile of universities, their OA…

Digital Libraries · Computer Science 2023-08-24 Niels Taubert , Anne Hobert , Najko Jahn , Andre Bruns , Elham Iravani

In this paper, we describe our decade-long experience of building and operating one of the most active Institutional Repository in the world: www.saber.ula.ve <http://www.saber.ula.ve> (University of the Andes, Merida-Venezuela). In order…

Digital Libraries · Computer Science 2009-12-11 Y. Briceno , H. Y. Contreras , L. A. Nunez , F. Salager-Meyer , A. Rojas , R. Torrens

We explicitly test if the reliability of credit ratings depends on the total number of admissible states. We analyse open access credit rating data and show that the effect of the number of states in the dynamical properties of ratings…

Risk Management · Quantitative Finance 2015-06-22 P. Lencastre , F. Raischel , P. G. Lind

Large language models (LLMs) are increasingly trained on tabular data, which, unlike unstructured text, often contains personally identifiable information (PII) in a highly structured and explicit format. As a result, privacy risks arise,…

Cryptography and Security · Computer Science 2025-07-24 Eyal German , Sagiv Antebi , Daniel Samira , Asaf Shabtai , Yuval Elovici

Reinforcement learning with outcome-based feedback faces a fundamental challenge: when rewards are only observed at trajectory endpoints, how do we assign credit to the right actions? This paper provides the first comprehensive analysis of…

Machine Learning · Computer Science 2025-07-25 Fan Chen , Zeyu Jia , Alexander Rakhlin , Tengyang Xie

In the last few decades, Machine Learning (ML) has achieved significant success across domains ranging from healthcare, sustainability, and the social sciences, to criminal justice and finance. But its deployment in increasingly…

Machine Learning · Computer Science 2025-09-03 Nathan Justin , Qingshi Sun , Andrés Gómez , Phebe Vayanos

An important component in deploying machine learning (ML) in safety-critic applications is having a reliable measure of confidence in the ML model's predictions. For a classifier $f$ producing a probability vector $f(x)$ over the candidate…

Machine Learning · Computer Science 2022-10-26 Gal Yona , Amir Feder , Itay Laish

Membership Inference Attack (MIA) determines the presence of a record in a machine learning model's training data by querying the model. Prior work has shown that the attack is feasible when the model is overfitted to its training data or…

Cryptography and Security · Computer Science 2018-02-15 Yunhui Long , Vincent Bindschaedler , Lei Wang , Diyue Bu , Xiaofeng Wang , Haixu Tang , Carl A. Gunter , Kai Chen

Reward models are central to aligning LLMs with human preferences, but they are costly to train, requiring large-scale human-labeled preference data and powerful pretrained LLM backbones. Meanwhile, the increasing availability of…

Computation and Language · Computer Science 2025-10-27 Yapei Chang , Yekyung Kim , Michael Krumdick , Amir Zadeh , Chuan Li , Chris Tanner , Mohit Iyyer

Membership inference attacks (MIAs) have become the standard tool for evaluating privacy leakage in machine learning (ML). Among them, the Likelihood-Ratio Attack (LiRA) is widely regarded as the state of the art when sufficient shadow…

Cryptography and Security · Computer Science 2026-03-10 Najeeb Jebreel , Mona Khalil , David Sánchez , Josep Domingo-Ferrer

This paper presents the experiments accomplished as a part of our participation in the MEDIQA challenge, an (Abacha et al., 2019) shared task. We participated in all the three tasks defined in this particular shared task. The tasks are viz.…

Computation and Language · Computer Science 2021-07-07 Dibyanayan Bandyopadhyay , Baban Gain , Tanik Saikh , Asif Ekbal

The sheer number of research outputs published every year makes systematic reviewing increasingly time- and resource-intensive. This paper explores the use of machine learning techniques to help navigate the systematic review process. ML…

We study whether latent motivation signals in short Spanish admission responses predict engagement and performance in an early quantum computing pathway run by QuantumHub Peru. We analyze N=241 applicants' open responses and link them to…

Physics Education · Physics 2026-02-24 Daniella Alexandra Crysti Vargas Saldana , Freddy Herrera Cueva

Missing data on response variables are common in clinical studies. Corresponding to the uncertainty of missing mechanism, theoretical frameworks on controlled imputation have been developed. In practice, it is recommended to conduct a…

Methodology · Statistics 2022-03-08 Tony Wang , Ying Liu

This paper provides a closed form expression for the pairwise score vector for the multivariate ordered probit model. This result has several implications in likelihood-based inference. It is indeed used both to speed-up gradient based…

Methodology · Statistics 2019-01-30 Martina Bravo , Antonio Canale

Membership inference attacks (MIA) aim to infer whether a particular data point is part of the training dataset of a model. In this paper, we propose a new task in the context of LLM privacy: entity-level discovery of membership risk…

Machine Learning · Computer Science 2025-11-04 Ali Satvaty , Suzan Verberne , Fatih Turkmen

In this work, we present the findings of an online study, where we explore the impact of utilizing embeddings to recommend job postings under real-time constraints. On the Austrian job platform Studo Jobs, we evaluate two popular…

Information Retrieval · Computer Science 2019-07-16 Markus Reiter-Haas , Emanuel Lacic , Tomislav Duricic , Valentin Slawicek , Elisabeth Lex
‹ Prev 1 8 9 10 Next ›