English
Related papers

Related papers: What is Wrong with Net Promoter Score

200 papers

Reliable evaluation protocols are of utmost importance for reproducible NLP research. In this work, we show that sometimes neither metric nor conventional human evaluation is sufficient to draw conclusions about system performance. Using…

Computation and Language · Computer Science 2021-01-25 Yevgeniy Puzikov

In black-box optimization, noise in the objective function is inevitable. Noise disrupts the ranking of candidate solutions in comparison-based optimization, possibly deteriorating the search performance compared with a noiseless scenario.…

Neural and Evolutionary Computing · Computer Science 2024-01-26 Daiki Morinaga , Youhei Akimoto

Recommender systems are known to suffer from the popularity bias problem: popular (i.e. frequently rated) items get a lot of exposure while less popular ones are under-represented in the recommendations. Research in this area has been…

Information Retrieval · Computer Science 2019-09-20 Himan Abdollahpouri , Masoud Mansoury , Robin Burke , Bamshad Mobasher

The state of neural network pruning has been noticed to be unclear and even confusing for a while, largely due to "a lack of standardized benchmarks and metrics" [3]. To standardize benchmarks, first, we need to answer: what kind of…

Computer Vision and Pattern Recognition · Computer Science 2023-02-23 Huan Wang , Can Qin , Yue Bai , Yun Fu

Traditional evaluation of information access systems has focused primarily on average utility across a set of information needs (information retrieval) or users (recommender systems). In this work, we argue that evaluating only with average…

Information Retrieval · Computer Science 2024-10-18 Fernando Diaz

Individuals often navigate several options with incomplete knowledge of their own preferences. Information provisioning tools such as public rankings and personalized recommendations have become central to helping individuals make choices,…

Theoretical Economics · Economics 2025-06-05 Omar Besbes , Yash Kanoria , Akshit Kumar

Recommendation systems are widespread, and through customized recommendations, promise to match users with options they will like. To that end, data on engagement is collected and used. Most recommendation systems are ranking-based, where…

Information Retrieval · Computer Science 2024-05-08 Omar Besbes , Yash Kanoria , Akshit Kumar

Ordinal classification models assign higher penalties to predictions further away from the true class. As a result, they are appropriate for relevant diagnostic tasks like disease progression prediction or medical image grading. The…

Computer Vision and Pattern Recognition · Computer Science 2023-09-19 Adrian Galdran

To address a looming crisis of unreproducible evaluation for named entity recognition, we propose guidelines and introduce SeqScore, a software package to improve reproducibility. The guidelines we propose are extremely simple and center…

Computation and Language · Computer Science 2021-11-08 Chester Palen-Michel , Nolan Holley , Constantine Lignos

Popularity bias is a long-standing challenge in recommender systems. Such a bias exerts detrimental impact on both users and item providers, and many efforts have been dedicated to studying and solving such a bias. However, most existing…

Information Retrieval · Computer Science 2022-08-03 Ziwei Zhu , Yun He , Xing Zhao , James Caverlee

The inner-product navigable small world graph (ip-NSW) represents the state-of-the-art method for approximate maximum inner product search (MIPS) and it can achieve an order of magnitude speedup over the fastest baseline. However, to date…

Information Retrieval · Computer Science 2019-12-10 Jie Liu , Xiao Yan , Xinyan Dai , Zhirong Li , James Cheng , Ming-Chang Yang

Despite the massive investments in information security technologies and research over the past decades, the information security industry is still immature. In particular, the prioritization of remediation efforts within vulnerability…

Cryptography and Security · Computer Science 2019-08-15 Jay Jacobs , Sasha Romanosky , Benjamin Edwards , Michael Roytman , Idris Adjerid

The many metrics employed for the evaluation of search engine results have not themselves been conclusively evaluated. We propose a new measure for a metric's ability to identify user preference of result lists. Using this measure, we…

Information Retrieval · Computer Science 2011-03-16 Pavel Sirotkin

Over the last decade proposal success rates in the fundamental sciences have dropped significantly. Astronomy and related fields funded by NASA and NSF are no exception. Data across agencies show that this is not principally the result of a…

Score matching is a recently developed parameter learning method that is particularly effective to complicated high dimensional density models with intractable partition functions. In this paper, we study two issues that have not been…

Machine Learning · Computer Science 2012-05-14 Siwei Lyu

We revisit skip-gram negative sampling (SGNS), one of the most popular neural-network based approaches to learning distributed word representation. We first point out the ambiguity issue undermining the SGNS model, in the sense that the…

Computation and Language · Computer Science 2019-01-15 Cun Mu , Guang Yang , Zheng Yan

Recently there has been a growing interest in fairness-aware recommender systems, including fairness in providing consistent performance across different users or groups of users. A recommender system could be considered unfair if the…

Information Retrieval · Computer Science 2019-10-17 Himan Abdollahpouri , Masoud Mansoury , Robin Burke , Bamshad Mobasher

I prove that competitive market outcomes require computational intractability. If P = NP, firms can efficiently solve the collusion detection problem, identifying deviations from cooperative agreements in complex, noisy markets and thereby…

Computer Science and Game Theory · Computer Science 2026-02-25 Philip Z. Maymin

We look at discovering the impact of market microstructure on equitability for market participants at public exchanges such as the New York Stock Exchange or NASDAQ. Are these environments equitable venues for low-frequency participants…

Multiagent Systems · Computer Science 2021-11-02 Kshama Dwarakanath , Svitlana S Vyetrenko , Tucker Balch

U.S. state education agencies mark schools displaying achievement gaps between demographic subgroups as needing improvement. Some schools may have few students in these subgroups, such that average end-of-year test scores only noisily…

Methodology · Statistics 2025-12-10 Joshua Wasserman , Michael R. Elliott , Ben B. Hansen
‹ Prev 1 4 5 6 7 8 10 Next ›