English
Related papers

Related papers: Using Fisher Information In Big Data

200 papers

The shortcomings of the traditional univariate distributions in the past greatly encouraged mathematical statisticians to develop new generalizations of distributions. The New Generalized Fisk distribution, a unique distribution presented…

Methodology · Statistics 2025-04-22 Veeranna Banoth

Federated learning aggregates model updates from distributed clients, but standard first order methods such as FedAvg apply the same scalar weight to all parameters from each client. Under non-IID data, these uniformly weighted updates can…

Machine Learning · Computer Science 2026-01-21 Zhipeng Chang , Ting He , Wenrui Hao

Mutual Information (MI) is a powerful statistical measure that quantifies shared information between random variables, particularly valuable in high-dimensional data analysis across fields like genomics, natural language processing, and…

Machine Learning · Computer Science 2024-12-02 Andre O. Falcao

The digital revolution has led to the digitization of human behavior, creating unprecedented opportunities to understand observable actions on an unmatched scale. Emerging phenomena such as crowdfunding and crowdsourcing have further…

Machine Learning · Computer Science 2023-06-27 Hannah H. Chang , Anirban Mukherjee

We describe an exercise of using Big Data to predict the Michigan Consumer Sentiment Index, a widely used indicator of the state of confidence in the US economy. We carry out the exercise from a pure ex ante perspective. We use the…

Statistical Finance · Quantitative Finance 2014-05-23 Rickard Nyman , Paul Ormerod

Informatics and technological advancements have triggered generation of huge volume of data with varied complexity in its management and analysis. Big Data analytics is the practice of revealing hidden aspects of such data and making…

Databases · Computer Science 2018-03-30 Bikram Karmakar , Indranil Mukhopadhyay

Supervised fine-tuning (SFT) is a standard approach to adapting large language models (LLMs) to new domains. In this work, we improve the statistical efficiency of SFT by selecting an informative subset of training examples. Specifically,…

Machine Learning · Computer Science 2025-05-22 Rohan Deb , Kiran Thekumparampil , Kousha Kalantari , Gaurush Hiranandani , Shoham Sabach , Branislav Kveton

Recent advances in artificial intelligence have enabled the generation of large-scale, low-cost predictions with increasingly high fidelity. As a result, the primary challenge in statistical inference has shifted from data scarcity to data…

Statistics Theory · Mathematics 2026-02-12 Shirong Xu , Will Wei Sun

Machine unlearning aims to revoke some training data after learning in response to requests from users, model developers, and administrators. Most previous methods are based on direct fine-tuning, which may neither remove data completely…

Machine Learning · Computer Science 2023-10-10 Yufang Liu , Changzhi Sun , Yuanbin Wu , Aimin Zhou

Quantum Fisher Information (QFI) is a ubiquitous quantity with applications ranging from quantum metrology and resource theories to condensed matter physics. In equilibrium local quantum many-body systems, the QFI of a subsystem with…

Quantum Physics · Physics 2025-04-30 Florent Ferro , Maurizio Fagotti

Background: When developing a clinical prediction model using time-to-event data, previous research focuses on the sample size to minimise overfitting and precisely estimate the overall risk. However, instability of individual-level risk…

Modern large-scale datasets are frequently said to be high-dimensional. However, their data point clouds frequently possess structures, significantly decreasing their intrinsic dimensionality (ID) due to the presence of clusters, points…

Machine Learning · Computer Science 2019-01-21 Luca Albergante , Jonathan Bac , Andrei Zinovyev

Researchers in many disciplines have previously used a variety of mathematical techniques for analyzing group interactions. Here we use a new metric for this purpose, called 'integrated information' or 'phi.' Phi was originally developed by…

Social and Information Networks · Computer Science 2018-11-21 David Engel , Thomas W. Malone

We prove two lower bounds for the complexity of non-log-concave sampling within the framework of Balasubramanian et al. (2022), who introduced the use of Fisher information (FI) bounds as a notion of approximate first-order stationarity in…

Machine Learning · Statistics 2022-10-07 Sinho Chewi , Patrik Gerber , Holden Lee , Chen Lu

Federated Learning (FL) has become increasingly popular across different sectors, offering a way for clients to work together to train a global model without sharing sensitive data. It involves multiple rounds of communication between the…

Machine Learning · Computer Science 2025-07-24 Amandeep Singh Bhatia , Sabre Kais

The Fisher information approximation (FIA) is an implementation of the minimum description length principle for model selection. Unlike information criteria such as AIC or BIC, it has the advantage of taking the functional form of a model…

Methodology · Statistics 2018-08-02 Daniel W. Heck , Morten Moshagen , Edgar Erdfelder

How to design a Markov Decision Process (MDP) based radar controller that makes small sacrifices in performance to mask its sensing plan from an adversary? The radar controller purposefully minimizes the Fisher information of its emissions…

Systems and Control · Electrical Eng. & Systems 2024-03-26 Shashwat Jain , Vikram Krishnamurthy , Muralidhar Rangaswamy , Bosung Kang , Sandeep Gogineni

Fisher's likelihood is widely used for statistical inference for fixed unknowns. This paper aims to extend two important likelihood-based methods, namely the maximum likelihood procedure for point estimation and the confidence procedure for…

Statistics Theory · Mathematics 2025-03-03 Hangbin Lee , Youngjo Lee

It is not unusual for a data analyst to encounter data sets distributed across several computers. This can happen for reasons such as privacy concerns, efficiency of likelihood evaluations, or just the sheer size of the whole data set. This…

Computation · Statistics 2018-05-22 Randy C. S. Lai , J. Hannig , Thomas C. M. Lee

Feature selection, as a data preprocessing strategy, has been proven to be effective and efficient in preparing data (especially high-dimensional data) for various data mining and machine learning problems. The objectives of feature…

Machine Learning · Computer Science 2018-08-28 Jundong Li , Kewei Cheng , Suhang Wang , Fred Morstatter , Robert P. Trevino , Jiliang Tang , Huan Liu
‹ Prev 1 3 4 5 6 7 10 Next ›