English
Related papers

Related papers: Optimal Boolean Locality-Sensitive Hashing

200 papers

We prove a $k^{-\Omega(\log(\varepsilon_2 - \varepsilon_1))}$ lower bound for adaptively testing whether a Boolean function is $\varepsilon_1$-close to or $\varepsilon_2$-far from $k$-juntas. Our results provide the first superpolynomial…

Data Structures and Algorithms · Computer Science 2023-04-24 Xi Chen , Shyamal Patel

We propose a new class of data-independent locality-sensitive hashing (LSH) algorithms based on the fruit fly olfactory circuit. The fundamental difference of this approach is that, instead of assigning hashes as dense points in a low…

Machine Learning · Computer Science 2018-12-06 Jaiyam Sharma , Saket Navlakha

Many problems in classification involve huge numbers of irrelevant features. Model selection reveals the crucial features, reduces the dimensionality of feature space, and improves model interpretation. In the support vector machine…

Methodology · Statistics 2021-10-18 Alfonso Landeros , Kenneth Lange

We study the problem of hypothesis selection under the constraint of local differential privacy. Given a class $\mathcal{F}$ of $k$ distributions and a set of i.i.d. samples from an unknown distribution $h$, the goal of hypothesis selection…

Machine Learning · Statistics 2026-03-05 Alireza F. Pour , Hassan Ashtiani , Shahab Asoodeh

In this paper, Hardy type operator $H_{\beta}$ on $\bR^{n}$ and its adjoint operator $H_{\beta}^{*}$ are investigated. We use novel methods to obtain two main results. One is that we obtain the operators $H_{\beta}$ and $H_{\beta}^{*}$…

Classical Analysis and ODEs · Mathematics 2021-02-03 Qianjun He , Dunyan Yan

The present paper provides exact expressions for the probability distributions of linear functionals of the two-parameter Poisson--Dirichlet process $\operatorname {PD}(\alpha,\theta)$. We obtain distributional results yielding exact forms…

Probability · Mathematics 2009-09-29 Lancelot F. James , Antonio Lijoi , Igor Prünster

We establish a regularity result for optimal sets of the isoperimetric problem with double density under mild ($\alpha$-)H\"older regularity assumptions on the density functions. Our main Theorem improves some previous results and allows to…

Analysis of PDEs · Mathematics 2023-08-15 Lisa Beck , Eleonora Cinti , Christian Seis

We prove that hashing $n$ balls into $n$ bins via a random matrix over $\mathbf{F}_2$ yields expected maximum load $O(\log n / \log \log n)$. This matches the expected maximum load of a fully random function and resolves an open question…

Data Structures and Algorithms · Computer Science 2025-05-21 Michael Jaber , Vinayak M. Kumar , David Zuckerman

We demonstrate the impact of a generic zero-free region and zero-density estimate on the error term in the prime number theorem. Consequently, we are able to improve upon previous work of Pintz and provide an essentially optimal error term…

Number Theory · Mathematics 2025-10-22 Daniel R. Johnston

Given a reproducing kernel Hilbert space H of real-valued functions and a suitable measure mu over the source space D (subset of R), we decompose H as the sum of a subspace of centered functions for mu and its orthogonal in H. This…

Machine Learning · Statistics 2012-12-10 Nicolas Durrande , David Ginsbourger , Olivier Roustant , Laurent Carraro

Sensitivity, block sensitivity and certificate complexity are basic complexity measures of Boolean functions. The famous sensitivity conjecture claims that sensitivity is polynomially related to block sensitivity. However, it has been…

Computational Complexity · Computer Science 2015-06-09 Andris Ambainis , Krišjānis Prūsis , Jevgēnijs Vihrovs

We study the approximation of expectations $\E(f(X))$ for Gaussian random elements $X$ with values in a separable Hilbert space $H$ and Lipschitz continuous functionals $f \colon H \to \R$. We consider restricted Monte Carlo algorithms,…

Numerical Analysis · Mathematics 2018-02-15 Michael B. Giles , Mario Hefter , Lukas Mayer , Klaus Ritter

Algorithms for hyperparameter optimization abound, all of which work well under different and often unverifiable assumptions. Motivated by the general challenge of sequentially choosing which algorithm to use, we study the more specific…

Machine Learning · Statistics 2016-04-12 Robert Nishihara , David Lopez-Paz , Léon Bottou

Consider an oracle which takes a point $x$ and returns the minimizer of a convex function $f$ in an $\ell_2$ ball of radius $r$ around $x$. It is straightforward to show that roughly $r^{-1}\log\frac{1}{\epsilon}$ calls to the oracle…

Optimization and Control · Mathematics 2020-03-19 Yair Carmon , Arun Jambulapati , Qijia Jiang , Yujia Jin , Yin Tat Lee , Aaron Sidford , Kevin Tian

We introduce simple, efficient algorithms for computing a MinHash of a probability distribution, suitable for both sparse and dense data, with equivalent running times to the state of the art for both cases. The collision probability of…

Data Structures and Algorithms · Computer Science 2019-01-04 Ryan Moulton , Yunjiang Jiang

The problem of minimizing convex functionals of probability distributions is solved under the assumption that the density of every distribution is bounded from above and below. A system of sufficient and necessary first-order optimality…

Information Theory · Computer Science 2018-12-05 Michael Fauss , Abdelhak M. Zoubir

We study the problem of testing discrete distributions with a focus on the high probability regime. Specifically, given samples from one or more discrete distributions, a property $\mathcal{P}$, and parameters $0< \epsilon, \delta <1$, we…

Data Structures and Algorithms · Computer Science 2020-09-15 Ilias Diakonikolas , Themis Gouleakis , Daniel M. Kane , John Peebles , Eric Price

Contrastive learning is a representational learning paradigm in which a neural network maps data elements to feature vectors. It improves the feature space by forming lots with an anchor and examples that are either positive or negative…

Computer Vision and Pattern Recognition · Computer Science 2025-05-26 Fabian Deuser , Philipp Hausenblas , Hannah Schieber , Daniel Roth , Martin Werner , Norbert Oswald

We derive a robust error estimate for a recently proposed numerical method for $\alpha$-dissipative solutions of the Hunter-Saxton equation, where $\alpha \in [0, 1]$. In particular, if the following two conditions hold: i) there exist a…

Numerical Analysis · Mathematics 2024-10-10 Thomas Christiansen

Decision trees are one of the most fundamental computational models for computing Boolean functions $f : \{0, 1\}^n \mapsto \{0, 1\}$. It is well-known that the depth and size of decision trees are closely related to time and number of…

Computational Complexity · Computer Science 2025-01-03 Deepu Benson , Balagopal Komarath , Jayalal Sarma , Nalli Sai Soumya