English
Related papers

Related papers: A generalized Goulden-Jackson cluster method and l…

200 papers

Pooling is an essential component of a wide variety of sentence representation and embedding models. This paper explores generalized pooling methods to enhance sentence embedding. We propose vector-based multi-head attention that includes…

Computation and Language · Computer Science 2022-02-24 Qian Chen , Zhen-Hua Ling , Xiaodan Zhu

The key in agglomerative clustering is to define the affinity measure between two sets. A novel agglomerative clustering method is proposed by utilizing the path integral to define the affinity measure. Firstly, the path integral descriptor…

Computer Vision and Pattern Recognition · Computer Science 2015-08-10 Wei-Ya Ren , Shuo-Hao Li , Qiang Guo , Guo-Hui Li , Jun Zhang

Clustering is an unsupervised machine learning methodology where unlabeled elements/objects are grouped together aiming to the construction of well-established clusters that their elements are classified according to their similarity. The…

Machine Learning · Statistics 2023-10-20 Dimitrios Saligkaras , Vasileios E. Papageorgiou

Algorithms for node clustering typically focus on finding homophilous structure in graphs. That is, they find sets of similar nodes with many edges within, rather than across, the clusters. However, graphs often also exhibit heterophilous…

Machine Learning · Computer Science 2023-08-15 Sudhanshu Chanpuriya , Cameron Musco

In the present paper we define the notion of generalized cumulants which gives a universal framework for commutative, free, Boolean, and especially, monotone probability theories. The uniqueness of generalized cumulants holds for each…

Probability · Mathematics 2015-05-13 Takahiro Hasebe , Hayato Saigo

This article concerns a class of generalized linear mixed models for clustered data, where the random effects are mapped uniquely onto the grouping structure and are independent between groups. We derive necessary and sufficient conditions…

Methodology · Statistics 2017-09-20 Jarod Y. L. Lee , Peter J. Green , Louise M. Ryan

We introduce a generalized forward-backward splitting method with penalty term for solving monotone inclusion problems involving the sum of a finite number of maximally monotone operators and the normal cone to the nonempty set of zeros of…

Optimization and Control · Mathematics 2018-07-31 Nimit Nimana , Narin Petrot

In this paper, we use subword complexes to provide a uniform approach to finite type cluster complexes and multi-associahedra. We introduce, for any finite Coxeter group and any nonnegative integer k, a spherical subword complex called…

Combinatorics · Mathematics 2013-07-11 Cesar Ceballos , Jean-Philippe Labbé , Christian Stump

Determining the number of clusters in a dataset is a fundamental issue in data clustering. Many methods have been proposed to solve the problem of selecting the number of clusters, considering it to be a problem with regard to model…

Machine Learning · Computer Science 2022-10-04 Ryosuke Motegi , Yoichi Seki

We study the possibility of producing and detecting continuous variable cluster states in an optical set-up in an extremely compact fashion. This method is based on a multi-pixel homodyne detection system recently demonstrated…

Quantum Physics · Physics 2015-06-15 Giulia Ferrini , Jean-Pierre Gazeau , Thomas Coudreau , Claude Fabre , Nicolas Treps

In this paper, we propose an alternative to deep neural networks for semantic information retrieval for the case of long documents. This new approach exploiting clustering techniques to take into account the meaning of words in Information…

Information Retrieval · Computer Science 2025-07-29 Paul Mbathe Mekontchou , Armel Fotsoh , Bernabe Batchakui , Eddy Ella

The input of most clustering algorithms is a symmetric matrix quantifying similarity within data pairs. Such a matrix is here turned into a quadratic set function measuring cluster score or similarity within data subsets larger than pairs.…

Discrete Mathematics · Computer Science 2015-09-30 Giovanni Rossi

This paper proposes a new linearized mixed data sampling (MIDAS) model and develops a framework to infer clusters in a panel regression with mixed frequency data. The linearized MIDAS estimation method is more flexible and substantially…

Econometrics · Economics 2021-02-04 Yeonwoo Rho , Yun Liu , Hie Joo Ahn

The Widom-Rowlinson model of a fluid mixture is studied using a new cluster algorithm that is a generalization of the invaded cluster algorithm previously applied to Potts models. Our estimate of the critical exponents for the two-component…

Statistical Mechanics · Physics 2009-10-30 Gregory Johnson , Harvey Gould , J. Machta , L. K. Chayes

Matrix factorization is a well-studied task in machine learning for compactly representing large, noisy data. In our approach, instead of using the traditional concept of matrix rank, we define a new notion of link-rank based on a…

Machine Learning · Statistics 2018-05-02 Pouya Pezeshkpour , Carlos Guestrin , Sameer Singh

We propose a model-based clustering algorithm for a general class of functional data for which the components could be curves or images. The random functional data realizations could be measured with error at discrete, and possibly random,…

Machine Learning · Statistics 2022-03-14 Steven Golovkine , Nicolas Klutchnikoff , Valentin Patilea

Local explanation methods highlight the input tokens that have a considerable impact on the outcome of classifying the document at hand. For example, the Anchor algorithm applies a statistical analysis of the sensitivity of the classifier…

Machine Learning · Computer Science 2024-01-15 Alon Mor , Yonatan Belinkov , Benny Kimelfeld

Motzkin paths are simple yet important combinatorial objects. In this paper, we consider families of Motzkin paths with restrictions on peak heights, valley heights, upward-run lengths, downward-run lengths, and flat-run lengths. This paper…

Combinatorics · Mathematics 2020-10-07 AJ Bu

To generate summaries that include multiple aspects or topics for text documents, most approaches use clustering or topic modeling to group relevant sentences and then generate a summary for each group. These approaches struggle to optimize…

Artificial Intelligence · Computer Science 2024-05-30 Xiaobo Guo , Jay Desai , Srinivasan H. Sengamedu

The (usual) Caldero-Chapoton map is a map from the set of objects of a category to a Laurent polynomial ring over the integers. In the case of a cluster category, it maps "reachable" indecomposable objects to the corresponding cluster…

Representation Theory · Mathematics 2018-12-14 Thorsten Holm , Peter Jorgensen