中文
相关论文

相关论文: Weighted Tanimoto Coefficient for 3D Molecule Stru…

200 篇论文

In this article the issues are discussed with the Bayesian approach, least-square fits, and most-likely fits. Trying to counter these issues, a method, based on weighted confidence, is proposed for estimating probabilities and other…

统计理论 · 数学 2017-01-26 Fetze Pijlman

Approximate nearest-neighbor search is a fundamental algorithmic problem that continues to inspire study due its essential role in numerous contexts. In contrast to most prior work, which has focused on point sets, we consider…

计算几何 · 计算机科学 2021-04-01 Ahmed Abdelkader , David M. Mount

We consider DNA codes based on the nearest-neighbor (stem) similarity model which adequately reflects the "hybridization potential" of two DNA sequences. Our aim is to present a survey of bounds on the rate of DNA codes with respect to a…

信息论 · 计算机科学 2016-11-17 A. D'yachkov , A. Voronina , A. Macula , T. Renz , V. Rykov

Proximity graph-based methods have emerged as a leading paradigm for approximate nearest neighbor (ANN) search in the system community. This paper presents fresh insights into the theoretical foundation of these methods. We describe an…

数据结构与算法 · 计算机科学 2025-09-10 Shangqi Lu , Yufei Tao

Advanced engineering materials design involves the exploration of massive multidimensional feature spaces, the correlation of materials properties and the processing parameters derived from disparate sources. The search for alternative…

人工智能 · 计算机科学 2013-01-03 Doreswamy , M. N. Vanajakshi

Pattern recognition constitutes a particularly important task underlying a great deal of scientific and technologica activities. At the same time, pattern recognition involves several challenges, including the choice of features to…

机器学习 · 计算机科学 2024-09-04 Alexandre Benatti , Luciano da F. Costa

Nearest-neighbor methods have become popular in statistics and play a key role in statistical learning. Important decisions in nearest-neighbor methods concern the variables to use (when many potential candidates exist) and how to measure…

统计方法学 · 统计学 2024-01-31 Marcello D'Orazio

Template matching is a basic method in image analysis to extract useful information from images. In this paper, we suggest a new method for pattern matching. Our method transform the template image from two dimensional image into one…

计算机视觉与模式识别 · 计算机科学 2014-09-11 Y. M. Fouda

Histopathology digital scans are large-size images that contain valuable information at the pixel level. Content-based comparison of these images is a challenging task. This study proposes a content-based similarity measure for…

图像与视频处理 · 电气工程与系统科学 2021-07-30 Mehdi Afshari , H. R. Tizhoosh

This paper introduces a new similarity measure based on edge counting in a taxonomy like WorldNet or Ontology. Measurement of similarity between text segments or concepts is very useful for many applications like information retrieval,…

人工智能 · 计算机科学 2012-11-21 Manjula Shenoy. K , K. C. Shet , U. Dinesh Acharya

Grain boundaries are dominant imperfections in nanocrystalline materials that form a complex 3-dimensional (3D) network. Solute segregation to grain boundaries is strongly coupled to the grain boundary character, which governs the stability…

A major bottleneck in nanoparticle measurements is the lack of comparability. Comparability of measurement results is obtained by metrological traceability, which is obtained by calibration. In the present work the calibration of…

应用统计 · 统计学 2018-12-24 J. Pétry , B. De Boeck , N. Sebaihi , M. Coenegrachts , T. Caebergs , M. Dobre

Straight-forward conformation generation models, which generate 3-D structures directly from input molecular graphs, play an important role in various molecular tasks with machine learning, such as 3D-QSAR and virtual screening in drug…

生物大分子 · 定量生物学 2022-03-16 Shuwen Yang , Tianyu Wen , Ziyao Li , Guojie Song

In recent years there have been a lot of interest to test for similarity between biological drug products, commonly known as biologics. Biologics are large and complex molecule drugs that are produced by living cells and hence these are…

统计方法学 · 统计学 2019-02-21 Lin Dong , Sujit K. Ghosh

We consider a similarity measure between two sets $A$ and $B$ of vectors, that balances the average and maximum cosine distance between pairs of vectors, one from set $A$ and one from set $B$. As a motivation for this measure, we present…

数据结构与算法 · 计算机科学 2021-08-31 Michael Leybovich , Oded Shmueli

A set of molecular descriptors whose length is independent of molecular size is developed for machine learning models that target thermodynamic and electronic properties of molecules. These features are evaluated by monitoring performance…

We propose a new "bi-metric" framework for designing nearest neighbor data structures. Our framework assumes two dissimilarity functions: a ground-truth metric that is accurate but expensive to compute, and a proxy metric that is cheaper…

信息检索 · 计算机科学 2024-06-06 Haike Xu , Sandeep Silwal , Piotr Indyk

Similarity search (nearest neighbor search) is a problem of pursuing the data items whose distances to a query item are the smallest from a large database. Various methods have been developed to address this problem, and recently a lot of…

数据结构与算法 · 计算机科学 2014-08-14 Jingdong Wang , Heng Tao Shen , Jingkuan Song , Jianqiu Ji

Measuring the three-dimensional (3D) distribution of chemistry in nanoscale matter is a longstanding challenge for metrological science. The inelastic scattering events required for 3D chemical imaging are too rare, requiring high beam…

Similarity metrics are a core component of many information retrieval and machine learning systems. In this work we propose a method capable of learning a similarity metric from data equipped with a binary relation. By considering only the…

机器学习 · 计算机科学 2016-04-06 Henry Gouk , Bernhard Pfahringer , Michael Cree