中文
相关论文

相关论文: Automatic ontology generation for data mining usin…

200 篇论文

We explore the implications of using fuzzy techniques (mainly those commonly used in the linguistic description/summarization of data discipline) from a natural language generation perspective. For this, we provide an extensive discussion…

人工智能 · 计算机科学 2016-05-18 A. Ramos-Soto , A. Bugarín , S. Barro

Recently ontologies have been exploited in a wide range of research areas for data modeling and data management. They greatly assists in defining the semantic model of the underlying data combined with domain knowledge. In this paper, we…

数据库 · 计算机科学 2021-06-08 Jiantao Wu , Fabrizio Orlandi , Declan O'Sullivan , Soumyabrata Dev

Ontologies play a central role in structuring knowledge across domains, supporting tasks such as reasoning, data integration, and semantic search. However, their large size and complexity, particularly in fields such as biomedicine,…

人机交互 · 计算机科学 2025-08-19 Vladimir Zhurov , John Kausch , Kamran Sedig , Mostafa Milani

Building taxonomies is often a significant part of building an ontology, and many attempts have been made to automate the creation of such taxonomies from relevant data. The idea in such approaches is either that relevant definitions of the…

人工智能 · 计算机科学 2023-12-12 Mathieu d'Aquin

Ontologies are pivotal for structuring knowledge bases to enhance question answering (QA) systems powered by Large Language Models (LLMs). However, traditional ontology creation relies on manual efforts by domain experts, a process that is…

人工智能 · 计算机科学 2025-06-03 Yash Tiwari , Owais Ahmad Lone , Mayukha Pal

In this paper, we describe an approach to populate an existing ontology with instance information present in the natural language text provided as input. An ontology is defined as an explicit conceptualization of a shared domain. This…

信息检索 · 计算机科学 2013-02-07 Raghu Anantharangachar , Srinivasan Ramani , S Rajagopalan

Cluster analysis is widely used in the areas of machine learning and data mining. Fuzzy clustering is a particular method that considers that a data point can belong to more than one cluster. Fuzzy clustering helps obtain flexible clusters,…

机器学习 · 计算机科学 2018-06-06 Aybükë Oztürk , Stéphane Lallich , Jérôme Darmont

This paper describe a methodology for semi-automatic classification schema definition (a classification schema is a taxonomy of categories useful for automatic document classification). The methodology is based on: (i) an extensional…

其他计算机科学 · 计算机科学 2009-10-06 Erika De Francesco , Salvatore Iiritano , Antonino Spagnolo , Marco Iannelli

Collocations are important for many tasks of Natural language processing such as information retrieval, machine translation, computational lexicography etc. So far many statistical methods have been used for collocation extraction. Almost…

计算与语言 · 计算机科学 2008-11-11 Raj Kishor Bisht , H. S. Dhami

Explainability is a key challenge and a major research theme in AI research for developing intelligent systems that are capable of working with humans more effectively. An obvious choice in developing explainable intelligent systems relies…

人工智能 · 计算机科学 2023-01-06 Erman Acar , Andrea De Domenico , Krishna Manoorkar , Mattia Panettiere

Capability ontologies are increasingly used to model functionalities of systems or machines. The creation of such ontological models with all properties and constraints of capabilities is very complex and can only be done by ontology…

人工智能 · 计算机科学 2024-10-21 Luis Miguel Vieira da Silva , Aljosha Köcher , Felix Gehlhoff , Alexander Fay

In this paper we present clustering method is very sensitive to the initial center values, requirements on the data set too high, and cannot handle noisy data the proposal method is using information entropy to initialize the cluster…

信息检索 · 计算机科学 2011-04-12 K. Suresh

Large Language Models (LLMs) achieve strong performance in analyzing and generating text, yet they struggle with explicit, transparent, and verifiable reasoning over complex texts such as those containing debates. In particular, they lack…

人工智能 · 计算机科学 2026-03-04 Gianvincenzo Alfano , Sergio Greco , Lucio La Cava , Stefano Francesco Monea , Irina Trubitsyna

The quest for acquiring a formal representation of the knowledge of a domain of interest has attracted researchers with various backgrounds into a diverse field called ontology learning. We highlight classical machine learning and data…

人工智能 · 计算机科学 2021-04-06 Ana Ozaki

In the new era of internet systems and applications, a concept of detecting distinguished topics from huge amounts of text has gained a lot of attention. These methods use representation of text in a numerical format -- called embeddings --…

计算与语言 · 计算机科学 2022-05-16 Danial Toufani-Movaghar , Mohammad-Reza Feizi-Derakhshi

The eXtensible Markup Language (XML) can be used as data exchange format in different domains. It allows different parties to exchange data by providing common understanding of the basic concepts in the domain. XML covers the syntactic…

数字图书馆 · 计算机科学 2012-06-05 Nora Yahia , Sahar A. Mokhtar , AbdelWahab Ahmed

A key challenge for Industry 4.0 applications is to develop control systems for automated manufacturing services that are capable of addressing both data integration and semantic interoperability issues, as well as monitoring and decision…

人工智能 · 计算机科学 2022-10-11 Massimo Carraturo , Andrea Mazzullo

The traditional prototype based clustering methods, such as the well-known fuzzy c-mean (FCM) algorithm, usually need sufficient data to find a good clustering partition. If the available data is limited or scarce, most of the existing…

机器学习 · 计算机科学 2016-04-06 Zhaohong Deng , Yizhang Jiang , Fu-Lai Chung , Hisao Ishibuchi , Kup-Sze Choi , Shitong Wang

Large Language Models (LLMs) have shown significant potential for ontology engineering. However, it is still unclear to what extent they are applicable to the task of domain-specific ontology generation. In this study, we explore the…

Presently, a very large number of public and private data sets are available around the local governments. In most cases, they are not semantically interoperable and a huge human effort is needed to create integrated ontologies and…

数据库 · 计算机科学 2020-05-11 Pierfrancesco Bellini , Paolo Nesi , Nadia Rauch