中文
相关论文

相关论文: Product/Brand extraction from WikiPedia

200 篇论文

Understanding how various external campaigns or events affect readership on Wikipedia is important to efforts aimed at improving awareness and access to its content. In this paper, we consider how to build time-series models aimed at…

计算机与社会 · 计算机科学 2019-03-27 Xiaoxi Chelsy Xie , Isaac Johnson , Anne Gomez

With increasing importance of e-commerce, many websites have emerged where users can express their opinions about products, such as movies, books, songs, etc. Such interactions can be modeled as bipartite graphs where the weight of the…

信息检索 · 计算机科学 2016-03-16 Abhinav Mishra

We propose a new approach to explain Bayesian Networks. The approach revolves around a new definition of a probabilistic argument and the evidence it provides. We define a notion of independent arguments, and propose an algorithm to extract…

人工智能 · 计算机科学 2021-12-03 Jaime Sevilla

Social tagging has become an interesting approach to improve search and navigation over the actual Web, since it aggregates the tags added by different users to the same resource in a collaborative way. This way, it results in a list of…

信息检索 · 计算机科学 2012-02-27 Arkaitz Zubiaga

Generating semantic lexicons semi-automatically could be a great time saver, relative to creating them by hand. In this paper, we present an algorithm for extracting potential entries for a category from an on-line corpus, based upon a…

计算与语言 · 计算机科学 2007-05-23 Brian Roark , Eugene Charniak

Search engines are a combination of hardware and computer software supplied by a particular company through the website which has been determined. Search engines collect information from the web through bots or web crawlers that crawls the…

信息检索 · 计算机科学 2014-10-22 Ahmad Josi , Leon Andretti Abdillah , Suryayusra

This paper describes a new method to extract relevant keywords from patent claims, as part of the task of retrieving other patents with similar claims (search for prior art). The method combines a qualitative analysis of the writing style…

信息检索 · 计算机科学 2019-06-19 Julien Rossi , Matthias Wirth , Evangelos Kanoulas

Studies of different term extractors on a corpus of the biomedical domain revealed decreasing performances when applied to highly technical texts. The difficulty or impossibility of customising them to new domains is an additional…

计算与语言 · 计算机科学 2007-05-23 Sophie Aubin , Thierry Hamon

With the advent of semantic web, various tools and techniques have been introduced for presenting and organizing knowledge. Concept hierarchies are one such technique which gained significant attention due to its usefulness in creating…

人工智能 · 计算机科学 2016-11-30 V. S. Anoop , S. Asharaf , P. Deepak

Because of the data deluge in scientific publication, finding relevant information is getting harder and harder for researchers and readers. Building an enhanced scientific search engine by taking semantic relations into account poses a…

信息检索 · 计算机科学 2017-09-29 Bastien Latard , Jonathan Weber , Germain Forestier , Michel Hassenforder

Machine learning applications to symbolic mathematics are becoming increasingly popular, yet there lacks a centralized source of real-world symbolic expressions to be used as training data. In contrast, the field of natural language…

机器学习 · 计算机科学 2022-07-06 Joanne T. Kim , Mikel Landajuela , Brenden K. Petersen

The instances of templates in Wikipedia form an interesting data set of structured information. Here I focus on the cite journal template that is primarily used for citation to articles in scientific journals. These citations can be…

数字图书馆 · 计算机科学 2008-06-12 Finn Aarup Nielsen

In this paper, we present a novel approach to identify feature specific expressions of opinion in product reviews with different features and mixed emotions. The objective is realized by identifying a set of potential features in the review…

信息检索 · 计算机科学 2012-09-19 Subhabrata Mukherjee , Pushpak Bhattacharyya

Generating factual, long-form text such as Wikipedia articles raises three key challenges: how to gather relevant evidence, how to structure information into well-formed text, and how to ensure that the generated text is factually correct.…

计算与语言 · 计算机科学 2022-04-13 Angela Fan , Claire Gardent

With the large volume of unstructured data that increases constantly on the web, the motivation of representing the knowledge in this data in the machine-understandable form is increased. Ontology is one of the major cornerstones of…

计算与语言 · 计算机科学 2021-05-10 Fatima N. AL-Aswadi , Huah Yong Chan , Keng Hoon Gan

Wikipedia is the world's largest online encyclopedia, but maintaining article quality through collaboration is challenging. Wikipedia designed a quality scale, but with such a manual assessment process, many articles remain unassessed. We…

计算与语言 · 计算机科学 2023-10-04 Pedro Miguel Moás , Carla Teixeira Lopes

Nowadays, information describing navigation behaviour of internet users are used in several fields, e-commerce, economy, sociology and data science. Such information can be extracted from different knowledge bases, including…

社会与信息网络 · 计算机科学 2020-08-18 Célestin Coquidé , Włodzimierz Lewoniewski

Consumers often read product reviews to inform their buying decision, as some consumers want to know a specific component of a product. However, because typical sentences on product reviews contain various details, users must identify…

计算与语言 · 计算机科学 2022-07-14 Shogo Anda , Masato Kikuchi , Tadachika Ozono

As the World Wide Web is growing rapidly, it is getting increasingly challenging to gather representative information about it. Instead of crawling the web exhaustively one has to resort to other techniques like sampling to determine the…

数据结构与算法 · 计算机科学 2009-02-11 Eda Baykan , Monika Henzinger , Stefan F. Keller , Sebastian De Castelberg , Markus Kinzler

This paper presents a deep learning based approach to extract product comparison information out of user reviews on various e-commerce websites. Any comparative product review has three major entities of information: the names of the…

信息检索 · 计算机科学 2023-11-01 Jatin Arora , Sumit Agrawal , Pawan Goyal , Sayan Pathak
‹ 上一页 1 8 9 10 下一页 ›