English
Related papers

Related papers: Rapidly Bootstrapping a Question Answering Dataset…

200 papers

During this pandemic situation, extracting any relevant information related to COVID-19 will be immensely beneficial to the community at large. In this paper, we present a very important resource, COVIDRead, a Stanford Question Answering…

Computation and Language · Computer Science 2021-10-19 Tanik Saikh , Sovan Kumar Sahoo , Asif Ekbal , Pushpak Bhattacharyya

The recent outbreak of the novel coronavirus is wreaking havoc on the world and researchers are struggling to effectively combat it. One reason why the fight is difficult is due to the lack of information and knowledge. In this work, we…

Computation and Language · Computer Science 2020-10-12 Jinhyuk Lee , Sean S. Yi , Minbyul Jeong , Mujeen Sung , Wonjin Yoon , Yonghwa Choi , Miyoung Ko , Jaewoo Kang

In response to the Kaggle's COVID-19 Open Research Dataset (CORD-19) challenge, we have proposed three transformer-based question-answering systems using BERT, ALBERT, and T5 models. Since the CORD-19 dataset is unlabeled, we have evaluated…

Computation and Language · Computer Science 2021-01-28 Hillary Ngai , Yoona Park , John Chen , Mahboobeh Parsapoor

We present COVID-Q, a set of 1,690 questions about COVID-19 from 13 sources, which we annotate into 15 question categories and 207 question clusters. The most common questions in our dataset asked about transmission, prevention, and…

Computation and Language · Computer Science 2023-09-12 Jerry Wei , Chengyu Huang , Soroush Vosoughi , Jason Wei

We present CAiRE-COVID, a real-time question answering (QA) and multi-document summarization system, which won one of the 10 tasks in the Kaggle COVID-19 Open Research Dataset Challenge, judged by medical experts. Our system aims to tackle…

Computation and Language · Computer Science 2020-12-09 Dan Su , Yan Xu , Tiezheng Yu , Farhad Bin Siddique , Elham J. Barezi , Pascale Fung

For the last two years, from 2020 to 2021, COVID-19 has broken disease prevention measures in many countries, including Vietnam, and negatively impacted various aspects of human life and the social community. Besides, the misleading…

Computation and Language · Computer Science 2022-09-15 Triet Minh Thai , Ngan Ha-Thao Chu , Anh Tuan Vo , Son T. Luu

The COVID-19 Open Research Dataset (CORD-19) is a growing resource of scientific papers on COVID-19 and related historical coronavirus research. CORD-19 is designed to facilitate the development of text mining and information retrieval…

Humans gather information by engaging in conversations involving a series of interconnected questions and answers. For machines to assist in information gathering, it is therefore essential to enable them to answer conversational questions.…

Computation and Language · Computer Science 2019-04-02 Siva Reddy , Danqi Chen , Christopher D. Manning

During the outbreak time of COVID-19, computed tomography (CT) is a useful manner for diagnosing COVID-19 patients. Due to privacy issues, publicly available COVID-19 CT datasets are highly difficult to obtain, which hinders the research…

Machine Learning · Computer Science 2020-06-19 Xingyi Yang , Xuehai He , Jinyu Zhao , Yichen Zhang , Shanghang Zhang , Pengtao Xie

In the last few years, open-domain question answering (ODQA) has advanced rapidly due to the development of deep learning techniques and the availability of large-scale QA datasets. However, the current datasets are essentially designed for…

Computation and Language · Computer Science 2022-02-23 Jiexin Wang , Adam Jatowt , Masatoshi Yoshikawa

We present a large, challenging dataset, COUGH, for COVID-19 FAQ retrieval. Similar to a standard FAQ dataset, COUGH consists of three parts: FAQ Bank, Query Bank and Relevance Set. The FAQ Bank contains ~16K FAQ items scraped from 55…

Computation and Language · Computer Science 2021-09-13 Xinliang Frederick Zhang , Heming Sun , Xiang Yue , Simon Lin , Huan Sun

We propose CodeQA, a free-form question answering dataset for the purpose of source code comprehension: given a code snippet and a question, a textual answer is required to be generated. CodeQA contains a Java dataset with 119,778…

Computation and Language · Computer Science 2021-09-20 Chenxiao Liu , Xiaojun Wan

Under the pandemic of COVID-19, people experiencing COVID19-related symptoms or exposed to risk factors have a pressing need to consult doctors. Due to hospital closure, a lot of consulting services have been moved online. Because of the…

Computation and Language · Computer Science 2020-06-19 Wenmian Yang , Guangtao Zeng , Bowen Tan , Zeqian Ju , Subrato Chakravorty , Xuehai He , Shu Chen , Xingyi Yang , Qingyang Wu , Zhou Yu , Eric Xing , Pengtao Xie

The emergence of the novel COVID-19 pandemic has had a significant impact on global healthcare and the economy over the past few months. The virus's rapid widespread has led to a proliferation in biomedical research addressing the pandemic…

Artificial Intelligence · Computer Science 2020-07-21 Chongyan Chen , Islam Akef Ebeid , Yi Bu , Ying Ding

Existing Scholarly Question Answering (QA) methods typically target homogeneous data sources, relying solely on either text or Knowledge Graphs (KGs). However, scholarly information often spans heterogeneous sources, necessitating the…

Computation and Language · Computer Science 2024-12-06 Tilahun Abedissa Taffa , Debayan Banerjee , Yaregal Assabie , Ricardo Usbeck

Question answering over knowledge bases (KBQA) has become a popular approach to help users extract information from knowledge bases. Although several systems exist, choosing one suitable for a particular application scenario is difficult.…

Computation and Language · Computer Science 2022-11-16 Khiem Vinh Tran , Hao Phu Phan , Khang Nguyen Duc Quach , Ngan Luu-Thuy Nguyen , Jun Jo , Thanh Tam Nguyen

Existing question answering (QA) datasets fail to train QA systems to perform complex reasoning and provide explanations for answers. We introduce HotpotQA, a new dataset with 113k Wikipedia-based question-answer pairs with four key…

Computation and Language · Computer Science 2018-09-26 Zhilin Yang , Peng Qi , Saizheng Zhang , Yoshua Bengio , William W. Cohen , Ruslan Salakhutdinov , Christopher D. Manning

This research presents a review of main datasets that are developed for COVID-19 research. We hope this collection will continue to bring together members of the computing community, biomedical experts, and policymakers in the pursuit of…

Computers and Society · Computer Science 2022-07-28 Syed Raza Bashir , Shaina Raza , Vidhi Thakkar , Usman Naseem

What are the latent questions on some textual data? In this work, we investigate using question generation models for exploring a collection of documents. Our method, dubbed corpus2question, consists of applying a pre-trained question…

Information Retrieval · Computer Science 2020-09-22 Gabriela Surita , Rodrigo Nogueira , Roberto Lotufo

We publicly release a new large-scale dataset, called SearchQA, for machine comprehension, or question-answering. Unlike recently released datasets, such as DeepMind CNN/DailyMail and SQuAD, the proposed SearchQA was constructed to reflect…

Computation and Language · Computer Science 2017-06-13 Matthew Dunn , Levent Sagun , Mike Higgins , V. Ugur Guney , Volkan Cirik , Kyunghyun Cho
‹ Prev 1 2 3 10 Next ›