中文
相关论文

相关论文: COVIDRead: A Large-scale Question Answering Datase…

200 篇论文

We present the Stanford Question Answering Dataset (SQuAD), a new reading comprehension dataset consisting of 100,000+ questions posed by crowdworkers on a set of Wikipedia articles, where the answer to each question is a segment of text…

计算与语言 · 计算机科学 2016-10-12 Pranav Rajpurkar , Jian Zhang , Konstantin Lopyrev , Percy Liang

We present CovidQA, the beginnings of a question answering dataset specifically designed for COVID-19, built by hand from knowledge gathered from Kaggle's COVID-19 Open Research Dataset Challenge. To our knowledge, this is the first…

计算与语言 · 计算机科学 2020-04-24 Raphael Tang , Rodrigo Nogueira , Edwin Zhang , Nikhil Gupta , Phuong Cam , Kyunghyun Cho , Jimmy Lin

We present COVID-Q, a set of 1,690 questions about COVID-19 from 13 sources, which we annotate into 15 question categories and 207 question clusters. The most common questions in our dataset asked about transmission, prevention, and…

计算与语言 · 计算机科学 2023-09-12 Jerry Wei , Chengyu Huang , Soroush Vosoughi , Jason Wei

The recent outbreak of the novel coronavirus is wreaking havoc on the world and researchers are struggling to effectively combat it. One reason why the fight is difficult is due to the lack of information and knowledge. In this work, we…

计算与语言 · 计算机科学 2020-10-12 Jinhyuk Lee , Sean S. Yi , Minbyul Jeong , Mujeen Sung , Wonjin Yoon , Yonghwa Choi , Miyoung Ko , Jaewoo Kang

Existing Scholarly Question Answering (QA) methods typically target homogeneous data sources, relying solely on either text or Knowledge Graphs (KGs). However, scholarly information often spans heterogeneous sources, necessitating the…

计算与语言 · 计算机科学 2024-12-06 Tilahun Abedissa Taffa , Debayan Banerjee , Yaregal Assabie , Ricardo Usbeck

For the last two years, from 2020 to 2021, COVID-19 has broken disease prevention measures in many countries, including Vietnam, and negatively impacted various aspects of human life and the social community. Besides, the misleading…

计算与语言 · 计算机科学 2022-09-15 Triet Minh Thai , Ngan Ha-Thao Chu , Anh Tuan Vo , Son T. Luu

The COVID-19 global pandemic has resulted in international efforts to understand, track, and mitigate the disease, yielding a significant corpus of COVID-19 and SARS-CoV-2-related publications across scientific disciplines. As of May 2020,…

信息检索 · 计算机科学 2020-06-18 Andre Esteva , Anuprit Kale , Romain Paulus , Kazuma Hashimoto , Wenpeng Yin , Dragomir Radev , Richard Socher

The COVID-19 Open Research Dataset (CORD-19) is a growing resource of scientific papers on COVID-19 and related historical coronavirus research. CORD-19 is designed to facilitate the development of text mining and information retrieval…

Community question answering and discussion platforms such as Reddit, Yahoo! answers or Quora provide users the flexibility of asking open ended questions to a large audience, and replies to such questions maybe useful both to the user and…

信息检索 · 计算机科学 2021-01-28 Manisha Verma , Kapil Thadani , Shaunak Mishra

We present a large, challenging dataset, COUGH, for COVID-19 FAQ retrieval. Similar to a standard FAQ dataset, COUGH consists of three parts: FAQ Bank, Query Bank and Relevance Set. The FAQ Bank contains ~16K FAQ items scraped from 55…

计算与语言 · 计算机科学 2021-09-13 Xinliang Frederick Zhang , Heming Sun , Xiang Yue , Simon Lin , Huan Sun

We present CAiRE-COVID, a real-time question answering (QA) and multi-document summarization system, which won one of the 10 tasks in the Kaggle COVID-19 Open Research Dataset Challenge, judged by medical experts. Our system aims to tackle…

计算与语言 · 计算机科学 2020-12-09 Dan Su , Yan Xu , Tiezheng Yu , Farhad Bin Siddique , Elham J. Barezi , Pascale Fung

Extractive reading comprehension systems can often locate the correct answer to a question in a context document, but they also tend to make unreliable guesses on questions for which the correct answer is not stated in the context. Existing…

计算与语言 · 计算机科学 2018-06-12 Pranav Rajpurkar , Robin Jia , Percy Liang

COVID-19 is one of the most important topic these days, specifically on search engines and news. While fake news are easily shared, scientific papers are reliable sources where information can be extracted. With about 24,000 scientific…

数字图书馆 · 计算机科学 2020-05-04 Bernard Dousset , Josiane Mothe

COVID-19 has resulted in an ongoing pandemic and as of 12 June 2020, has caused more than 7.4 million cases and over 418,000 deaths. The highly dynamic and rapidly evolving situation with COVID-19 has made it difficult to access accurate,…

信息检索 · 计算机科学 2020-06-25 David Oniani , Yanshan Wang

Reading comprehension has been widely studied. One of the most representative reading comprehension tasks is Stanford Question Answering Dataset (SQuAD), on which machine is already comparable with human. On the other hand, accessing large…

计算与语言 · 计算机科学 2018-04-03 Chia-Hsuan Li , Szu-Lin Wu , Chi-Liang Liu , Hung-yi Lee

During the outbreak time of COVID-19, computed tomography (CT) is a useful manner for diagnosing COVID-19 patients. Due to privacy issues, publicly available COVID-19 CT datasets are highly difficult to obtain, which hinders the research…

机器学习 · 计算机科学 2020-06-19 Xingyi Yang , Xuehai He , Jinyu Zhao , Yichen Zhang , Shanghang Zhang , Pengtao Xie

The COVID-19 pandemic is accompanied by a massive "infodemic" that makes it hard to identify concise and credible information for COVID-19-related questions, like incubation time, infection rates, or the effectiveness of vaccines. As a…

计算与语言 · 计算机科学 2022-04-20 Johannes Graf , Gino Lancho , Patrick Zschech , Kai Heinrich

The COVID-19 pandemic has had adverse effects on both physical and mental health. During this pandemic, numerous studies have focused on gaining insights into health-related perspectives from social media. In this study, our primary…

机器学习 · 计算机科学 2024-12-02 Mahathir Mohammad Bishal , Md. Rakibul Hassan Chowdory , Anik Das , Muhammad Ashad Kabir

We present TriviaQA, a challenging reading comprehension dataset containing over 650K question-answer-evidence triples. TriviaQA includes 95K question-answer pairs authored by trivia enthusiasts and independently gathered evidence…

计算与语言 · 计算机科学 2017-05-16 Mandar Joshi , Eunsol Choi , Daniel S. Weld , Luke Zettlemoyer

The onset of the COVID-19 pandemic accentuated the need for access to biomedical literature to answer timely and disease-specific questions. During the early days of the pandemic, one of the biggest challenges we faced was the lack of…

计算与语言 · 计算机科学 2023-09-29 Chumki Basu , Himanshu Garg , Allen McIntosh , Sezai Sablak , John R. Wullert
‹ 上一页 1 2 3 10 下一页 ›