English
Related papers

Related papers: PolyIE: A Dataset of Information Extraction from P…

200 papers

Scientific information extraction (SciIE) is critical for converting unstructured knowledge from scholarly articles into structured data (entities and relations). Several datasets have been proposed for training and validating SciIE models.…

Computation and Language · Computer Science 2024-10-29 Qi Zhang , Zhijia Chen , Huitong Pan , Cornelia Caragea , Longin Jan Latecki , Eduard Dragut

Automatically extracting key information from scientific documents has the potential to help scientists work more efficiently and accelerate the pace of scientific progress. Prior work has considered extracting document-level entity…

Digital Libraries · Computer Science 2021-06-04 Vijay Viswanathan , Graham Neubig , Pengfei Liu

Extracting information from full documents is an important problem in many domains, but most previous work focus on identifying relationships within a sentence or a paragraph. It is challenging to create a large-scale information extraction…

Computation and Language · Computer Science 2020-05-04 Sarthak Jain , Madeleine van Zuylen , Hannaneh Hajishirzi , Iz Beltagy

The number of published articles in the field of materials science is growing rapidly every year. This comparatively unstructured data source, which contains a large amount of information, has a restriction on its re-usability, as the…

Computation and Language · Computer Science 2021-01-26 Souradip Guha , Ankan Mullick , Jatin Agrawal , Swetarekha Ram , Samir Ghui , Seung-Cheol Lee , Satadeep Bhattacharjee , Pawan Goyal

In the rapidly evolving field of scientific research, efficiently extracting key information from the burgeoning volume of scientific papers remains a formidable challenge. This paper introduces an innovative framework designed to automate…

Information Retrieval · Computer Science 2024-01-31 Yangyang Liu , Shoubin Li

Open Information Extraction (OpenIE) aims to extract structured relational tuples (subject, relation, object) from sentences and plays critical roles for many downstream NLP applications. Existing solutions perform extraction at sentence…

Computation and Language · Computer Science 2021-05-12 Kuicai Dong , Yilin Zhao , Aixin Sun , Jung-Jae Kim , Xiaoli Li

Extracting key information from scientific papers has the potential to help researchers work more efficiently and accelerate the pace of scientific progress. Over the last few years, research on Scientific Information Extraction (SciIE)…

Computation and Language · Computer Science 2023-12-19 Yuhan Li , Jian Wu , Zhiwei Yu , Börje F. Karlsson , Wei Shen , Manabu Okumura , Chin-Yew Lin

Scientific information extraction (SciIE) has primarily relied on entity-relation extraction in narrow domains, limiting its applicability to interdisciplinary research and struggling to capture the necessary context of scientific…

Computation and Language · Computer Science 2025-09-22 Bofu Dong , Pritesh Shah , Sumedh Sonawane , Tiyasha Banerjee , Erin Brady , Xinya Du , Ming Jiang

Automatic extraction of information from publications is key to making scientific knowledge machine readable at a large scale. The extracted information can, for example, facilitate academic search, decision making, and knowledge graph…

Computation and Language · Computer Science 2024-04-02 Tarek Saier , Mayumi Ohta , Takuto Asakura , Michael Färber

Scientific Information Extraction (ScientificIE) is a critical task that involves the identification of scientific entities and their relationships. The complexity of this task is compounded by the necessity for domain-specific knowledge…

Computation and Language · Computer Science 2023-12-27 Dong Pham , Xanh Ho , Quang-Thuy Ha , Akiko Aizawa

Open Information Extraction (OIE) is the task of the unsupervised creation of structured information from text. OIE is often used as a starting point for a number of downstream tasks including knowledge base construction, relation…

Computation and Language · Computer Science 2018-08-23 Paul Groth , Michael Lauruhn , Antony Scerri , Ron Daniel

Information extraction (IE) in scientific literature has facilitated many down-stream tasks. OpenIE, which does not require any relation schema but identifies a relational phrase to describe the relationship between a subject and an object,…

Computation and Language · Computer Science 2021-08-05 Joseph Kuebler , Lingbo Tong , Meng Jiang

Existing scholarly information extraction (SIE) datasets focus on scientific papers and overlook implementation-level details in code repositories. README files describe datasets, source code, and other implementation-level artifacts,…

Computation and Language · Computer Science 2026-03-09 Genet Asefa Gesese , Zongxiong Chen , Shufan Jiang , Mary Ann Tan , Zhaotai Liu , Sonja Schimmler , Harald Sack

Information Extraction (IE) from scientific texts can be used to guide readers to the central information in scientific documents. But narrow IE systems extract only a fraction of the information captured, and Open IE systems do not perform…

Computation and Language · Computer Science 2020-05-27 Ruben Kruiper , Julian F. V. Vincent , Jessica Chen-Burger , Marc P. Y. Desmulliez , Ioannis Konstas

Biomedical information extraction (BioIE) is important to many applications, including clinical decision support, integrative biology, and pharmacovigilance, and therefore it has been an active research. Unlike existing reviews covering a…

Computation and Language · Computer Science 2016-06-28 Feifan Liu , Jinying Chen , Abhyuday Jagannatha , Hong Yu

Developing large-scale foundational datasets is a critical milestone in advancing artificial intelligence (AI)-driven scientific innovation. However, unlike AI-mature fields such as natural language processing, materials science,…

Chemical Physics · Physics 2025-11-18 Ryo Yoshida , Yoshihiro Hayashi , Hidemine Furuya , Ryohei Hosoya , Kazuyoshi Kaneko , Hiroki Sugisawa , Yu Kaneko , Aiko Takahashi , Yoh Noguchi , Shun Nanjo , Keiko Shinoda , Tomu Hamakawa , Mitsuru Ohno , Takuya Kitamura , Misaki Yonekawa , Stephen Wu , Masato Ohnishi , Chang Liu , Teruki Tsurimoto , Arifin , Araki Wakiuchi , Kohei Noda , Junko Morikawa , Teruaki Hayakawa , Junichiro Shiomi , Masanobu Naito , Kazuya Shiratori , Tomoki Nagai , Norio Tomotsu , Hiroto Inoue , Ryuichi Sakashita , Masashi Ishii , Isao Kuwajima , Kenji Furuichi , Norihiko Hiroi , Yuki Takemoto , Takahiro Ohkuma , Keita Yamamoto , Naoya Kowatari , Masato Suzuki , Naoya Matsumoto , Seiryu Umetani , Hisaki Ikebata , Yasuyuki Shudo , Mayu Nagao , Shinya Kamada , Kazunori Kamio , Taichi Shomura , Kensaku Nakamura , Yudai Iwamizu , Atsutoshi Abe , Koki Yoshitomi , Yuki Horie , Katsuhiko Koike , Koichi Iwakabe , Shinya Gima , Kota Usui , Gikyo Usuki , Takuro Tsutsumi , Keitaro Matsuoka , Kazuki Sada , Masahiro Kitabata , Takuma Kikutsuji , Akitaka Kamauchi , Yusuke Iijima , Tsubasa Suzuki , Takenori Goda , Yuki Takabayashi , Kazuko Imai , Yuji Mochizuki , Hideo Doi , Koji Okuwaki , Hiroya Nitta , Taku Ozawa , Hitoshi Kamijima , Toshiaki Shintani , Takuma Mitamura , Massimiliano Zamengo , Yuitsu Sugami , Seiji Akiyama , Yoshinari Murakami , Atsushi Betto , Naoya Matsuo , Satoru Kagao , Tetsuya Kobayashi , Norie Matsubara , Shosei Kubo , Yuki Ishiyama , Yuri Ichioka , Mamoru Usami , Satoru Yoshizaki , Seigo Mizutani , Yosuke Hanawa , Shogo Kunieda , Mitsuru Yambe , Takeru Nakamura , Hiromori Murashima , Kenji Takahashi , Naoki Wada , Masahiro Kawano , Yosuke Harada , Takehiro Fujita , Erina Fujita , Ryoji Himeno , Hiori Kino , Kenji Fukumizu

Objectives: Despite the recent adoption of large language models (LLMs) for biomedical information extraction, challenges in prompt engineering and algorithms persist, with no dedicated software available. To address this, we developed…

Machine Learning · Computer Science 2025-04-02 Enshuo Hsu , Kirk Roberts

No existing dataset adequately tests how well language models can incrementally update entity summaries - a crucial ability as these models rapidly advance. The Incremental Entity Summarization (IES) task is vital for maintaining accurate,…

Computation and Language · Computer Science 2024-06-10 Eunjeong Hwang , Yichao Zhou , Beliz Gunel , James Bradley Wendt , Sandeep Tata

Open Information Extraction (OIE) systems seek to compress the factual propositions of a sentence into a series of n-ary tuples. These tuples are useful for downstream tasks in natural language processing like knowledge base creation,…

Computation and Language · Computer Science 2021-01-28 Jacob Solawetz , Stefan Larson

The rapid growth of scientific literature has made manual extraction of structured knowledge increasingly impractical. To address this challenge, we introduce SCILIRE, a system for creating datasets from scientific literature. SCILIRE has…

Computation and Language · Computer Science 2026-03-16 Necva Bölücü , Jessica Irons , Changhyun Lee , Brian Jin , Maciej Rybinski , Huichen Yang , Andreas Duenser , Stephen Wan
‹ Prev 1 2 3 10 Next ›