English
Related papers

Related papers: Low-resourced Languages and Online Knowledge Repos…

200 papers

Natural Language Processing is a crucial frontier in artificial intelligence, with broad applications in many areas, including public health, agriculture, education, and commerce. However, due to the lack of substantial linguistic…

Computation and Language · Computer Science 2025-01-22 Audrey Mbogho , Quin Awuor , Andrew Kipkebut , Lilian Wanzare , Vivian Oloo

Large Language Models (LLMs) have rapidly increased in size and apparent capabilities in the last three years, but their training data is largely English text. There is growing interest in multilingual LLMs, and various efforts are striving…

Knowledge is useless without structure. While the classification of knowledge has been an enduring philosophical enterprise, it recently found applications in computer science, notably for artificial intelligence. The availability of large…

Physics and Society · Physics 2018-03-06 Maxime Gabella

Despite much progress in recent years, the vast majority of work in natural language processing (NLP) is on standard languages with many speakers. In this work, we instead focus on low-resource languages and in particular non-standardized…

Computation and Language · Computer Science 2023-04-20 Verena Blaschke , Hinrich Schütze , Barbara Plank

For many people, Wikipedia represents one of the primary sources of knowledge about foreign cultures. Yet, different Wikipedia language editions offer different descriptions of cultural practices. Unveiling diverging representations of…

Social and Information Networks · Computer Science 2015-07-14 Paul Laufer , Claudia Wagner , Fabian Flöck , Markus Strohmaier

Although the multilingual capability of LLMs offers new opportunities to overcome the language barrier, do these capabilities translate into real-life scenarios where linguistic divide and knowledge conflicts between multilingual sources…

Computation and Language · Computer Science 2025-06-26 Nikhil Sharma , Kenton Murray , Ziang Xiao

In this paper, we address the data scarcity problem in automatic data-driven glossing for low-resource languages by coordinating multiple sources of linguistic expertise. We supplement models with translations at both the token and sentence…

Computation and Language · Computer Science 2024-06-18 Changbing Yang , Garrett Nicolai , Miikka Silfverberg

Information systems usually show as a particular point of failure the vagueness between user search terms and the knowledge orders of the information space in question. Some kind of guided searching therefore becomes more and more important…

Information Retrieval · Computer Science 2014-06-02 Peter Mutschke , Andrea Scharnhorst , Christophe Guéret , Philipp Mayr , Preben Hansen , Aida Slavic

Current science communication has a number of drawbacks and bottlenecks which have been subject of discussion lately: Among others, the rising number of published articles makes it nearly impossible to get an overview of the state of the…

Digital Libraries · Computer Science 2020-11-24 Arthur Brack , Anett Hoppe , Markus Stocker , Sören Auer , Ralph Ewerth

Knowledge-based visual question answering (VQA) requires answering questions with external knowledge in addition to the content of images. One dataset that is mostly used in evaluating knowledge-based VQA is OK-VQA, but it lacks a gold…

Computation and Language · Computer Science 2021-09-10 Man Luo , Yankai Zeng , Pratyay Banerjee , Chitta Baral

Cultural Heritage texts contain rich knowledge that is difficult to query systematically due to the challenges of converting unstructured discourse into structured Knowledge Graphs (KGs). This paper introduces ATR4CH (Adaptive Text-to-RDF…

Computation and Language · Computer Science 2026-04-09 Andrea Schimmenti , Valentina Pasqual , Fabio Vitali , Marieke van Erp

This work explores fine-tuning OpenAI's Whisper automatic speech recognition (ASR) model for Amharic, a low-resource language, to improve transcription accuracy. While the foundational Whisper model struggles with Amharic due to limited…

Computing and Internet access are substantially growing markets in Southern Africa, which brings with it increasing demands for local content and tools in indigenous African languages. Since most of those languages are low-resourced,…

Computation and Language · Computer Science 2022-10-24 C. Maria Keet

Large language models hold promise for addressing medical challenges, such as medical diagnosis reasoning, research knowledge acquisition, clinical decision-making, and consumer health inquiry support. However, they often generate…

Computation and Language · Computer Science 2025-06-03 Zhe Chen , Yusheng Liao , Shuyang Jiang , Pingjie Wang , Yiqiu Guo , Yanfeng Wang , Yu Wang

Assigning relevant keywords to documents is very important for efficient retrieval, clustering and management of the documents. Especially with the web corpus deluged with digital documents, automation of this task is of prime importance.…

Information Retrieval · Computer Science 2017-06-20 Ayush Singhal , Ravindra Kasturi , Ankit Sharma , Jaideep Srivastava

RALMs (Retrieval-Augmented Language Models) broaden their knowledge scope by incorporating external textual resources. However, the multilingual nature of global knowledge necessitates RALMs to handle diverse languages, a topic that has…

Computation and Language · Computer Science 2024-10-30 Suhang Wu , Jialong Tang , Baosong Yang , Ante Wang , Kaidi Jia , Jiawei Yu , Junfeng Yao , Jinsong Su

Despite improved digital access to scholarly literature in the last decades, the fundamental principles of scholarly communication remain unchanged and continue to be largely document-based. Scholarly knowledge remains locked in…

Digital Libraries · Computer Science 2022-06-06 Mohamad Yaser Jaradeh , Allard Oelen , Manuel Prinz , Markus Stocker , Sören Auer

Most of the existing information retrieval systems are based on bag of words model and are not equipped with common world knowledge. Work has been done towards improving the efficiency of such systems by using intelligent algorithms to…

Artificial Intelligence · Computer Science 2015-03-17 Pekka Malo , Pyry Siitari , Ankur Sinha

A number of serious reasons will convince an increasing amount of researchers to store their relevant material in centers which we will call "language resource archives". They combine the duty of taking care of long-term preservation as…

Computation and Language · Computer Science 2009-01-20 Peter Wittenburg , Daan Broeder , Wolfgang Klein , Stephen Levinson , Laurent Romary

Traditional code search engines often do not perform well with natural language queries since they mostly apply keyword matching. These engines thus need carefully designed queries containing information about programming APIs for code…

Software Engineering · Computer Science 2018-07-10 Mohammad Masudur Rahman , Chanchal K. Roy , David Lo
‹ Prev 1 8 9 10 Next ›