English
Related papers

Related papers: Disease Identification From Unstructured User Inpu…

200 papers

In biomedical Subgroup Discovery, practitioners are interested in discovering interpretable and homogeneous subgroups within a group of patients. In this paper, assuming that healthy subjects (i.e., controls) share common but irrelevant…

Machine Learning · Computer Science 2026-05-21 Robin Louiset , Edouard Duchesnay , Benoit Dufumier , Antoine Grigis , Pietro Gori

The identification of rare diseases from clinical notes with Natural Language Processing (NLP) is challenging due to the few cases available for machine learning and the need of data annotation from clinical experts. We propose a method…

Computation and Language · Computer Science 2021-07-30 Hang Dong , Víctor Suárez-Paniagua , Huayu Zhang , Minhong Wang , Emma Whitfield , Honghan Wu

We present an algorithm that takes an unannotated corpus as its input, and returns a ranked list of probable morphologically related pairs as its output. The algorithm tries to discover morphologically related pairs by looking for pairs…

Computation and Language · Computer Science 2007-05-23 Marco Baroni , Johannes Matiasek , Harald Trost

This article presents a method for prompt-based mental health screening from a large and noisy dataset of social media text. Our method uses GPT 3.5. prompting to distinguish publications that may be more relevant to the task, and then uses…

Computation and Language · Computer Science 2024-05-14 Wesley Ramos dos Santos , Ivandre Paraboni

A large percentage of medical information is in unstructured text format in electronic medical record systems. Manual extraction of information from clinical notes is extremely time consuming. Natural language processing has been widely…

Information Retrieval · Computer Science 2019-08-16 Dianbo Liu , Dmitriy Dligach , Timothy Miller

Social scientists are increasingly turning to unstructured datasets to unlock new empirical insights, e.g., estimating descriptive statistics of or causal effects on quantitative measures derived from text, audio, or video data. In many…

Econometrics · Economics 2026-05-06 Jacob Carlson

We propose a categorical matrix factorization method to infer latent diseases from electronic health records (EHR) data in an unsupervised manner. A latent disease is defined as an unknown biological aberration that causes a set of common…

Applications · Statistics 2019-02-15 Yang Ni , Peter Mueller , Yuan Ji

Machine learning models are increasingly used in the medical domain to study the association between risk factors and diseases to support practitioners in predicting health outcomes. In this paper, we showcase the use of machine-learned…

With the advancement in the technology sector spanning over every field, a huge influx of information is inevitable. Among all the opportunities that the advancements in the technology have brought, one of them is to propose efficient…

Information Retrieval · Computer Science 2021-12-14 Sudhanshu , Narinder Singh Punn , Sanjay Kumar Sonbhadra , Sonali Agarwal

It is an important subject how deal with the symptom's data, input data, to improve the accuracy and efficiency of the diagnostic algorithm in the medical decision support systems. In this paper, we described a method for the numerical…

Computers and Society · Computer Science 2018-11-19 Won-Il Song , Taek-Jong Kim

On social media platforms like Twitter, users regularly share their opinions and comments with software vendors and service providers. Popular software products might get thousands of user comments per day. Research has shown that such…

Software Engineering · Computer Science 2021-08-20 Christoph Stanik , Tim Pietz , Walid Maalej

Clinical notes are an efficient way to record patient information but are notoriously hard to decipher for non-experts. Automatically simplifying medical text can empower patients with valuable information about their health, while saving…

Predicting patient mortality is an important and challenging problem in the healthcare domain, especially for intensive care unit (ICU) patients. Electronic health notes serve as a rich source for learning patient representations, that can…

Computation and Language · Computer Science 2019-10-16 Shaika Chowdhury , Chenwei Zhang , Philip S. Yu , Yuan Luo

This paper presents a predictive model for Influenza-Like-Illness, based on Twitter traffic. We gather data from Twitter based on a set of keywords used in the Influenza wikipedia page, and perform feature selection over all words used in 3…

Social and Information Networks · Computer Science 2021-11-23 Katerina Katsani-Geronymaki , Polyvios Pratikakis

Mental health poses a significant challenge for an individual's well-being. Text analysis of rich resources, like social media, can contribute to deeper understanding of illnesses and provide means for their early detection. We tackle a…

Computation and Language · Computer Science 2020-03-18 Ivan Sekulić , Michael Strube

Psychiatric illnesses are often associated with multiple symptoms, whose severity must be graded for accurate diagnosis and treatment. This grading is usually done by trained clinicians based on human observations and judgments made within…

Sound · Computer Science 2017-03-17 Rita Singh , Justin Baker , Luciana Pennant , Louis-Philippe Morency

We present a system that uses a learned autocompletion mechanism to facilitate rapid creation of semi-structured clinical documentation. We dynamically suggest relevant clinical concepts as a doctor drafts a note by leveraging features from…

Machine Learning · Computer Science 2020-07-31 Divya Gopinath , Monica Agrawal , Luke Murray , Steven Horng , David Karger , David Sontag

Causal understanding is a fundamental goal of evidence-based medicine. When randomization is impossible, causal inference methods allow the estimation of treatment effects from retrospective analysis of observational data. However, such…

Machine Learning · Computer Science 2024-11-06 Samuel Lee , Zach Wood-Doughty

Medical concept normalization helps in discovering standard concepts in free-form text i.e., maps health-related mentions to standard concepts in a vocabulary. It is much beyond simple string matching and requires a deep semantic…

Computation and Language · Computer Science 2020-06-09 Katikapalli Subramanyam Kalyan , S. Sangeetha

The majority of big data is unstructured and of this majority the largest chunk is text. While data mining techniques are well developed and standardized for structured, numerical data, the realm of unstructured data is still largely…

Artificial Intelligence · Computer Science 2015-06-24 Carlo A. Trugenberger