Enhances medical code predictions for multi-morbidity patients using text classification.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Deep learning predicts readmissions from less structured data.
Improved NER in medical text with few examples.
Clinical notes are text documents that are created by clinicians for each patient encounter. They are typically accompanied by medical codes, which describe the diagnosis and treatment. Annotating these codes is labor intensive and error prone; furthermore, the connection between the codes and the text is not annotated…
Early detection of preventable diseases is important for better disease management, improved inter-ventions, and more efficient health-care resource allocation. Various machine learning approacheshave been developed to utilize information in Electronic Health Record (EHR) for this task. Majorityof previous attempts, ho…
The majority of medical documents and electronic health records (EHRs) are in text format that poses a challenge for data processing and finding relevant documents. Looking for ways to automatically retrieve the enormous amount of health and medical knowledge has always been an intriguing topic. Powerful methods have b…
Paper uses NLP to cluster patient visits for diagnosis validation.
System assesses patient urgency and recommends care based on medical notes.
Develops a deep multimodal classifier for medical images and reports.
This new research explores the effects of various training methods on a Polish to English Statistical Machine Translation system for medical texts. Various elements of the EMEA parallel text corpora from the OPUS project were used as the basis for training of phrase tables and language models and for development, tunin…
The quality of machine translation is rapidly evolving. Today one can find several machine translation systems on the web that provide reasonable translations, although the systems are not perfect. In some specific domains, the quality may decrease. A recently proposed approach to this domain is neural machine translat…
Improves medication name inference for telemedicine and conversational agents.
Framework extracts symptoms from EHRs for rapid disease outbreak detection.
Cross-modal data programming speeds medical machine learning.
Predicting diagnoses from Electronic Health Records (EHRs) is an important medical application of multi-label learning. We propose a convolutional residual model for multi-label classification from doctor notes in EHR data. A given patient may have multiple diagnoses, and therefore multi-label learning is required. We …
Medical applications challenge today's text categorization techniques by demanding both high accuracy and ease-of-interpretation. Although deep learning has provided a leap ahead in accuracy, this leap comes at the sacrifice of interpretability. To address this accuracy-interpretability challenge, we here introduce, fo…
We develop a model using deep learning techniques and natural language processing on unstructured text from medical records to predict hospital-wide -day unplanned readmission, with c-statistic . Our model is constructed to allow physicians to interpret the significant features for prediction.
Word embeddings are a popular approach to unsupervised learning of word relationships that are widely used in natural language processing. In this article, we present a new set of embeddings for medical concepts learned using an extremely large collection of multimodal medical data. Leaning on recent theoretical insigh…
Research benchmarks LLMs in medical domain to reduce hallucinations.
We present PubMed 200k RCT, a new dataset based on PubMed for sequential sentence classification. The dataset consists of approximately 200,000 abstracts of randomized controlled trials, totaling 2.3 million sentences. Each sentence of each abstract is labeled with their role in the abstract using one of the following …
New method reduces costs in text classification with partial labels.
MedCAT extracts valuable medical information from unstructured text.
Survey on understanding neural networks for medical applications.
Automated generation of medical reports from chest x-rays using expert annotations.
Hybrid AI and rule-based framework de-identifies medical imaging data.
Deep learning predicts ICU mortality with enhanced interpretability.
The paper presents a systematic review of state-of-the-art approaches to identify patient cohorts using electronic health records. It gives a comprehensive overview of the most commonly de-tected phenotypes and its underlying data sets. Special attention is given to preprocessing of in-put data and the different modeli…
Study uses multimodal machine learning to predict ICD-10 codes.
In this work, we present the Grounded Recurrent Neural Network (GRNN), a recurrent neural network architecture for multi-label prediction which explicitly ties labels to specific dimensions of the recurrent hidden state (we call this process "grounding"). The approach is particularly well-suited for extracting large nu…
Enhances machine learning with background knowledge through feature generation.
Motivated by the need to automate medical information extraction from free-text radiological reports, we present a bi-directional long short-term memory (BiLSTM) neural network architecture for modelling radiological language. The model has been used to address two NLP tasks: medical named-entity recognition (NER) and …
Enhanced word embedding creates new consumer-friendly health terms.
Improved negation detection in Dutch clinical texts using machine learning.
LLmFPCA-detect detects anomalies in sparse longitudinal text data using LLMs and mFPCA.
Improved named entity recognition in EHRs with transfer learning.
Proposes BONMI for integrating noisy matrices from multi-source data.
Deep learning models improved sentence similarity in medical records.
The digitalization of stored information in hospitals now allows for the exploitation of medical data in text format, as electronic health records (EHRs), initially gathered for other purposes than epidemiology. Manual search and analysis operations on such data become tedious. In recent years, the use of natural langu…
New method predicts drug interactions from drug images.
Interventional cancer clinical trials are generally too restrictive, and some patients are often excluded on the basis of comorbidity, past or concomitant treatments, or the fact that they are over a certain age. The efficacy and safety of new treatments for patients with these characteristics are, therefore, not defin…
BERT-XML automates ICD coding from EHR notes using BERT pretraining.
Artificial intelligence (AI) generally and machine learning (ML) specifically demonstrate impressive practical success in many different application domains, e.g. in autonomous driving, speech recognition, or recommender systems. Deep learning approaches, trained on extremely large data sets or using reinforcement lear…
As deep neural networks continue to revolutionize various application domains, there is increasing interest in making these powerful models more understandable and interpretable, and narrowing down the causes of good and bad predictions. We focus on recurrent neural networks, state of the art models in speech recogniti…
Many predictive tasks, such as diagnosing a patient based on their medical chart, are ultimately defined by the decisions of human experts. Unfortunately, encoding experts' knowledge is often time consuming and expensive. We propose a simple way to use fuzzy and informal knowledge from experts to guide discovery of int…
Medical deconfounder uses EHRs to estimate treatment effects without confounders.
Two large medical dialogue datasets for improving healthcare.
Proposes guidelines for developing medical AI products.
Study uses social media analytics to identify exercise-related topics.