Paper reviews neurolinguistics and language technologies, emphasizing mutual enrichment.
problem Understanding brain activity during language processing.
method Brain imaging studies and natural language representations.
result Development of brain-aware natural language representations.
This review compares extractive and abstractive summarization methods.
problem Improving abstractive summarization in natural language processing.
method Compared various approaches including supervised and unsupervised methods, deep learning, and NLP.
result Current research uses combinations of approaches, but abstractive summarization remains unsolved.
Survey reviews code-switched speech and language processing.
problem Processing code-switched text and speech for multilingual communities.
method Reviews computational approaches and lists available resources.
result Essential for building intelligent agents that interact in multilingual settings.
Systematic review of electronic health record phenotyping approaches.
problem Detecting patient cohorts using electronic health records.
method Comprehensive literature review of preprocessing and modeling approaches.
result Natural language processing shows promise for electronic phenotyping.
Study evaluates human vs. machine review generation, finds human assessments correlate better with lexical overlaps.
problem Evaluating natural language generation models for online reviews is challenging and inconsistent.
method Compared human evaluators with various automated evaluation methods, including discriminative and word overlap metrics.
result Human evaluators do not correlate well with discriminative evaluators, but correlate better with lexical overlaps.
Improved collaborative filtering with neural network models of reviews.
problem Improve collaborative filtering performance using side information from reviews.
method Introduced two neural network models (product-of-experts and recurrent neural network) to incorporate reviews into collaborative filtering.
result The product-of-experts model achieved state-of-the-art performance, outperforming LDA-based approach.
This paper reviews different word embeddings for sentiment classification using deep learning.
problem Handling large textual data with simple ML algorithms.
method Word embedding strategies implemented on an Amazon Review Dataset.
result Different word embeddings improve accuracy in sentiment classification.
Algorithm improves SLR efficiency in financial narratives.
problem Fragmented understanding of financial narratives.
method NLP, clustering, interpretability tools.
result Unified narrative modeling approaches needed.
NLP techniques improve drug discovery by analyzing chemical and protein text.
problem Improving drug discovery through better analysis of chemical and protein text.
method Natural language processing techniques applied to biochemical entities.
result Enhanced prediction of molecular properties and design of novel molecules.
Deep learning models outperform classical methods in text classification.
problem Improving text classification accuracy using deep learning.
method Comprehensive review of deep learning models and datasets for text classification.
result Deep learning models outperform classical methods on various text classification tasks.
Project analyzes drug reviews to predict ratings using machine learning.
problem Predicting drug ratings from text reviews.
method Implemented supervised machine learning algorithms with TFIDF and Count Vectors.
result Good results in predicting test data sets for popular conditions.
Workshop reviews techniques to understand neural NLP models.
problem Understanding the inner workings of neural networks in natural language processing.
method Systematic manipulation of inputs, decoding intermediate representations, modifying architectures, and testing on simplified languages.
result Various techniques can improve explainability of neural network models.
This paper reviews LLMs for credit risk assessment, creating a taxonomy.
problem Assessing credit risk using financial text analysis.
method Systematic review of 60 papers, focusing on model architectures, data types, and explainability mechanisms.
result Developed a taxonomy of LLM-based credit risk models.
LR-Robot automates SLRs with AI, expert oversight, and multidimensional analysis.
problem Efficient but contextually limited outputs from existing SLR frameworks.
method Human-in-the-loop process, structured knowledge sources, retrieval-augmented generation.
result Empirical demonstration of AI-driven literature synthesis in option pricing.
Vision and language tasks often fail to test AI comprehensively.
problem Current vision and language tasks are flawed due to dataset and evaluation issues.
method Review of current state and proposal for improvement.
result State-of-the-art systems perform well due to dataset and evaluation flaws.
This paper reviews and proposes a unified framework for contrastive learning.
problem The origins and development of contrastive learning across various fields.
method A comprehensive literature review and a general Contrastive Representation Learning framework.
result A unified framework simplifies and unifies contrastive learning methods.
ZeroSCROLLS benchmarks zero-shot natural language understanding over long texts.
problem Evaluate natural language understanding models over long texts without training data.
method Adapt six tasks from SCROLLS benchmark and add four new datasets, including novel aggregation tasks.
result Claude outperforms ChatGPT, and GPT-4 achieves highest average score.
This paper reviews sentiment analysis on Indian languages.
problem Understanding sentiment in multilingual web data.
method Reviews and discusses approaches for sentiment analysis on Indian languages.
result Challenges in sentiment analysis on indigenous languages.
This paper reviews GANs algorithms, theory, and applications.
problem Lack of comprehensive study on GANs connections and evolution.
method Detailed introduction of GANs algorithms, theoretical investigation, and application illustrations.
result Comprehensive review of GANs from algorithms, theory, and applications perspectives.
SSMBA generates synthetic data to improve robustness in natural language tasks.
problem Improving out-of-domain generalization of models trained on natural language data.
method SSMBA uses corruption and reconstruction functions to generate synthetic data points on the manifold assumption.
result SSMBA consistently outperforms existing methods on robustness benchmarks across multiple tasks and datasets.
Natural language processing improves COVID-19 hospitalization identification.
problem Identifying patients hospitalized due to COVID-19 among those with positive SARS-CoV-2 tests.
method Used natural language processing on provider notes and structured EHR data elements to create classification algorithms.
result Classification algorithms using provider notes outperformed those using only structured EHR data elements, with AUROC of 0.894 compared to 0.841.
Predicts Yelp star reviews using deep learning and network structure.
problem Predicting Yelp star reviews based on network structure and features.
method Compared multiple models including deep learning on network and item features.
result Deep learning models combining node-level and network features outperform others.
LR-Robot accelerates SLRs by combining expert oversight and AI, revealing trends and patterns in financial research.
problem Manual SLRs are impractical due to the scale and complexity of modern financial research.
method Domain experts define taxonomies and constraints, LLMs execute classification, and human evaluation ensures reliability.
result AI can understand and synthesize literature, revealing trends and core research directions.
Paper proposes a method to improve language model performance on unknown distributions.
problem Language models trained on diverse data can perform poorly on unseen distributions.
method Distributionally robust optimization (DRO) to minimize worst-case performance over a mixture of potential test distributions.
result Topic CVaR approach reduces perplexity by 5.5 points compared to standard maximum likelihood.
Survey of deep learning methods for image captioning.
problem Generating accurate and complex image descriptions.
method Comprehensive review of deep learning techniques for image captioning.
result Analysis of strengths, limitations, and popular datasets in deep learning image captioning.
Survey of LLMs in finance tasks, including adoption and performance.
problem Utilizing large language models in financial tasks.
method Review of current approaches, decision framework for adoption.
result Synthesizes state-of-the-art for LLMs in finance.
Deep reinforcement learning combines deep learning with reinforcement learning for complex tasks.
problem Complex tasks with high-dimensional data.
method Combining deep learning architectures (autoencoders, CNN, RNN) with reinforcement learning.
result Successful learning of useful representations for high-dimensional data.
Paper finds mislabeled instances in various datasets.
problem Mislabeled instances in labeled datasets.
method Non-parametric end-to-end pipeline for finding mislabeled instances.
result Average precision of more than 0.84 for finding top 1% mislabeled instances.
This paper reviews methods for interpreting deep learning models with sequential data.
problem Limited interpretability of deep learning models in sequential data domains.
method Reviews and compares techniques for sequential interpretability.
result Current techniques have limitations and future research is needed.
Paper uses pre-trained models and active learning to analyze customer reviews quickly.
problem Automatic review analysis with limited labeled data and time.
method Pre-trained language representation and active learning framework.
result Fully automatic review analysis achieved at a faster pace.
UQE uses LLMs to analyze unstructured data efficiently.
problem Efficient analytics on unstructured data.
method Proposes UQE, a query engine that uses LLMs to interpret UQL queries.
result Demonstrates efficient analytics on various unstructured data types.
Paper predicts Airbnb prices using machine learning and customer reviews.
problem Predicting optimal Airbnb prices with limited property information.
method Uses machine learning, sentiment analysis, and various models.
result Develops a model to help both property owners and customers with price evaluation.
Pars-ABSA dataset for Persian aspect-based sentiment analysis.
problem Lack of public dataset for Persian aspect-based sentiment analysis.
method Manually annotated dataset with 5,114 positive, 3,061 negative, and 1,827 neutral samples.
result State-of-the-art performance of deep learning methods on Pars-ABSA compared to similar English datasets.
LLMs improve stock price forecasting from financial news and reports.
problem Predicting stock prices with high accuracy and robustness.
method Analyzing financial news, reports, and transcripts using LLMs.
result LLMs can improve stock price forecasting but face practical challenges.
Paper presents LLM-enhanced contract metadata extraction.
problem Automatic detection and annotation of legal clauses in contracts.
method Integration of publicly available and proprietary datasets with advanced LLM methodologies.
result Substantial improvements in clause identification accuracy and efficiency.
This paper analyzes text in financial disclosures to improve financial analysis.
problem Insufficient analysis of unstructured text in financial disclosures.
method Reviews and explores methods in computational linguistics and NLP.
result Highlights limitations of sentiment metrics and suggests future research areas.
This review explores ChatGPT in accounting and finance.
problem Understanding the current state of research on ChatGPT in accounting and finance.
method A scoping review of recent publications and working papers.
result Identifies three themes: applications, research tools, and implications.
LLM extracts actionable insights from customer reviews.
problem Extracting actionable insights from customer reviews.
method Large language model approach distinguishing perceptual attributes from actionable features.
result High consistency and predictive validity of LLM insights compared to human coders.
A new approach to rationalization identifies true rationales by considering causal relationships.
problem Existing rationalization methods struggle with spuriousness, where snippets with similar contributions are hard to distinguish.
method The method leverages causal inference to identify non-spurious rationales, defining probabilities of causation based on a structural causal model.
result The proposed causal rationalization outperforms existing methods on real-world datasets.
We review ideas on temporal dependences and recurrences in discrete time series from several areas of natural and social sciences. We revisit existing studies and redefine the relevant observables in the language of copulas (joint laws of the ranks). We propose that copulas provide an appropriate mathematical framework…
This paper reviews deep learning's latest progress and applications.
problem Challenges in deep learning models and applications.
method Analysis of existing models and new emerging models.
result Summarizes deep learning's applications in various AI fields.
The paper explores how to make neural networks extrapolate longer sequences.
problem Neural networks struggle to extrapolate beyond seen data, especially for longer sequences.
method The authors propose a model with separate content- and location-based attention mechanisms.
result Models with the proposed attention mechanisms are better at extrapolating longer sequences.
The abstract reviews financial concepts using physics.
problem Financial pricing and risk management.
method Discrete time formalism, path integral, Green's function formulas.
result Formulas for pricing and risk mitigation methods.
The paper uses geometry to assess how hard examples are for NLP models.
problem Challenges in NLP datasets and classifiers, especially with shallow features.
method Information geometry to quantify example difficulty, exploring BERT, CNN, and fasttext.
result Deep learning models are vulnerable to word substitutions in difficult examples.
Machine learning reduces workload in healthcare systematic reviews by 70%.
problem Efficiently screening abstracts for systematic reviews in healthcare.
method Training SVM classifiers on labeled abstracts to classify RCTs.
result SVM classifier achieved 90% accuracy and 0.84 F1 score.
Survey of self-supervised learning methods in computer vision, NLP, and graph learning.
problem Manual labeling and vulnerability to attacks in supervised learning.
method Generative, contrastive, and generative-contrastive approaches.
result Self-supervised learning achieves high performance in representation learning.
This paper reviews feature selection in KGs for improved ML model performance.
problem Improving feature selection in KGs for better machine learning model efficacy.
method Comprehensive review of feature selection methodologies in KGs.
result Advancement in scalability, accuracy, and interpretability of feature selection techniques.
Myia compiler optimizes ML models with efficient AD for array programming.
problem Efficient automatic differentiation for array programming in ML.
method Introduces a new graph-based IR that supports function calls, higher-order functions, and recursion.
result Myia compiler enables efficient AD using source transformation without a tape, supporting higher-order derivatives.