Deep learning classifies railroad accident causes from narrative reports.
problem Classifying accident causes from narrative reports is challenging.
method Applied deep learning with word embeddings to classify accident causes.
result Deep learning accurately classifies accident causes from narratives and identifies inconsistencies.
Analyst reports contain valuable information for investment decisions.
problem Investment value in analyst reports is not fully understood or utilized.
method Embedded analyst reports with LLMs and ML forecasts of future returns.
result Portfolios formed on analyst report narratives outperform numerical forecasts and established factors.
Developing an AI economist agent using RAG, knowledge graphs, and LLMs for economic scenario analysis.
problem Economic scenario analysis using large language models and knowledge graphs.
method Proposing an RAG-based AI economist framework that utilizes knowledge graphs and LLMs.
result Improves economic coherence and traceability in generated reports.
Narrative disclosures in 10-K filings improve bankruptcy prediction beyond accounting ratios.
problem Traditional bankruptcy prediction models rely on accounting ratios, which may not capture early warning signals.
method Developed a PB Stress Score based on distress-specific language in 10-K narratives, evaluated against accounting and dictionary benchmarks.
result Adding the PB Stress Score increases AUC from 0.8323 to 0.9019 and improves top-decile bankruptcy capture from 44.12% to 64.71%.
Predicting emotional state from narratives using contextual information.
problem Predicting valence from personal narratives using contextual information.
method Investigated multiple machine learning techniques to model narratives, focusing on textual information.
result Models capture inter-individual differences, leading to more accurate predictions of emotional state.
Study finds strong link between crypto narratives and prices.
problem Understanding the impact of crypto narratives on prices.
method Topic modeling of Twitter data combined with sentiment analysis.
result Strong correlation between narratives and crypto prices.
Platform combines RL and language models to study narrative influence on AI decisions.
problem Understanding how narrative elements shape AI decision-making.
method Dual-system architecture with reinforcement learning and language model integration.
result Initial experiments show narrative frameworks can influence AI decision-making.
Improved earnings predictions through text-morphed earnings calls.
problem Improving earnings prediction models using narrative information.
method Introducing a text-morphing methodology to generate counterfactual transcripts.
result Analysts over-react to sentiment and under-react to risk and uncertainty.
Algorithm improves SLR efficiency in financial narratives.
problem Fragmented understanding of financial narratives.
method NLP, clustering, interpretability tools.
result Unified narrative modeling approaches needed.
Paper fine-tunes a language model to predict long-term stock buy signals.
problem Predicting long-term stock price movements with narrative text.
method Fine-tuning a small language model on 10-K reports for buy/sell decisions.
result Buy signals generated from 10-K text are most precise at 6 and 9 months, providing 4.8-9% improvement over random selection.
A new platform models how narratives influence financial markets.
problem Explaining irrational market behaviors through narratives.
method Integrated opinion dynamics and agent-based modeling.
result Initial results show how narratives shape financial outcomes.
Data mining reveals power structures in Bangladeshi newspapers.
problem Understanding the power dynamics and narrative structure in news reporting.
method Named entity recognition to create temporal actor networks from news statements.
result Cliquishness among powerful political leaders in news articles.
We describe an exercise of using Big Data to predict the Michigan Consumer Sentiment Index, a widely used indicator of the state of confidence in the US economy. We carry out the exercise from a pure ex ante perspective. We use the methodology of algorithmic text analysis of an archive of brokers' reports over the peri…
Paper analyzes systematic jump risk around the clock using news narratives.
problem Identifying and managing priced risks in real-time market conditions.
method Combining high-frequency market data with news narratives classified by an LLM.
result Significant heterogeneity in risk premia, with macroeconomic news commanding the largest premium.
CrystalCandle creates user-friendly explanations for machine learning models.
problem Low trust in predictive models due to lack of interpretability.
method End-to-end pipeline for model interpretation, including Model Importer, Interpreter, Narrative Generator, and Exporter.
result CrystalCandle leads to higher adoption rates and improved downstream metrics.
Study finds whitepaper narratives do not predict market factor structure.
problem Predicting market behavior from cryptocurrency whitepaper claims.
method Zero-shot NLP classification combined with CP tensor decomposition of market data.
result Weak alignment between whitepaper claims and market statistics and latent factors.
Models predict emotional valence from narratives, matching human raters.
problem Predicting emotional valence from multimodal time-series data.
method Adapted attention-based mechanisms (Transformer, Memory Fusion Network) to emotional narratives.
result Models perform well, matching human raters on emotional valence prediction.
Detects systematic anomalies in consumer complaints using NLP.
problem Detecting small, frequent anomalies in consumer complaints.
method NLP conversion of narratives, followed by anomaly detection algorithm.
result Demonstrates effectiveness of NLP for detecting systematic anomalies.
Framework detects influential actors in disinformation networks.
problem Identifying and countering hostile influence operations on social media.
method Combines NLP, ML, graph analytics, and causal inference.
result 96% precision, 79% recall, 96% PR-curve area for IO detection.
Study uses LLM to extract and compare segment disclosures from financial filings.
problem Challenges in completeness and comparability of segment disclosures in financial reports.
method Developed a large language model framework to extract and preserve segment information from Form 10-K filings.
result The LLM accurately extracts segment-level information and addresses cross-period knowledge questions.
Study finds no significant alignment between whitepaper claims and market structure.
problem Correlation between cryptocurrency whitepaper narratives and market behavior.
method Developed a contamination-aware pipeline for measuring structural correspondence, combining NLP classification and market statistics.
result No significant claims-market alignment detected in the sample.
Paper uses Apprenticeship Learning to model player behavior in interactive narratives.
problem Understanding and simulating player behavior in interactive narratives.
method Receding Horizon IRL (RHIRL) to learn reward functions and policies.
result RHIRL can learn action sequences and generate behavior similar to specific players.
QRAFTI uses multi-agent framework to improve equity factor research.
problem Replicating and developing new equity factors in large financial datasets.
method Integrates a research toolkit with MCP servers for data access and custom coding operations.
result Improves performance and explainability in multi-step empirical tasks.
Study detects SLI in children from spontaneous narrative transcripts.
problem Detecting Specific Language Impairment (SLI) in children.
method Three-stage pipeline: feature extraction, dimensionality reduction, and classification.
result 97.13% accuracy in identifying SLI from transcripts.
FinTech framework clusters innovations for financial services.
problem Lack of comprehensive definition and analysis of FinTech.
method Narrative review of over 100 studies, clustering framework development.
result Developed a comprehensive FinTech clustering framework.
LLMs can help explain credit risk models but not autonomously.
problem Leveraging LLMs for post-hoc explainability in credit risk models.
method Comparison of LLM outputs with SHAP and coefficient-based attributions on three LMs.
result LLMs reliably preserve feature-importance rankings but poorly align with autonomous explanations.
Representation learning improves EHR data for healthcare tasks.
problem Transforming EHR data into useful representations for machine learning.
method Deep learning and disentangling underlying factors from EHR data.
result Better representations improve machine learning performance in healthcare.
This report aims to improve trust in AI by explaining machine learning models.
problem Understanding and trusting automated decision-making systems.
method Survey and distillation of literature on explainable machine learning.
result Survey findings help practitioners understand and apply explainable methods.
Automatically extracts phenotypes from cancer clinical notes for genetic studies.
problem Lack of structured patient representations in EHRs.
method Clustering of medical terms and sentences in clinical notes.
result 341 significant associations between clinical features and somatic mutations.
Study shows GPT's earnings forecasts are human-like but not always accurate.
problem Information friction in AI-generated financial analysis.
method Examined GPT's earnings forecasts following corporate earnings releases and proposed a diagnostic framework.
result GPT's narrative attention is consistent and human-like but not always associated with higher forecast accuracy.
Hierarchical organization is a cornerstone of complexity and multifractality constitutes its central quantifying concept. For model uniform cascades the corresponding singularity spectra are symmetric while those extracted from empirical data are often asymmetric. Using the selected time series representing such divers…
New framework tackles deep financial reporting bottleneck by improving hallucination and coherence.
problem Statistical smoothing trap in LLMs limits deep financial reporting quality.
method DeepNews Framework integrates information foraging, schema-guided planning, and adversarial prompting.
result DeepNews system achieves 25% acceptance rate in blind test, significantly outperforming SOTA.
Financial LLMs need explicit bias consideration to avoid invalid results.
problem Finance-specific biases inflate performance and contaminate backtests.
method Identified five recurring biases and proposed a Structural Validity Framework.
result Explicit bias consideration is necessary for valid deployment claims.
FinCausal 2020 task detects financial document causality.
problem Detect causal relationships in financial documents.
method Binary classification and relation extraction tasks.
result 16 teams participated, 13 submitted system descriptions.
A single BLSTM network tackles ambiguous words in text data.
problem Ambiguity in text data, especially in technical domains.
method Proposes a single Bidirectional LSTM network for all ambiguous words.
result Comparable performance to top WSD algorithms on SensEval-3 benchmark.
Analyzes news graphs to predict financial market dislocations.
problem Predicting financial market dislocations using news content.
method Extracts entities from news articles, aggregates them into graphs, applies network analysis, and uses sentiment analysis.
result Identifies high entropy in news graphs correlates with financial market dislocations.
Detects crime series using RBM embeddings from crime narratives.
problem Detecting related crime series from crime records.
method Unsupervised learning of latent feature embeddings using Gaussian-Bernoulli RBM.
result Related cases are closer in feature space, unrelated cases are far apart.
Media seems to have become more partisan, often providing a biased coverage of news catering to the interest of specific groups. It is therefore essential to identify credible information content that provides an objective narrative of an event. News communities such as digg, reddit, or newstrust offer recommendations,…
Paper develops a method for valid inference using language model predictions from verbal autopsy narratives.
problem Valid inference from verbal autopsy narratives for public health decision-making.
method Develops multiPPI++ method for valid inference using NLP techniques for COD prediction.
result Demonstrates the effectiveness of multiPPI++ in handling transportability issues and recovering ground truth estimates.
Model shows how past consumption affects household confidence, leading to varied economic outcomes.
problem Exploring how past consumption impacts current confidence and economic activity in a multi-household model.
method Developed a DSGE model where past consumption influences individual household confidence and consumption propensity.
result The model demonstrates a range of economic outcomes including high output with no crises, high output with increased volatility, and alternation of high and low output states.
LDA identifies latent topics in CFPB consumer complaints over time.
problem Identify latent topics in CFPB consumer complaints for better regulation effectiveness.
method Latent Dirichlet Allocation (LDA) for topic modeling of consumer complaints.
result Time trends of latent topics reveal regulatory effectiveness and consumer protection issues.
LLMs outperform human analysts in predicting earnings direction.
problem Evaluating financial statements without narrative or industry-specific information.
method Trained GPT4 on standardized, anonymous financial statements and instructed to predict earnings direction.
result LLMs predict earnings directionally with accuracy comparable to narrowly trained ML models.
Enhances neural language processing with a hierarchical context-aware model.
problem Limited context in neural language processing systems.
method Hierarchical recurrent neural network with multi-level context representation.
result Improves semantic error detection by 12.75% relative for unsupervised models and 20.37% relative for supervised models.
Study finds monthly SIPs outperform first-day SIPs in Nifty 50 by 0.5-2.5% annually.
problem Underexplored impact of SIP timing in India's equity market.
method 22-year analysis using multi-layered statistical framework (non-parametric tests, effect size metrics, SSD).
result Monthly SIPs (EXP-SIP) outperform first-day SIPs (FTD-SIP) by 0.5-2.5% annually over short-to-medium-term horizons.
Detects radical content on Twitter using textual, psychological, and behavioral signals.
problem Limit the spread of extremist narratives on social media.
method Analyzed extremist material, created contextual text-based model, inferred psychological properties, evaluated on Twitter.
result Radical users exhibit distinguishable textual, psychological, and behavioral properties.
Algorithm reduces audit costs by identifying best service configurations from biased textual evidence.
problem Designing service systems from textual evidence requires accurate selection despite biased automated scoring.
method Developed PP-LUCB algorithm combining LLM scores and selective audits to minimize costs.
result Correctly identified the best model in 40/40 trials with 90% cost reduction.
Framework improves clinical timeline reconstruction from text and tables.
problem Temporal precision and event timing in clinical narratives and EHRs.
method Retrieval-augmented multimodal alignment framework.
result Consistently improves absolute timestamp accuracy and temporal concordance.
The Lady Maisry ballads afford us a framework within which to segment a storyline into its major components. Segments and as a consequence nodal points are discussed for nine different variants of the Lady Maisry story of a (young) woman being burnt to death by her family, on account of her becoming pregnant by a forei…