Large language models learn company embeddings from SEC filings.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Neural model learns company embeddings from data and news.
New machine learning method classifies companies effectively.
In the context of the current financial crisis, when more companies are facing bankruptcy or insolvency, the paper aims to find methods to identify distressed firms by using financial ratios. The study will focus on identifying a group of Romanian listed companies, for which financial data for the year 2008 were availa…
AI agent predicts industry and product/service codes for companies.
CNN predicts stock fluctuations using company news headlines.
In this paper we consider an information theoretic approach for the accounting classification process. We propose a matrix formalism and an algorithm for calculations of information theoretic measures associated to accounting classification. The formalism may be useful for further generalizations and computer-based imp…
Study classifies liability insurance policies using machine learning.
This paper proposes a use of an ordinal classifier to evaluate the financial solidity of non-life insurance companies as strong, moderate, weak, and insolvency. This study constructed an efficient classification model that can be used by regulators to evaluate the financial solidity and to determine the priority of fur…
Improved default prediction for mid-cap companies using transformer models.
StonkBERT predicts stock price movements using company text data.
As the number of publicly traded companies as well as the amount of their financial data grows rapidly, it is highly desired to have tracking, analysis, and eventually stock selections automated. There have been few works focusing on estimating the stock prices of individual companies. However, many of those have worke…
Research proposes classifiers to distinguish tweets with conflicting cashtags.
FinBERT-XRC model assesses financial report risk, offering transparent explanations.
This paper introduces a Decision Tree Learner as an early warning system for classification of the non-life insurance companies according to their financial solid as strong, moderate, weak, or insolvency. In this study, we ran several experiments to show that the proposed model can achieve a good result using standard …
Text classification systems will help to solve the text clustering problem in the Azerbaijani language. There are some text-classification applications for foreign languages, but we tried to build a newly developed system to solve this problem for the Azerbaijani language. Firstly, we tried to find out potential practi…
This research develops a dynamic risk management system for industrial companies.
Study finds similar companies in Dhaka Stock Exchange using technical data.
Developed Merton's model for public companies using observed liabilities.
Company2Vec creates embeddings from company websites for fine-grained business analytics.
Develops Merton's model for private companies using DDM.
This paper develops a valuation model for private companies.
This thesis identifies share buybacks and predicts their impact on stock performance.
New model uses financial filings to predict bankruptcy, even without MDA sections.
We present an analytical study of an insurance company. We model the company's performance on a statistical basis and evaluate the predicted annual income of the company in terms of insurance parameters namely the premium, total number of the insured, average loss claims etc. We restrict ourselves to a single insurance…
This paper evaluates financial competitiveness of Indian real estate companies using entropy method.
Paper uses time series transformers to predict investment success.
The study predicts bankruptcy in Indian companies using financial ratios.
Employing profits data of Japanese companies in 2002 and 2003, we confirm that Pareto's law and the Pareto index are derived from the law of detailed balance and Gibrat's law. The last two laws are observed beyond the region where Pareto's law holds. By classifying companies into job categories, we find that companies …
Companies may be achieving only a third of the value they could be getting from data science in industry applications. In this paper, we propose a methodology for categorizing and answering 'The Big Three' questions (what is going on, what is causing it, and what actions can I take that will optimize what I care about)…
The paper analyzes how news sentiment of companies can affect market movements.
In a stock market, the price fluctuations are interactive, that is, one listed company can influence others. In this paper, we seek to study the influence relationships among listed companies by constructing a directed network on the basis of Chinese stock market. This influence network shows distinct topological prope…
The model is aimed to discriminate the 'good' and the 'bad' companies in Russian corporate sector based on their financial statements data based on Russian Accounting Standards. The data sample consists of 126 Russian public companies- issuers of Ruble bonds which represent about 36% of total number of corporate bonds …
In this study we consider relations between companies in Poland taking into account common branches they belong to. It is clear that companies belonging to the same branch compete for similar customers, so the market induces correlations between them. On the other hand two branches can be related by companies acting in…
A new method to value IPOed companies.
Study compares sentiment spillover networks from news and social media in tech companies.
Audit fees change based on company and economic factors during auditor switching.
This paper focuses on a comparative evaluation of the most common and modern methods for text classification, including the recent deep learning strategies and ensemble methods. The study is motivated by a challenging real data problem, characterized by high-dimensional and extremely sparse data, deriving from incoming…
We consider the problem of evaluating the quality of startup companies. This can be quite challenging due to the rarity of successful startup companies and the complexity of factors which impact such success. In this work we collect data on tens of thousands of startup companies, their performance, the backgrounds of t…
Model predicts default risk based on company's financial forecasts and credit conditions.
The real estate is a pillar industry of China's national economy. Due to changes in policy and market conditions, the real estate companies are facing greater pressures to survive in a competitive environment. They must improve their financial competitiveness. Based on the conceptual framework of financial competitiven…
To understand the relationship between news sentiment and company stock price movements, and to better understand connectivity among companies, we define an algorithm for measuring sentiment-based network risk. The algorithm ranks companies in networks of co-occurrences, and measures sentiment-based risk, by calculatin…
Modelling all possible life cycles of a company in a highly competitive economic environment gives a significant advantage to the owner in his business investment activities. This article proposes and analyses a dynamic model of a company's life cycle with known action costs and transition probabilities, that can be af…
Deep learning models outperform classical methods in forecasting company fundamentals.
Study shows activist board representation improves Japanese companies' performance.
We first estimate the average growth of a company's annual income and its variance by using both real company data and a numerical model which we already introduced a couple of years ago. Investment strategies expecting for income growth is evaluated based on the numerical model. Our numerical simulation suggests the p…
Customer churn is a major problem and one of the most important concerns for large companies. Due to the direct effect on the revenues of the companies, especially in the telecom field, companies are seeking to develop means to predict potential customer to churn. Therefore, finding factors that increase customer churn…
A pairwise clustering approach is applied to the analysis of the Dow Jones index companies, in order to identify similar temporal behavior of the traded stock prices. To this end, the chaotic map clustering algorithm is used, where a map is associated to each company and the correlation coefficients of the financial ti…