Causalfe estimates treatment effects in panel data with fixed effects.
problem Spurious heterogeneity in treatment effect estimates due to fixed effects in panel data.
method CFFE approach with node-level residualization during tree construction.
result Validates the estimator's performance through simulation studies.
Develops DML for nonlinear panel data models with fixed effects.
problem Estimating causal effects in nonlinear panel data models with fixed effects.
method Double machine learning (DML) procedures for approximating nuisance functions.
result First-differencing yields the least constraints on fixed effects distribution.
R package xtdml uses DML for panel data models with fixed effects.
problem Estimating structural parameters in panel data models with fixed effects.
method Combines machine learning with statistical estimation for inference.
result Demonstrates improved performance in learning nuisance functions.
A new estimator reduces bias and improves efficiency for staggered adoption studies.
problem Bias in difference-in-differences estimates for staggered adoption studies.
method Fused Extended Two-Way Fixed Effects (FETWFE) estimator with automatic parameter selection.
result FETWFE identifies correct restrictions with probability tending to one, improving efficiency.
MOMENT selects and estimates mixed-effects models using moment identities.
problem Selecting and estimating random-effects covariance matrix and fixed-effects coefficients in multiresponse linear mixed-effects models.
method MOMENT is a stage-wise moment-based framework that reduces the random-effects selection problem to a smooth constrained convex optimization problem.
result MOMENT performs competitively and can outperform separate univariate analyses for correlated responses.
Gradient boosting algorithm for spatial panel models improves estimation in high-dimensional settings.
problem Estimation failure in high-dimensional spatial panel models.
method Model-based gradient boosting algorithm for spatial panel models with random and fixed effects.
result Feasibility and interpretability in both low- and high-dimensional settings.
Study examines impact of capital structure on Indian auto companies' profitability.
problem Understanding the impact of capital structure on profitability of Indian auto companies.
method Used fixed and random effect models with 10 years of data from 17 companies.
result Optimal capital structure improves company performance and maintains capital adequacy.
We consider the problem of learning predictive models from longitudinal data, consisting of irregularly repeated, sparse observations from a set of individuals over time. Such data often exhibit {\em longitudinal correlation} (LC) (correlations among observations for each individual over time), {\em cluster correlation…
The aim of this study is to investigate quantitatively whether share prices deviated from company fundamentals in the stock market crash of 2008. For this purpose, we use a large database containing the balance sheets and share prices of 7,796 worldwide companies for the period 2004 through 2013. We develop a panel reg…
Bayesian ARMA model with directional shifts captures structural breaks in compositional time series.
problem Structural breaks in compositional time series due to external shocks or policy changes.
method Developed a Bayesian Dirichlet ARMA model augmented with a directional-shift intervention mechanism.
result The model captures structural breaks through interpretable parameters and produces coherent probabilistic forecasts.
This paper concerns the development of an inferential framework for high-dimensional linear mixed effect models. These are suitable models, for instance, when we have n repeated measurements for M subjects. We consider a scenario where the number of fixed effects p is large (and may be larger than M), but the n…
GBMixed boosts mixed models for clustered data, estimating mean and variance flexibly.
problem Flexible estimation of mean and variance components in clustered data.
method Gradient Boosting framework for linear mixed models with likelihood-based gradients.
result GBMixed accurately recovers complex nonlinear fixed effects and covariances.
gKRLS accelerates KRLS estimation for complex models.
problem Limited flexibility and high computation for KRLS.
method Re-formulate KRLS as a hierarchical model and implement random sketching.
result gKRLS can fit models on large datasets in minutes.
Little is known about how different types of advertising affect brand attitudes. We investigate the relationships between three brand attitude variables (perceived quality, perceived value and recent satisfaction) and three types of advertising (national traditional, local traditional and digital). The data represent t…
Many scientific and engineering challenges -- ranging from pharmacokinetic drug dosage allocation and personalized medicine to marketing mix (4Ps) recommendations -- require an understanding of the unobserved heterogeneity in order to develop the best decision making-processes. In this paper, we develop a hypothesis te…
In this paper we investigate panel regression models with interactive fixed effects. We propose two new estimation methods that are based on minimizing convex objective functions. The first method minimizes the sum of squared residuals with a nuclear (trace) norm regularization. The second method minimizes the nuclear …
The paper discusses fairness in bank stress tests, comparing various methods to address institutional differences.
problem Fair aggregation of bank-specific stress test models into a common model.
method Comparing various notions of regression fairness, including estimating and discarding centered bank fixed effects.
result The method of estimating and discarding centered bank fixed effects is preferable for linear models, improving forecast accuracy and equal treatment.
A new method combines machine learning with mixed-effects models for better repeated measurement analysis.
problem Inference of linear coefficients in partially linear mixed-effects models with complex interactions and high-dimensional variables.
method Double machine learning approach to estimate nonparametrically nonlinear variables, then use standard linear mixed-effects techniques to estimate the linear coefficient.
result The estimated fixed effects coefficient converges at the parametric rate and is semiparametrically efficient.
Paper speeds up Gaussian process inference using Matérn kernels.
problem Efficiently performing Gaussian process inference for large datasets.
method Exact Matérn kernel decomposition into empirical cumulative distribution functions, combined with divide-and-conquer approach.
result The proposed algorithm significantly speeds up Gaussian process inference for low-dimensional problems with hundreds of thousands of data points.
Paper proposes CIV estimator for categorical instruments in small sample settings.
problem Estimation with categorical instruments in settings with few observations per category.
method CIV estimator leveraging regularization assumption for latent categorical variable.
result CIV estimator is asymptotically normal, efficient, and semiparametrically efficient under homoskedasticity.
Study on CEF discount in Bangladesh, finds size and maturity impact, turnover negative.
problem Exploring the discount puzzle in closed-end mutual funds in Bangladesh.
method Fixed effects panel regression with diagnostic tests.
result Fund size and maturity positively impact CEF discount, turnover negatively impacts.
Study analyzes factors affecting capital adequacy in Bangladesh's banks.
problem Factors influencing capital adequacy in commercial banks in Bangladesh.
method Fixed Effect, Random Effect, and Pooled Ordinary Least Square (POLS) methods.
result Several independent variables significantly affect capital adequacy, with specific relationships between leverage, liquidity risk, and other factors.
Digital transformation boosts corporate financial asset allocation, especially short-term.
problem Understanding how digital transformation affects corporate financial decisions.
method Fixed-effects models and staggered DID design using A-share listed companies data.
result Digital transformation significantly promotes corporate financial asset allocation, more pronounced in short-term.
Study reveals which startup valuation factors are most critical.
problem Understanding the complex factors influencing startup valuations.
method Hierarchical prediction models using decision trees and random forests.
result Identifies which factors most significantly impact startup valuations.
Study finds carbon emissions affect stock value, but not bought emissions.
problem Determining if carbon emissions impact stock value and whether this is due to direct or indirect emissions.
method Fixed-effects analysis with propensity score weighting to control for selection bias.
result Firms with higher Scope 1 emissions have a statistically significant positive carbon premium, but Scope 2 emissions do not.
Proposes a method to estimate policy values in reinforcement learning with unmeasured confounders.
problem Estimating policy values in reinforcement learning with unmeasured confounders.
method Develops a two-way deconfounder algorithm using a neural tensor network to learn unmeasured confounders and system dynamics.
result Consistent policy value estimation through model-based estimator.
The methodology presented provides a quantitative way to characterize investor behavior and price dynamics within a particular asset class and time period. The methodology is applied to a data set consisting of over 250,000 data points of the S&P 100 stocks during 2004-2018. Using a two-way fixed-effects model, we unco…
Overparameterized MLR fits hyper-curves, improving model robustness.
problem Improper predictors degrade model generalizability.
method Parameterizing with a scalar and monomial basis, fitting hyper-curves.
result Hyper-curve approach yields robust predictions for noisy data.
This paper investigates how to measure common market risk factors using newly proposed Panel Quantile Regression Model for Returns. By exploring the fact that volatility crosses all quantiles of the return distribution and using penalized fixed effects estimator we are able to control for otherwise unobserved heterogen…
FinTech negatively impacts Chinese banks' financial sustainability.
problem Impact of FinTech on financial sustainability of Chinese commercial banks.
method Three-stage network DEA-Malmquist model and two-way fixed effects model.
result FinTech primarily undermines financial sustainability by eroding loan efficiency and profitability.
New matrix completion method for arbitrary sampling patterns using network flows.
problem Matrix completion under arbitrary sampling patterns.
method Network flow approach to matrix completion.
result Minimax optimal estimation for individual entries.
The choice of sentence encoder architecture reflects assumptions about how a sentence's meaning is composed from its constituent words. We examine the contribution of these architectures by holding them randomly initialised and fixed, effectively treating them as as hand-crafted language priors, and evaluating the resu…
Combines boosting with Gaussian process and mixed effects models.
problem Model misspecifications and independence assumptions in boosting.
method Relaxes zero or linearity assumption in Gaussian process and mixed effects models, and independence assumption in boosting.
result Increased prediction accuracy compared to existing approaches.
With this work it is analyzed the import and export of horticultural products between Portugal and the other world countries. It is used data about Portuguese international trade of vegetables from 2006 to 2010. The data were obtained from the INE (Statistics Portugal), gently given by the AICEP (Trade & Investment Age…
Logit-link models reveal socio-temporal effects on microfinance delinquency.
problem Understanding and quantifying socio-temporal factors affecting microfinance loan delinquency.
method Developed and evaluated discrete-time logit-link models with fixed-effects and frailty extensions.
result Simple random intercept structures capture latent heterogeneity in microfinance repayment behavior.
The paper improves machine learning for heavy-tailed panel data.
problem Improving estimates for financial and economic data with fat tails.
method Sparse-group LASSO regularization and Fuk-Nagaev concentration inequality.
result Oracle inequalities for panel data estimators.
The paper explains how to predict returns based on firm characteristics.
problem Predicting returns based on firm characteristics in equilibrium models.
method Reverse-engineering equilibrium construction process with linear demands in characteristics.
result Linear expressions for returns are derived from scaled net aggregate demands and their variations.
Improved causal inference with panel data using deep learning.
problem Causal inference challenges in social science with panel data.
method Adapted N-BEATS deep neural architecture for time series forecasting.
result SyNBEATS estimator outperforms existing methods in panel data settings.
We build a simple diagnostic criterion for approximate factor structure in large cross-sectional equity datasets. Given a model for asset returns with observable factors, the criterion checks whether the error terms are weakly cross-sectionally correlated or share at least one unobservable common factor. It only requir…
Paper develops robust econometric methods for staggered adoption studies.
problem Estimation challenges in event studies with staggered adoption.
method Design-first framework with exact probability limits, diagnostics, and orthogonal score constructions.
result Uniformly valid inference under restricted violations of parallel trends.
We statistically investigate the distribution of share price and the distributions of three common financial indicators using data from approximately 8,000 companies publicly listed worldwide for the period 2004-2013. We find that the distribution of share price follows Zipf's law; that is, it can be approximated by a …
LLMs can predict CFO responses to economic surveys
problem Measuring business sentiment
method Prompting an LLM to role-play as a CFO
result LLM reproduces individual human responses
Study finds WACC negatively impacts firm profitability in Bangladesh's food industry.
problem Determining the impact of Weighted Average Cost of Capital (WACC) on firm profitability.
method Fixed Effects Panel Regression Model using 12 food and allied industry companies from 2005-2019.
result WACC negatively correlates with firm profitability (ROA), significant relationship.
Multi-task learning models using Gaussian processes (GP) have been developed and successfully applied in various applications. The main difficulty with this approach is the computational cost of inference using the union of examples from all tasks. Therefore sparse solutions, that avoid using the entire data directly a…
Study finds environmental liability insurance reduces industrial carbon emissions.
problem Reduction of industrial carbon emissions.
method Two-way fixed effect model using provincial (city) level panel data from 2010 to 2020.
result Environmental liability insurance reduces industrial carbon emissions at both direct and indirect levels, with varying effects.
This work relates the framework of model-based clustering for spatial functional data where the data are surfaces. We first introduce a Bayesian spatial spline regression model with mixed-effects (BSSR) for modeling spatial function data. The BSSR model is based on Nodal basis functions for spatial regression and accom…
ChatGPT snapshots predict future stock returns.
problem Predicting future stock returns using pre-cutoff text.
method Extracted LLM outlook scores from OpenAI snapshots.
result Outlook scores positively correlate with future stock returns.
BSA-TNP improves NP scalability and accuracy for spatiotemporal data.
problem Scalability and accuracy trade-off in Neural Processes.
method Introduces KRBlocks, group-invariant attention biases, and BSA for scalable spatiotemporal inference.
result BSA-TNP matches or exceeds accuracy of best models while training faster.