This document describes an approach to the problem of predicting dangerous seismic events in active coal mines up to 8 hours in advance. It was developed as a part of the AAIA'16 Data Mining Challenge: Predicting Dangerous Seismic Events in Active Coal Mines. The solutions presented consist of ensembles of various pred…
New benchmark and COAL model tackle class-imbalanced domain adaptation.
problem Aligning feature and label distributions across domains with different label distributions.
method COAL model combining feature and label distribution alignment.
result COAL model outperforms recent domain adaptation methods.
For researching the association between coal enterprise management and return in financial market, this paper applies the method of time difference relevance and PageRank method to seek the leader-index of a stock set containing 21 coal enterprises in A-share market and score those stocks. Based on the return in 2011, …
We design an active learning algorithm for cost-sensitive multiclass classification: problems where different errors have different costs. Our algorithm, COAL, makes predictions by regressing to each label's cost and predicting the smallest. On a new example, it uses a set of regressors that perform well on past data t…
How does dynamic price information flow among Northern European electricity spot prices and prices of major electricity generation fuel sources? We use time series models combined with new advances in causal inference to answer these questions. Applying our methods to weekly Nordic and German electricity prices, and oi…
Study reveals trade dynamics in dry bulk shipping networks, highlighting their randomness and periodic changes.
problem Understanding the randomness and periodic changes in dry bulk shipping networks.
method Analysis of micro-level trade flow data from 2015 to 2023, focusing on grain, coal, and iron ore networks.
result Dry bulk shipping networks exhibit small-world phenomena and periodic life cycles, influenced by importing ports and global events.
Proposes CEP to better represent financial products' carbon impact.
problem Binary 'Green' label inadequately represents financial products' carbon impact.
method Introduces Carbon Equivalence Principle (CEP) for financial products.
result Financial products' carbon impact can be included as a linked term sheet.
We present the first fully variational Bayesian inference scheme for continuous Gaussian-process-modulated Poisson processes. Such point processes are used in a variety of domains, including neuroscience, geo-statistics and astronomy, but their use is hindered by the computational cost of existing inference schemes. Ou…
Model predicts volatility and dependencies in EUA and energy prices.
problem Analyzing uncertainty and dependencies in European carbon and energy prices.
method Probabilistic multivariate conditional time series model with VECM-Copula-GARCH structure.
result Forecasting performance evaluated in an extensive rolling-window study.
One major hurdle in the road toward a low carbon economy is the present entanglement of developed economies with oil. This tight relationship is mirrored in the correlation between most of economic indicators with oil price. This paper addresses the role of oil compared to the other three main energy commodities -coal,…
Digital Rock Imaging is constrained by detector hardware, and a trade-off between the image field of view (FOV) and the image resolution must be made. This can be compensated for with super resolution (SR) techniques that take a wide FOV, low resolution (LR) image, and super resolve a high resolution (HR), high FOV ima…
Water pollution is a major global environmental problem, and it poses a great environmental risk to public health and biological diversity. This work is motivated by assessing the potential environmental threat of coal mining through increased sulfate concentrations in river networks, which do not belong to any simple …
Proposes using DII to identify non-linear causal relationships in EU Allowances returns.
problem Identifying causal relationships in non-linear data of EU Allowances returns.
method Uses Differentiable Information Imbalance (DII) for non-parametric causal discovery compared to multivariate Granger causality.
result Significant overlap and differences in causal variables identified by linear and non-linear methods.
Three years ago we found a statistically reliable link between ConocoPhillips' (NYSE: COP) stock price and the difference between the core and headline CPI in the United States. In this article, the original relationship is revisited with new data available since 2009. The agreement between the observed monthly closing…
Model predicts EU carbon prices using market and political factors.
problem Predict future carbon prices for EU market management.
method Support vector regression with grid search and cross validation.
result Model predicts carbon prices accurately for 2030.
Study optimal timing to divest from assets with uncertain future scenarios.
problem Optimal timing to divest from assets with uncertain future scenarios.
method Smooth model of decision making under ambiguity aversion, optimal stopping problem with learning.
result Proves a minimax result reducing the problem to standard optimal stopping problems with learning.
The study models and forecasts natural gas prices using skewed, heavy-tailed distributions.
problem Modeling and forecasting natural gas prices with heavy tails and conditional heteroscedasticity.
method State-space time series models under skewed, heavy-tailed distributions.
result The proposed model reduces out-of-sample CRPS by 13% for Day-Ahead and 9% for Month-Ahead forecasts.
Paper predicts ecological footprint using energy parameters.
problem Forecasting the ecological footprint using energy parameters.
method Time series vector autoregression model.
result Predictions indicate increasing consumption and declining coal energy.
Deep learning improves weather modeling for electricity load forecasting.
problem Accurate load and renewable energy forecasting requires complex spatio-temporal weather modeling.
method Automated spatio-temporal feature extraction using deep neural networks.
result Deep learning outperforms traditional methods in French national load forecasting.
Model forecasts natural gas consumption with Fourier series and feedback.
problem Forecast natural gas consumption for risk minimization.
method Modulated Fourier series with temperature deviations, day-ahead feedback.
result Model outperforms time series methods for long-term projections.
In the first section, this report analyses Nuclear Power Plants (NPPs) in the context of megaprojects, explaining why they are often delivered over budget and late. In the second section, the report discusses how Small Modular Reactors (SMRs) might address these issues. Megaprojects are extremely risky and often implem…
The paper introduces BCART models for aggregate claim amount, improving frequency-severity and joint modeling.
problem Modeling aggregate claim amount with frequency-severity and joint dependencies.
method Developed three types of BCART models: frequency-severity, sequential, and joint models. Used various distributions for claim severity data.
result Weibull distribution outperforms gamma and lognormal for right-skewed, heavy-tailed claim severity data.
The paper uses model-based trees to create interpretable surrogate models for complex machine learning models.
problem Interpreting complex machine learning models.
method Using model-based trees to partition feature space and create interpretable models.
result Model-based trees generate optimal surrogate models that balance interpretability and performance.
Gauge Flow Models use a learnable Gauge Field in Generative Flow Models.
problem Improving generative model performance.
method Integrates a learnable Gauge Field into Flow ODEs.
result Gauge Flow Models outperform traditional Flow Models in Flow Matching experiments.
The study examines how model predictions hold up under model extensions.
problem Model predictions may not be robust under model extensions, limiting their applicability.
method The study uses causal ordering to assess robustness of qualitative model predictions and characterizes model extensions that preserve predictions.
result Conditions and techniques are provided to assess robustness of model predictions under model extensions.
Revises Bayesian model averaging for foundation models.
problem Ensemble pre-trained and lightly-finetuned foundation models for improved classification performance.
method Introduces trainable linear classifiers and computationally cheaper model averaging scheme (OMA).
result Ensembled models can better predict on various datasets.
Paper introduces symmetric divergence link models for probability distributions.
problem Symmetric divergence measures for probability distributions.
method Two general classes of link models: one for survival functions and another for cumulative probability distribution functions.
result Advantages of symmetric divergence measures over asymmetric measures for model averaging and feature assessment.
New method to handle credit portfolio model uncertainties.
problem Model risk in credit portfolio models.
method Demonstrates comprehensive yet easy-to-implement approach to uncertainty in model parameters.
result Comprehensive method to deal with model uncertainties.
The paper tests stock return models and uses LSTM to predict stock returns.
problem Validating stock return models and predicting stock returns.
method Used Fama-French three-factor, four-factor, and five-factor models; also used LSTM model.
result Fama-French five-factor model shows better validity for stock returns.
Researchers review challenges in interpreting additive models, especially neural additive models.
problem Challenges in interpreting additive models, particularly neural additive models.
method Review of generalized additive models and discussion of nonidentifiability.
result Challenges in claiming interpretability or suitability for safety-critical applications of additive models.
Proposes a decision-theoretic approach for enhancing model interpretability in Bayesian frameworks.
problem Challenges the traditional approach of restricting model structure for interpretability in Bayesian frameworks.
method Introduces an interpretability utility function and a two-step method involving a reference model and a proxy model.
result Demonstrates that the proposed method generates more accurate models with the same level of interpretability.
Novel hybrid modeling combines ML and physics for real-time diagnosis.
problem Real-time diagnosis of complex systems.
method Combines machine learning and physics-based models to create reduced-order models.
result Generated models are two orders of magnitude simpler, improving efficiency.
CRS model improves ranking data modeling with theoretical guarantees.
problem Lack of rich, multimodal models for ranking data.
method Contextual Repeated Selection (CRS) model for multimodal ranking data.
result CRS model significantly outperforms existing methods in various ranking contexts.
Sigma models linked to Gross-Neveu models via quiver varieties.
problem Understanding the relationship between sigma models and Gross-Neveu models.
method Exploring the mathematical correspondence between sigma models and Gross-Neveu models, including their geometric and trigonometric/elliptic deformations.
result Sigma models are mathematically equivalent to Gross-Neveu models under certain conditions.
Interpretable machine learning has become a strong competitor for traditional black-box models. However, the possible loss of the predictive performance for gaining interpretability is often inevitable, putting practitioners in a dilemma of choosing between high accuracy (black-box models) and interpretability (interpr…
Simple models are preferred over complex models, but over-simplistic models could lead to erroneous interpretations. The classical approach is to start with a simple model, whose shortcomings are assessed in residual-based model diagnostics. Eventually, one increases the complexity of this initial overly simple model a…
Matryoshka hides secret models in a carrier model, achieving high capacity and robustness.
problem Stealing functionality of private ML data by hiding models in a carrier model.
method Parameter sharing approach exploiting the learning capacity of the carrier model.
result Hides a 26x larger secret model or 8 secret models in the carrier model.
Eigen-stratified models reduce model size and improve performance.
problem Large model size in Laplacian-regularized stratified models.
method Formulate eigen-stratified models with linear combinations of bottom eigenvectors of the graph Laplacian.
result Significant reduction in model size with eigen-stratified models.
Seq2Seq models speed up epidemic model predictions.
problem Complex epidemic models are computationally expensive.
method Used deep seq2seq models as surrogates for complex models.
result Surrogates predict scenarios up to several thousand times faster.
This work develops scalable model selection methods with fast update and selection.
problem Efficient model selection for large pools of candidate models.
method Isolated model embedding, which supports asymptotically fast update and selection.
result Standardized Embedder achieves competitive model selection performances.
Paper proposes BMPO to optimize policies using bidirectional models.
problem Model-based reinforcement learning's reliance on forward model accuracy.
method Develops BMPO using both forward and backward models for policy optimization.
result BMPO outperforms state-of-the-art methods in sample efficiency and asymptotic performance.
Copulas outperform marginal models in multivariate risk forecasting, reducing model risk by narrowing down the set of models.
problem Model risk in multivariate risk forecasting, especially during crises.
method Comprehensive empirical study comparing Copula-GARCH models with fixed marginals, copulas, or neither.
result Model risk is almost entirely due to copula choice, not marginal models.
This paper distills a complex travel mode choice model into simpler, interpretable models.
problem Lack of interpretability in complex machine learning models for travel behavior.
method Model distillation combined with market segmentation.
result Generated interpretable models that closely match the predictions of the original complex model.
BayesBlend blends multiple models' predictions for better insurance loss predictions.
problem Improving insurance loss predictions by combining multiple models.
method Pseudo-Bayesian model averaging, stacking, and hierarchical stacking.
result BayesBlend provides a user-friendly way to blend model predictions and estimate weights.
The paper identifies when larger models improve predictions and proposes a switcher model.
problem Understanding when larger models benefit from added complexity.
method Numerical studies on T5 architecture to analyze predictive uncertainty and model performance.
result Large models improve on examples where small models are uncertain, but not on certain examples.
Aggregates models from different datasets using shared latent structures.
problem Aggregating models from heterogeneous datasets with shared latent structures.
method Bayesian nonparametrics for identifying correspondences among local model parameterizations.
result Framework successfully aggregates various model types across different applications.
Improved diffusion model generation speed with speculative sampling.
problem Generating samples from computationally expensive diffusion models.
method Extending speculative sampling to diffusion models, using fast draft models for candidate token generation.
result Significant speedup in generation, halving the number of function evaluations.
DBNs improve accuracy of biological ODE models with missing data.
problem Uncertainty in biological ODE models with missing data.
method Converted ODE models to DBNs and used Particle Filtering for parameter estimation.
result DBNs can accurately infer model variables with missing data.