We propose a probabilistic model for refining coarse-grained spatial data by utilizing auxiliary spatial data sets. Existing methods require that the spatial granularities of the auxiliary data sets are the same as the desired granularity of target data. The proposed model can effectively make use of auxiliary data set…
We introduce a hierarchical architecture for video understanding that exploits the structure of real world actions by capturing targets at different levels of granularity. We design the model such that it first learns simpler coarse-grained tasks, and then moves on to learn more fine-grained targets. The model is train…
Paper presents UrbanFM and UrbanPy models for inferring fine-grained urban flows.
problem Reduce cost of urban flow monitoring while maintaining data accuracy and granularity.
method Develops UrbanFM and UrbanPy models to infer fine-grained urban flows from coarse-grained observations.
result UrbanPy model demonstrates favorable performance for larger-scale inference tasks.
Tensor network architecture for classification and regression using wavelet transformations.
problem Efficiently performing classification and regression tasks on complex data.
method Tensor network layers based on MERA and MPS, with adaptive fine-graining.
result Adaptive fine-graining improves model performance without loss in accuracy.
Sparsity helps reduce the computational complexity of deep neural networks by skipping zeros. Taking advantage of sparsity is listed as a high priority in next generation DNN accelerators such as TPU. The structure of sparsity, i.e., the granularity of pruning, affects the efficiency of hardware accelerator design as w…
Improved investment performance with fine-grained LLM tasks.
problem Abstract financial trading systems often overlook real-world workflow intricacies, leading to degraded performance.
method Proposes a multi-agent LLM trading framework that decomposes investment analysis into fine-grained tasks.
result Fine-grained task decomposition significantly improves risk-adjusted returns compared to coarse-grained designs.
Considering event structure information has proven helpful in text-based stock movement prediction. However, existing works mainly adopt the coarse-grained events, which loses the specific semantic information of diverse event types. In this work, we propose to incorporate the fine-grained events in stock movement pred…
Fine-grained pretraining improves neural network's ability to learn rare features.
problem Improving generalization in deep learning models.
method Introducing a hierarchical multi-view structure to confine input data distribution.
result Fine-grained pretraining leads to better accuracy on hard downstream test samples.
The paper develops a framework for abstracting causal models using category theory.
problem Difficulties in changing the variables used to describe a system, especially from fine-grained to coarse-grained.
method Introduces a category of interventional causal models and uses enriched category theory to prove compositionality properties.
result Compositionality of model transformations is established, with bounded errors for each step.
We discuss a Bayesian formulation to coarse-graining (CG) of PDEs where the coefficients (e.g. material parameters) exhibit random, fine scale variability. The direct solution to such problems requires grids that are small enough to resolve this fine scale variability which unavoidably requires the repeated solution of…
LDA improves image classification accuracy with fewer features.
problem Fine-grained image classification with pretrained features.
method Supervised dimensionality reduction with LDA before linear probing.
result LDA improves accuracy over full features in 11 out of 12 configurations.
The combination of high-dimensionality and disparity of time scales encountered in many problems in computational physics has motivated the development of coarse-grained (CG) models. In this paper, we advocate the paradigm of data-driven discovery for extract- ing governing equations by employing fine-scale simulation …
Model compression techniques on Deep Neural Network (DNN) have been widely acknowledged as an effective way to achieve acceleration on a variety of platforms, and DNN weight pruning is a straightforward and effective method. There are currently two mainstreams of pruning methods representing two extremes of pruning reg…
New method uses LLMs to generate detailed scientific hypotheses.
problem Generating detailed, actionable scientific hypotheses from coarse initial directions.
method Hierarchical search method that incrementally adds details to hypotheses.
result Hierarchical search method consistently outperforms strong baselines on expert-annotated hypotheses.
The precise diagnosis is of great significance in developing precise treatment plans to restore neck function and reduce the burden posed by the cervical spondylosis (CS). However, the current available neck function assessment method are subjective and coarse-grained. In this paper, based on the relationship among CS,…
Generative framework learns effective, lower-dimensional models from high-dimensional data.
problem Predicting long-term behavior of complex, multiscale systems with limited data.
method Physics-aware probabilistic model order reduction with latent variables.
result Guaranteed long-term stability and predictive accuracy in multiscale physical systems.
Paper develops fine-grain spatiotemporal risk scores using high-resolution mobility data.
problem Developing reliable spatiotemporal risk scores for safe economic reopening.
method Hawkes process-based technique leveraging high-resolution cell-phone location signals.
result Fine-grain spatiotemporal risk scores based on high-resolution mobility data provide useful insights for safe re-opening.
We present a representation learning method that learns features at multiple different levels of scale. Working within the unsupervised framework of denoising autoencoders, we observe that when the input is heavily corrupted during training, the network tends to learn coarse-grained features, whereas when the input is …
Aspect-level sentiment classification (ASC) aims at identifying sentiment polarities towards aspects in a sentence, where the aspect can behave as a general Aspect Category (AC) or a specific Aspect Term (AT). However, due to the especially expensive and labor-intensive labeling, existing public corpora in AT-level are…
Machine learning generates coarse-grained force fields for molecular dynamics.
problem Creating thermodynamically consistent coarse-grained models for larger systems.
method Hybrid architecture using graph neural networks to learn molecular features.
result Framework reproduces thermodynamics for small biomolecular systems.
GDML learns effective CG models from all-atom data.
problem Learning effective coarse-grained force fields efficiently.
method Ensemble learning with stratified sampling and GDML.
result GDML yields smaller free energy error than neural networks.
Atomistic or ab-initio molecular dynamics simulations are widely used to predict thermodynamics and kinetics and relate them to molecular structure. A common approach to go beyond the time- and length-scales accessible with such computationally expensive simulations is the definition of coarse-grained molecular models.…
Decision-tree-based ensemble classification methods (DTEMs) are a prevalent tool for supervised anomaly detection. However, due to the continued growth of datasets, DTEMs result in increasing drawbacks such as growing memory footprints, longer training times, and slower classification latencies at lower throughput. In …
Paper reduces neural network complexity for image classification.
problem High computational complexity in deep neural networks.
method Proposes a two-step classification process: coarse-grain and fine-grain.
result Achieves similar accuracy with less computational complexity.
Data coarse graining improves model performance by filtering out less relevant features.
problem Lossy data transformations lose information but can improve model generalization.
method Data coarse graining schemes that systematically discard features based on relevance to the learning task.
result A 'high-pass' scheme helps models generalize better by filtering out less relevant features.
DataRater learns which data points are most valuable for training models.
problem Training model efficiency depends on high-quality training data.
method Meta-learning to estimate the value of data points for training.
result Meta-learning improves compute efficiency by filtering data effectively.
New framework embeds physics in coarse-grained models without big data.
problem Lack of big data and computational demand in data-driven coarse-graining.
method Proposes a novel objective based on reverse Kullback-Leibler divergence that incorporates physics in the form of force fields.
result Generative coarse-grained model predicts atomistic configurations and reveals physicochemical CVs.
Molecular dynamics simulations provide theoretical insight into the microscopic behavior of materials in condensed phase and, as a predictive tool, enable computational design of new compounds. However, because of the large temporal and spatial scales involved in thermodynamic and kinetic phenomena in materials, atomis…
PSimGNN partitions graphs into subgraphs for efficient graph similarity computation.
problem Efficiently compute graph similarity scores for large graphs.
method Graph partitioning followed by subgraph-level and node-level comparisons using a graph neural network.
result PSimGNN outperforms state-of-the-art methods in graph similarity computation tasks.
Financial markets analyzed by reducing correlation matrix complexity.
problem Understanding complex financial market correlations.
method Coarse graining Pearson correlation matrices into Guhr matrices by market sectors.
result Significant reduction in the number of relevant variables.
Temporal coarse-graining of latent default paths explains effective correlation in corporate defaults.
problem Understanding effective default correlation in corporate defaults.
method Temporal coarse-graining of latent default-probability paths, applied to corporate default-count data.
result Temporal coarse-graining provides a scale-consistent baseline that improves identifiability and reduces over-allocation of long-horizon fluctuations.
Novel algorithm detects causal macrovariables from high-dimensional data.
problem Leveraging high-dimensional observational datasets for coarse-grained causal models.
method Inspired by information bottlenecks, novel algorithm detects macrovariables and investigates causal relationships through additive noise models.
result Algorithm robustly detects and infers causal relationships in both synthetic and real climate datasets.
CG-BGs combine flow-based models with PMFs to sample large systems efficiently.
problem Sampling equilibrium molecular configurations from the Boltzmann distribution is challenging.
method Coarse-grained Boltzmann Generators (CG-BGs) use flow-based models and learned PMFs for efficient sampling.
result CG-BGs provide a practical route for sampling larger molecular systems efficiently.
DiAMoNDBack models protein backmapping from coarse-grained Cα traces.
problem Restoring all-atom details from coarse-grained protein representations.
method Autoregressive denoising diffusion model for residue-by-residue backmapping.
result Achieves state-of-the-art reconstruction performance in diverse applications.
Framework preserves emergent physics in non-equilibrium systems from particle trajectories.
problem Linking short spatiotemporal scales to emergent bulk physics in multiscale systems.
method Metriplectic bracket formalism for structure-preserving coarse-graining.
result Preservation of thermodynamic laws and conservation in machine-learned dynamics.
Temporal aggregation reveals latent default correlation from monthly data.
problem Understanding effective default correlation from monthly default data.
method Temporal coarse-graining of latent default-probability paths.
result Temporal coarse-graining improves identifiability and reduces over-allocation of long-horizon fluctuations.
New method uses normalizing flows to improve force fields for coarse-grained molecular dynamics.
problem Lack of reference atomistic forces makes force matching infeasible for MLCG force fields.
method Introduces noise-based kernels adapted to low-data regimes using normalizing flows.
result Flow-based kernels reduce local distortions while preserving global accuracy.
Bayesian model uses physics constraints for semi-supervised surrogate learning.
problem Lack of labeled data in fine-grained model training.
method Probabilistic generative model with virtual observables.
result Enables semi-supervised training with unlabeled data.
Neural HMM with AGA captures multi-scale dynamics in financial markets.
problem Capturing multi-scale temporal dynamics in financial markets.
method Parallel multi-resolution encoders, adaptive gating, and multi-head attention.
result Outperforms fixed-resolution baselines in predicting price movements and liquidity shocks.
Modern applications of machine learning (ML) deal with increasingly heterogeneous datasets comprised of data collected from overlapping latent subpopulations. As a result, traditional models trained over large datasets may fail to recognize highly predictive localized effects in favour of weakly predictive global patte…
The paper tackles physical constraints in probabilistic machine learning for CG models of high-dimensional systems.
problem Introducing physical constraints in probabilistic machine learning objectives for coarse-graining dynamical systems.
method Formulating coarse-graining process using probabilistic state-space model and accounting for constraints as virtual observables.
result Probabilistic inference tools can identify coarse-grained variables without needing a fine-to-coarse projection or time-derivatives.
We propose a probabilistic model for inferring the multivariate function from multiple areal data sets with various granularities. Here, the areal data are observed not at location points but at regions. Existing regression-based models can only utilize the sufficiently fine-grained auxiliary data sets on the same doma…
Graph neural network predicts optimal coarse-grained mapping operators.
problem Optimal coarse-grained mapping operators selection for molecular dynamics simulations.
method Graph Neural Network (DSGPM) trained on expert-annotated data.
result DSGPM outperforms state-of-the-art methods in graph segmentation.
We propose a new approach for analyzing price fluctuations in their strongly correlated regime ranging from minutes to months. This is done by employing a self-similarity assumption for the magnitude of coarse-grained price fluctuation or volatility. The existence of a Cramer function, the characteristic function for s…
Extracts coarse-grained PDEs from microscopic simulations.
problem Discovering effective PDEs for macro-scale processes from microscopic data.
method Combining neural networks with equation-free numerics and data-driven approaches.
result Efficiently discovers macro-scale PDEs from microscopic simulations.
Survey of weak form's role in equation learning, parameter estimation, and coarse graining.
problem Noise robustness, accuracy, and computational efficiency in weak form applications.
method Survey and recent developments in weak form versions of equation learning, parameter estimation, and coarse graining.
result Surprising noise robustness, accuracy, and computational efficiency in weak form applications.
A new machine-learned CG model predicts protein structures efficiently.
problem Developing a universal, computationally efficient protein simulation model.
method Combining deep learning with all-atom protein simulations to create a transferable CG force field.
result The model predicts protein structures, intermediates, and fluctuations efficiently.
ElasTST improves time-series forecasting across varying horizons.
problem Robust forecasting across different time horizons in varied industrial sectors.
method Elastic Time-Series Transformer (ElasTST) with non-autoregressive design, rotary position embedding, and multi-scale patching.
result ElasTST provides robust forecasts across varying horizons without retraining.