Deep-learning CNN automates Cu alloy grain size evaluation.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Pipelined Backpropagation trains large models without batches efficiently.
The paper reduces the complexity of financial market correlation matrices to a 2x2 matrix.
Study examines spillovers between BRICS and U.S. staple grain futures markets.
Study examines grain futures connectedness during Russia-Ukraine conflict.
New model predicts grain boundary migration in metals.
The paper studies scaling laws for associative memory mechanisms.
Fine-grained analysis of gradient descent with momentum provides modified loss equations.
Fine-grained atlases improve fMRI analysis of brain activity.
Machine learning generates coarse-grained force fields for molecular dynamics.
We examine the correlation of the limit price with the order book, when a limit order comes. We analyzed the Rebuild Order Book of Stock Exchange Electronic Trading Service, which is the centralized order book market of London Stock Exchange. As a result, the limit price is broadly distributed around the best price acc…
We develop an explicit and tractable representation of a twist-grain-boundary phase of a smectic A liquid crystal. This allows us to calculate the interaction energy between grain boundaries and the relative contributions from the bending and compression deformations. We discuss the special stability of the 90 degree g…
GDML learns effective CG models from all-atom data.
Atomistic or ab-initio molecular dynamics simulations are widely used to predict thermodynamics and kinetics and relate them to molecular structure. A common approach to go beyond the time- and length-scales accessible with such computationally expensive simulations is the definition of coarse-grained molecular models.…
We propose a probabilistic model for refining coarse-grained spatial data by utilizing auxiliary spatial data sets. Existing methods require that the spatial granularities of the auxiliary data sets are the same as the desired granularity of target data. The proposed model can effectively make use of auxiliary data set…
Considering event structure information has proven helpful in text-based stock movement prediction. However, existing works mainly adopt the coarse-grained events, which loses the specific semantic information of diverse event types. In this work, we propose to incorporate the fine-grained events in stock movement pred…
As entity type systems become richer and more fine-grained, we expect the number of types assigned to a given entity to increase. However, most fine-grained typing work has focused on datasets that exhibit a low degree of type multiplicity. In this paper, we consider the high-multiplicity regime inherent in data source…
The ubiquitous deployment of monitoring devices in urban flow monitoring systems induces a significant cost for maintenance and operation. A technique is required to reduce the number of deployed devices, while preventing the degeneration of data accuracy and granularity. In this paper, we present an approach for infer…
Modeling financial markets with sandpile model to understand price volatility and arbitrage constraints.
Study finds intrinsic multifractality in maize and barley spot markets, but not in wheat and rice.
Fine-grained pretraining improves neural network's ability to learn rare features.
We introduce a hierarchical architecture for video understanding that exploits the structure of real world actions by capturing targets at different levels of granularity. We design the model such that it first learns simpler coarse-grained tasks, and then moves on to learn more fine-grained targets. The model is train…
PSimGNN partitions graphs into subgraphs for efficient graph similarity computation.
Data coarse graining improves model performance by filtering out less relevant features.
Large twist-angle grain boundaries in layered structures are often described by Scherk's first surface whereas small twist-angle grain boundaries are usually described in terms of an array of screw dislocations. We show that there is no essential distinction between these two descriptions and that, in particular, their…
Fine-grained gap-dependent regret bounds for reinforcement learning.
Molecular dynamics simulations provide theoretical insight into the microscopic behavior of materials in condensed phase and, as a predictive tool, enable computational design of new compounds. However, because of the large temporal and spatial scales involved in thermodynamic and kinetic phenomena in materials, atomis…
In recent years, dock-less shared bikes have been widely spread across many cities in China and facilitate people's lives. However, at the same time, it also raises many problems about dock-less shared bike management due to the mismatching between demands and real distribution of bikes. Before deploying dock-less shar…
This research examines how the error rate of nearest neighbor classifiers varies with dataset size.
Financial markets analyzed by reducing correlation matrix complexity.
Temporal coarse-graining of latent default paths explains effective correlation in corporate defaults.
Sparsity helps reduce the computational complexity of deep neural networks by skipping zeros. Taking advantage of sparsity is listed as a high priority in next generation DNN accelerators such as TPU. The structure of sparsity, i.e., the granularity of pruning, affects the efficiency of hardware accelerator design as w…
Cellular regulatory dynamics is driven by large and intricate networks of interactions at the molecular scale, whose sheer size obfuscates understanding. In light of limited experimental data, many parameters of such dynamics are unknown, and thus models built on the detailed, mechanistic viewpoint overfit and are not …
We study the problem of object detection over scanned images of scientific documents. We consider images that contain objects of varying aspect ratios and sizes and range from coarse elements such as tables and figures to fine elements such as equations and section headers. We find that current object detectors fail to…
FIGARO generates symbolic music with fine-grained control.
Improved investment performance with fine-grained LLM tasks.
Recent works have cast some light on the mystery of why deep nets fit any data and generalize despite being very overparametrized. This paper analyzes training and generalization for a simple 2-layer ReLU net with random initialization, and provides the following improvements over recent works: (i) Using a tighter char…
CG-BGs combine flow-based models with PMFs to sample large systems efficiently.
Inspired by coarse-graining approaches used in physics, we show how similar algorithms can be adapted for data. The resulting algorithms are based on layered tree tensor networks and scale linearly with both the dimension of the input and the training set size. Computing most of the layers with an unsupervised algorith…
We present an algorithm for supervised learning using tensor networks, employing a step of preprocessing the data by coarse-graining through a sequence of wavelet transformations. We represent these transformations as a set of tensor network layers identical to those in a multi-scale entanglement renormalization ansatz…
DistPre predicts traffic speeds efficiently for large networks.
Existing methods for retrieving k-nearest neighbours suffer from the curse of dimensionality. We argue this is caused in part by inherent deficiencies of space partitioning, which is the underlying strategy used by most existing methods. We devise a new strategy that avoids partitioning the vector space and present a n…
We present a novel learning framework that consistently embeds underlying physics while bypassing a significant drawback of most modern, data-driven coarse-grained approaches in the context of molecular dynamics (MD), i.e., the availability of big data. The generation of a sufficiently large training dataset poses a co…
DiAMoNDBack models protein backmapping from coarse-grained Cα traces.
We revisit the stochastic variance-reduced policy gradient (SVRPG) method proposed by Papini et al. (2018) for reinforcement learning. We provide an improved convergence analysis of SVRPG and show that it can find an -approximate stationary point of the performance function within trajectories. This s…
Framework preserves emergent physics in non-equilibrium systems from particle trajectories.
Temporal aggregation reveals latent default correlation from monthly data.
ECN framework improves training on noisy structured labels.