Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

135270405540 · Jun 202019922001200920172026
48 results for Positional information

Neural networks have been proposed recently for positioning and channel charting of user equipments (UEs) in wireless systems. Both of these approaches process channel state information (CSI) that is acquired at a multi-antenna base-station in order to learn a function that maps CSI to location information. CSI-based p…

2019-09-29abs ↗pdf ↗

A family of probability distributions parametrized by an open domain ΛΛ in RnR^n defines the Fisher information matrix on this domain which is positive semi-definite. In information geometry the standard assumption has been that the Fisher information matrix tensor is positive definite defining in this way a Riemannia…

2015-03-29abs ↗pdf ↗

While invasively recorded brain activity is known to provide detailed information on motor commands, it is an open question at what level of detail information about positions of body parts can be decoded from non-invasively acquired signals. In this work it is shown that index finger positions can be differentiated fr…

2015-12-14abs ↗pdf ↗

Forecasting the future traffic flow distribution in an area is an important issue for traffic management in an intelligent transportation system. The key challenge of traffic prediction is to capture spatial and temporal relations between future traffic flows and historical traffic due to highly dynamical patterns of h…

2019-04-12abs ↗pdf ↗

cMIM improves representation learning without positive-pair augmentations.

problem Learning robust representations for diverse tasks.
method Contrastive Mutual Information Machine (cMIM) framework.
result cMIM outperforms MIM and InfoNCE on classification and regression tasks.

BiPE blends intra-segment and inter-segment encodings for better length extrapolation.

problem Improving length extrapolation in language models.
method Bilevel Positional Encoding (BiPE) that separates intra-segment and inter-segment encodings.
result BiPE enhances length extrapolation across various text modalities.

Graph neural controlled differential equations learn graph dynamics from vertex observations.

problem Predicting future states of dynamical systems on graphs with limited vertex data.
method Incorporates graph topology information into NCDE to predict graph dynamics.
result Informed NCDE requires fewer parameters and lower MAE compared to previous methods.

This paper tackles negative transfer in multi-task learning by introducing class-wise weights.

problem Negative transfer hampers function from achieving optimality in multi-task learning.
method Introduces class-wise weights to drive positive transfer and suppress negative transfer.
result Demonstrates improved performance in multi-task learning by reducing negative transfer.

A new DRL model for intraday trading incorporating positional context.

problem Neglecting positional context in existing DRL intraday trading strategies.
method Introducing positional features into the state space of a DRL model.
result Significant improvement in profitability and risk-adjusted metrics.

ChatGPT can summarize corporate disclosures more concisely and effectively, improving stock market reactions.

problem Information asymmetry and inefficiency in stock markets due to bloated disclosures.
method Comparing ChatGPT-generated summaries to original disclosures, analyzing their impact on stock market reactions.
result ChatGPT-generated summaries are more effective at explaining stock market reactions to disclosed information.

LLMs learn new tasks from unstructured data, but it depends on word co-occurrence and positional information.

problem Understanding how LLMs can learn new tasks from unstructured data without explicit training.
method Examined the capabilities of LLMs trained on unstructured data, focusing on sequence model requirements and training data structure.
result Many ICL capabilities can emerge from word co-occurrence in unstructured data, but positional information is crucial for certain tasks.

New insights into Markov chain geometry via positive transition measures.

problem Lack of statistical meaning in the space of transition probabilities.
method Constructing an extension of the space of transition probabilities using Amari's theory of positive measures.
result Introduction of a new dually flat structure for the space of positive transition measures.

MEANTIME improves sequential recommendation by using multi-temporal embeddings and attention mechanisms.

problem Limited use of timestamp information and information bottleneck in sequential recommendation models.
method MEANTIME employs multiple types of temporal embeddings and attention mechanisms to capture diverse patterns from user behavior sequences.
result MEANTIME outperforms state-of-the-art sequential recommendation methods.

We introduce a method for creating a special type of tree, called a tree position, from a weighted graph. Leaves of the tree correspond to vertices of the original graph, and the tree edges contain information which can be used to partition these vertices. By repeatedly applying reducing operations to the tree position…

2014-08-15abs ↗pdf ↗

New method recovers graph latent positions under edge differential privacy.

problem Recovering latent graph information from privatized graphs.
method Applying geometric insights to adjust statistical inference for privatized graphs.
result Achieves consistent recovery of latent positions under local edge differential privacy constraints.

The article examines entropy-information inequalities for continuous-time Markov chains under curvature-dimension conditions.

problem Proving Li-Yau inequalities and modified logarithmic Sobolev inequalities for reversible Markov chains.
method Introducing the CDΥ(κ,F)CD_Υ(κ,F) condition and deriving entropy-information inequalities.
result Derives functional inequalities relating entropy to Fisher information.

As algorithmic prediction systems have become widespread, fears that these systems may inadvertently discriminate against members of underrepresented populations have grown. With the goal of understanding fundamental principles that underpin the growing number of approaches to mitigating algorithmic discrimination, we …

2019-04-22abs ↗pdf ↗

FisherNet extends Autoencoder using Fisher information for better data reconstruction.

problem Data reconstruction accuracy and model scalability in high-dimensional latent spaces.
method Introduces FisherNet architecture that uses Fisher information to quantify and account for latent space uncertainty.
result FisherNet produces more accurate reconstructions and scales better with latent space dimensions compared to VAE.

The paper sets information-theoretic lower bounds for neural networks' parameter recovery and excess risk.

problem Establishing sample complexity lower bounds for neural network parameters and excess risk.
method Using information-theoretic tools, the paper proves lower bounds by constructing a generative network.
result Proves information-theoretic lower bounds for exact parameter recovery and positive excess risk.

New geometric structures defined on SPD matrices for better understanding.

problem Understanding SPD matrices and their geometric properties.
method Introducing Finslerian and dual information-geometric structures on James' bicone domain.
result Geodesics correspond to straight lines in coordinate systems, and new dissimilarities generalize existing ones.

Advocates against over-smoothing and over-squashing in GNNs, suggesting they are less critical than previously thought.

problem Over-smoothing and over-squashing in Graph Neural Networks (GNNs).
method Challenged the prevailing focus on these phenomena, proposing that performance decreases are due to uninformative receptive fields and localised information distribution.
result Performance decreases are mostly uncorrelated with over-smoothing and over-squashing, and optimal model depths remain small.

PiNGDA learns beneficial noise for graph augmentation stability.

problem Challenges in generating effective and stable graph augmentations.
method PiNGDA uses positive-incentive noise to scientifically analyze and generate beneficial graph augmentations.
result PiNGDA improves GCL performance by learning beneficial noise on graph topology and attributes.

The paper studies 3D manifolds with positive scalar curvature and volume growth.

problem Understanding the geometry of 3D manifolds with positive scalar curvature.
method Analyzes volume and geometric properties of 3D complete manifolds with positive scalar curvature, considering different curvature conditions.
result Volume growth estimates for 3D manifolds with positive scalar curvature, answering Gromov's question affirmatively.

A (positive) locally convex curve in the 2-sphere is a curve with positive geodesic curvature (i.e., which always turns left). In the 3-sphere, it is a curve with positive torsion. In this work we discussed the topology of spaces of such curves with prescribed initial and final jets. The case of the 2-sphere is underst…

2016-08-16abs ↗pdf ↗

New method efficiently learns positive-definite curvature for neural nets.

problem Efficiently learn positive-definite curvature for neural net training.
method Spectral-factorized positive-definite curvature learning approach.
result Efficiently applies arbitrary matrix roots and generic curvature learning.

The information metric arises in statistics as a natural inner product on a space of probability distributions. In general this inner product is positive semi-definite but is potentially degenerate. By associating to an instanton its energy density, we can examine the information metric {\bf g} on the moduli spaces $\M…

1996-11-25abs ↗pdf ↗

We examine geometric properties of a knot J that are unchanged by taking a (p,q)-cable K of J. Specifically, we relate w(K) to w(J), where w(K) is the width of K in the sense of Gabai. We use this information to demonstrate that thin position is a minimal bridge position of J if and only if the same is true for K, and …

2010-10-15abs ↗pdf ↗

Negative step sizes improve second-order methods for neural networks.

problem Second-order methods discard negative curvature, limiting their effectiveness.
method Introduce negative step sizes in second-order methods combined with Wolfe line search.
result Negative step sizes lead to global convergence and improved performance.

A new multi-label CPC method improves mutual information estimation and representation learning.

problem Underestimation of mutual information in contrastive predictive coding.
method Introducing a multi-label classification problem to overcome the logm\log m bound in mutual information estimation.
result The new method exceeds the logm\log m bound and leads to better mutual information estimation and improved unsupervised representation learning.

Inferring the structural properties of a protein from its amino acid sequence is a challenging yet important problem in biology. Structures are not known for the vast majority of protein sequences, but structure is critical for understanding function. Existing approaches for detecting structural similarity between prot…

2019-02-22abs ↗pdf ↗

Let G be a discrete group, and let M be a closed spin manifold of dimension m>3 with pi_1(M)=G. We assume that M admits a Riemannian metric of positive scalar curvature. We discuss how to use the L2-rho invariant and the delocalized eta invariant associated to the Dirac operator on M in order to get information about t…

2006-04-13abs ↗pdf ↗

A new mutual information optimization method using self-supervised binary contrastive learning.

problem Improving self-supervised contrastive learning for better model performance.
method Proposes a novel loss function for contrastive learning that optimizes mutual information in positive and negative pairs.
result The proposed method outperforms state-of-the-art self-supervised contrastive frameworks on various benchmark datasets.

GraphReach improves GNN performance by incorporating node positions.

problem Existing GNNs fail to capture node positions, leading to inaccurate predictions.
method GraphReach uses reachability estimations from anchor nodes to capture global node positions.
result GraphReach achieves up to 40% relative improvement in accuracy compared to state-of-the-art GNNs.

Agol recently introduced the notion of a veering triangulation, and showed that such triangulations naturally arise as layered triangulations of fibered hyperbolic 3-manifolds. We prove, by a constructive argument, that every veering triangulation admits positive angle structures, recovering a result of Hodgson, Rubins…

2010-12-23abs ↗pdf ↗