Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

3877115153 · Jun 202019922001200920172026
48 results for sparse parities

Paper solves k-sparse parity problem with sign SGD, matching SQ lower bound.

problem Solving k-sparse parity problems efficiently.
method Sign stochastic gradient descent on neural networks.
result Matches Statistical Query lower bound for solving k-sparse parity problems.

Curriculum learning helps neural networks learn parities more efficiently.

problem Improving learning efficiency for neural networks on parity targets.
method Using a curriculum learning approach with a mixture of sparse and dense inputs.
result A 2-layer ReLU neural network can learn parities more efficiently than a fully connected network.

Solves a model for sudden problem-solving ability in deep learning.

problem Emergence of new problem-solving abilities in deep learning models.
method Solves a simple multi-linear model in a skill-basis, finding analytic expressions for emergence and scaling laws.
result Simple model captures sigmoidal emergence of multiple new skills in neural networks.

Extracting a curriculum from a teacher network improves distillation efficiency.

problem Efficiently training a small network using a large teacher network's output.
method Random projection of teacher network's hidden representations to progressively train the student network.
result Extracted curriculum significantly outperforms one-shot distillation and achieves similar performance to progressive distillation.

We introduce the 2-colour parity. It is a theory of parity for a large class of virtual links, defined using the interaction between orientations of the link components and a certain type of colouring. The 2-colour parity is an extension of the Gaussian parity, to which it reduces on virtual knots. We show that the 2-c…

2019-01-22abs ↗pdf ↗

In the present paper, we develop the parity theory invented in \cite{ManSb}; we construct new parities for two-component (virtual and free) links. New parities significantly depend on geometrical properties of diagrams; in particular, they are mutation-sensitive. New parities can be used practically in all problems, wh…

2015-08-23abs ↗pdf ↗

In \cite {FrKn,Sbornik} it was shown that in some knot theories the crucial role is played by {\em parity}, i.e.\ a function on crossings valued in {0,1}\{0,1\} and behaving nicely with respect to Reidemeister moves. Any parity allows one to construct functorial mappings from knots to knots, to refine many invariants and …

2011-02-24abs ↗pdf ↗

Derives FACT, an alternative to NFA for neural networks, explaining feature learning.

problem Understanding how neural networks learn representations.
method First-principles approach using first-order optimality conditions.
result FACT explains why NFA holds and provides a principled alternative.

This paper tackles fair Bayes-optimal classifiers under predictive parity, proving their limitations and proposing a new algorithm.

problem Ensuring fair Bayes-optimal classifiers under predictive parity, especially when group performance levels vary widely.
method Proving the limitations of fair Bayes-optimal classifiers under predictive parity and proposing a new adaptive thresholding algorithm, FairBayes-DPP.
result Fair Bayes-optimal classifiers under predictive parity may not hold if group performance levels vary widely, leading to within-group unfairness.

We use crossing parity to construct a generalization of biquandles for virtual knots which we call Parity Biquandles. These structures include all biquandles as a standard example referred to as the even parity biquandle. Additionally, we find all Parity Biquandles arising from the Alexander Biquandle and Quaternionic …

2011-03-15abs ↗pdf ↗

Diversified risk parity strategies outperform equally-weighted portfolios in various asset universes.

problem Finding optimal portfolio allocations that balance risk and reward.
method Integrates various reward-risk measures and generic allocation rules into diversified risk parity.
result Diversified reward-risk parity strategies exhibit higher average returns, Sharpe ratios, and Calmar ratios compared to equally-weighted risk portfolios.

We learn higher-order Markov random fields from evolving data, bypassing computational barriers.

problem Learning graphical models from temporally correlated samples, especially with noisy data.
method Using the trajectory data from Glauber dynamics, we develop an algorithm to recover the graph and parameters efficiently.
result We demonstrate efficient learning of higher-order Markov random fields from trajectory data, overcoming computational hardness.

CoT improves transformer sample efficiency by reducing input token dependencies and attention sparsity.

problem Transformer sample inefficiency in simple tasks.
method Demonstrated through parity-learning setup, showing CoT reduces required samples from exponential to polynomial.
result Transformer learns function within polynomial samples with CoT, requiring exponential samples without CoT.

Functorial maps and weak parities are equivalent descriptions of rules of substitution virtual crossings for classical in diagrams of a knot in a way compatible with Reidemeister moves. We introduce the notion of maximal weak parity and describe it for knots in a given closed oriented surface. This weak parity defines …

2012-11-02abs ↗pdf ↗

In [3] we constructed the parity-biquandle bracket valued in {\em pictures} (linear combinations of 44-valent graphs). We gave no example of classical links such that the parity-biquandle bracket of which is not trivial. In the present paper we slightly change the notation of the parity-biquandle bracket and give exam…

2019-11-17abs ↗pdf ↗

Study shows physical drift affects put-call parity enforcement, not just option payoffs.

problem Inconsistency between quoted put-call parity and actual market behavior.
method Examined SPX and RUT index options, used drift-preserving GBM term to improve fit.
result Physical drift enters the enforcement of risk-neutral parity, not just option payoffs.

Parity mappings from the chords of a Gauss diagram to the integers is defined. The parity of the chords is used to construct families of invariants of Gauss diagrams and virtual knots. One family consists of degree nn Vassiliev invariants.

2012-03-13abs ↗pdf ↗

We define counting and cocycle enhancement invariants of virtual knots using parity biquandles. These invariants are determined by pairs consisting of a biquandle 2-cocycle φ^0 and a map φ^1 with certain compatibility conditions leading to one-variable or two-variable polynomial invariants of virtual knots. We provide …

2015-07-20abs ↗pdf ↗

Transformers can generalize to a large task family with only a few demonstrations.

problem Can learning from a small set of tasks generalize to a large task family?
method Investigating autoregressive compositional structure where each task is a composition of TT operations, each from a finite family of DD subtasks.
result Transformers can generalize to DTD^T tasks with only O~(D)\widetilde{O}(D) demonstrations.

Transformers solve parity problems efficiently with step-by-step reasoning.

problem Training transformers to solve complex, recursive problems like parity.
method Training a one-layer transformer to solve kk-parity, incorporating intermediate parities into the loss function, and using teacher forcing or augmented data.
result Transformers can learn parity in one gradient update with intermediate supervision or self-consistency checks.

2-dimensional knots and links are studied in the article. The notion of parity is introduced via techniques similar to the ones used by the second named author in 1-dimensional case. By using parity new invariants are constructed and known invariants are refined.

2016-06-22abs ↗pdf ↗

Transformers learn sparse Boolean functions through RL and SFT, revealing distinct learning behaviors.

problem Learning sparse Boolean functions with Transformers.
method Reinforcement Learning (RL) with process rewards and Supervised Fine-Tuning (SFT).
result RL learns the whole CoT chain simultaneously, while SFT learns step by step.

New method controls bias in training data for fair outcomes.

problem Ensuring equal treatment between different groups in machine learning.
method Contrastive information estimation to control mutual information between representations and protected attributes.
result Our method provides strong theoretical guarantees on the parity of any downstream algorithm.

The article develops a model for skewness risk in risk parity portfolios.

problem Managing skewness risk in asset allocation models.
method Modeling asset returns with skewness and jumps, deriving analytical formulas for risk contributions.
result Skewness-based risk parity portfolios outperform volatility-based portfolios in managing jump risks.

We mathematically compare four competing definitions of group-level nondiscrimination: demographic parity, equalized odds, predictive parity, and calibration. Using the theoretical framework of Friedler et al., we study the properties of each definition under various worldviews, which are assumptions about how, if at a…

2018-08-26abs ↗pdf ↗

We consider knot theories possessing a {\em parity}: each crossing is decreed {\em odd} or {\em even} according to some universal rule. If this rule satisfies some simple axioms concerning the behaviour under Reidemeister moves, this leads to a possibility of constructing new invariants and proving minimality and non-t…

2009-12-29abs ↗pdf ↗

We investigate an application of crossing parity for the bracket expansion of the Jones polynomial for virtual knots. In addition we consider an application of parity for the arrow polynomial as well as for the categorifications of both polynomials. We present a number of examples found through our calculations. We pro…

2011-10-21abs ↗pdf ↗

Neural networks outperform NTK on compositional tasks, revealing a complexity gap.

problem Understanding the performance gap between neural networks and NTK on tasks with compositional structure.
method Characterized Fourier and architectural complexities, and analyzed the minimax rates of the architecture class.
result The NTK estimator is exponentially sub-optimal compared to the minimax floor when complexities decouple.

This paper is an introduction to virtual knot theory and an exposition of new ideas and constructions, including the parity bracket polynomial, the arrow polynomial, the parity arrow polynomial and categorifications of the arrow polynomial. The paper is relatively self-contained and it describes virtual knot theory bot…

2011-01-04abs ↗pdf ↗