No best k-layer neural network approximations exist in general for common activations.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
For any positive integer , there exist neural networks with layers, nodes per layer, and distinct parameters which can not be approximated by networks with layers unless they are exponentially large --- they must possess nodes. This result is proved here for a class o…
New sparse penalty improves biclustering for gene expression data.
Proposes D-LADMM for solving constrained optimization problems.
Transformers can learn Markov processes with constant depth, surprising results.
Probabilistic bounds on neuron death in deep networks, showing depth can be increased indefinitely.
A usual reinsurance policy for insurance companies admits one or two layers of the payment deductions. Under optimal criterion of minimizing the conditional tail expectation (CTE) risk measure of the insurer's total risk, this article generalized an optimal stop-loss reinsurance policy to an optimal multi-layer reinsur…