Few-shot visual reasoning model learns analogical relationships from small data.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
SCL discovers compositional structures in analogical reasoning tasks.
Building on a specific formalization of analogical relationships of the form "A relates to B as C relates to D", we establish a connection between two important subfields of artificial intelligence, namely analogical reasoning and kernel-based machine learning. More specifically, we show that so-called analogical propo…
ADR helps LLMs find and use historical analogies for foresight analysis.
LLMs learn new tasks from unstructured data, but it depends on word co-occurrence and positional information.
Object ranking or "learning to rank" is an important problem in the realm of preference learning. On the basis of training data in the form of a set of rankings of objects represented as feature vectors, the goal is to learn a ranking function that predicts a linear order of any new set of objects. In this paper, we pr…
Tab-TRM uses recursive model for insurance pricing on tabular data.
Transformers solve parity problems efficiently with step-by-step reasoning.
We describe a post hoc test for the Sharpe ratio, analogous to Tukey's test for pairwise equality of means. The test can be applied after rejection of the hypothesis that all population Signal-Noise ratios are equal. The test is applicable under a simple correlation structure among asset returns. Simulations indicate t…
When we are faced with challenging image classification tasks, we often explain our reasoning by dissecting the image, and pointing out prototypical aspects of one class or another. The mounting evidence for each of the classes helps us make our final decision. In this work, we introduce a deep network architecture -- …
Optimal trading patterns adjust based on market efficiency and slippage costs.
In 2011 Enders, Müller and Topping showed that any blow up sequence of a Type I Ricci flow near a singular point converges to a non-trivial gradient Ricci soliton, leading them to conclude that for such flows all reasonable definitions of singular points agree with each other. We prove the analogous result for the harm…
An important problem in multi-label classification is to capture label patterns or underlying structures that have an impact on such patterns. This paper addresses one such problem, namely how to exploit hierarchical structures over labels. We present a novel method to learn vector representations of a label space give…
Reinforcement learning has gained wide popularity as a technique for simulation-driven approximate dynamic programming. A less known aspect is that the very reasons that make it effective in dynamic programming can also be leveraged for using it for distributed schemes for certain matrix computations involving non-nega…
In the eyes of a rationalist like Descartes or Spinoza, human reasoning is flawless, marching toward uncovering ultimate truth. A few centuries later, however, culminating in the work of Kahneman and Tversky, human reasoning was portrayed as anything but flawless, filled with numerous misjudgments, biases, and cognitiv…
The concept of explainability is envisioned to satisfy society's demands for transparency on machine learning decisions. The concept is simple: like humans, algorithms should explain the rationale behind their decisions so that their fairness can be assessed. While this approach is promising in a local context (e.g. to…
Dropout is a popular regularization technique in deep learning. Yet, the reason for its success is still not fully understood. This paper provides a new interpretation of Dropout from a frame theory perspective. By drawing a connection to recent developments in analog channel coding, we suggest that for a certain famil…
In Lorentzian manifolds of any dimension the concept of causal tensors is introduced. Causal tensors have positivity properties analogous to the so-called ``dominant energy condition''. Further, it is shown how to build, from ANY given tensor , a new tensor quadratic in and ``positive'', in the sense that it is …
Study of generalized knots and links, proving inequality involving crossing number and braid index.
Online learners track optimal solutions with constant step-size.
Entropy is a natural geometric quantity measuring the complexity of a surface embedded in . For dynamical reasons relating to mean curvature flow, Colding-Ilmanen-Minicozzi-White conjectured that the entropy of any closed surface is at least that of the self-shrinking two-sphere. We prove this conjecture …
Bob predicts a future observation based on a sample of size one. Alice can draw a sample of any size before issuing her prediction. How much better can she do than Bob? Perhaps surprisingly, under a large class of loss functions, which we refer to as the Cover-Hart family, the best Alice can do is to halve Bob's risk. …
We consider the problem of developing suitable learning representations (embeddings) for library packages that capture semantic similarity among libraries. Such representations are known to improve the performance of downstream learning tasks (e.g. classification) or applications such as contextual search and analogica…
This dissertation explores the integration of learning and analogy-making through the development of a computer program, called Analogator, that learns to make analogies by example. By "seeing" many different analogy problems, along with possible solutions, Analogator gradually develops an ability to make new analogies…
Analog forecasting uses local dynamics to predict chaotic systems.
Analog methods improve forecast accuracy in complex models.
The paper evaluates the probability distributions of analog-to-target distances for multiple analogs.
Raven's Progressive Matrices are one of the widely used tests in evaluating the human test taker's fluid intelligence. Analogously, this paper introduces geometric generalization based zero-shot learning tests to measure the rapid learning ability and the internal consistency of deep generative models. Our empirical re…
In our previous paper (see this arxiv math.DG/0402171) for generic rank 2 vector distributions on n-dimensional manifold (n greater or equal to 5) we constructed a special differential invariant, the fundamental form. In the case n=5 this differential invariant has the same algebraic nature, as the covariant binary biq…
The homology groups of many natural sequences of groups (e.g. general linear groups, mapping class groups, etc.) stabilize as . Indeed, there is a well-known machine for proving such results that goes back to early work of Quillen. Church and Farb discovered that many sequ…
According to an analogy to quasi-Fuchsian groups, we investigate topological and combinatorial structures of Lyubich and Minsky's affine and hyperbolic 3-laminations associated with the hyperbolic and parabolic quadratic maps. We begin by showing that hyperbolic rational maps in the same hyperbolic component have quasi…
LaTRO optimizes latent reasoning in LLMs without external reward.
Motivated by safety-critical applications, test-time attacks on classifiers via adversarial examples has recently received a great deal of attention. However, there is a general lack of understanding on why adversarial examples arise; whether they originate due to inherent properties of data or due to lack of training …
Auto-CEI improves LLM reasoning by balancing assertiveness and conservativeness.
A framework isolates VQA reasoning from perception for better model evaluation.
A new method for math reasoning that allows for iterative correction.
Inferring new facts from existing knowledge graphs (KG) with explainable reasoning processes is a significant problem and has received much attention recently. However, few studies have focused on relation types unseen in the original KG, given only one or a few instances for training. To bridge this gap, we propose Co…
Transformers learn multi-step reasoning through gradient descent.
Transformers with CoT don't enhance reasoning power across all tasks.
Early stopping methods reduce unnecessary reasoning steps in LLMs by monitoring uncertainty signals.
Achieving artificial visual reasoning - the ability to answer image-related questions which require a multi-step, high-level process - is an important step towards artificial general intelligence. This multi-modal task requires learning a question-dependent, structured reasoning process over images from language. Stand…
FinTradeBench benchmarks LLMs for financial reasoning combining company fundamentals and market signals.
Hybrid framework injects TSLM insights into GRLM for robust time-series reasoning.
Defines an odd analog of Plamenevskaya's invariant for transverse links.
Neural Logic Reasoning integrates deep learning and symbolic logic for better prediction tasks.
Forward-prediction models enhance physical reasoning, but only for specific tasks.
Analog method solves portfolio optimization problems faster and more efficiently.
Without relevant human priors, neural networks may learn uninterpretable features. We propose Dynamics of Attention for Focus Transition (DAFT) as a human prior for machine reasoning. DAFT is a novel method that regularizes attention-based reasoning by modelling it as a continuous dynamical system using neural ordinary…