SMAI framework tests and integrates single-cell data alignability.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Conformal Alignment ensures trustworthy outputs from foundation models.
DTA aligns multimodal data with prior correspondence knowledge.
We present Manifold Alignment Determination (MAD), an algorithm for learning alignments between data points from multiple views or modalities. The approach is capable of learning correspondences between views as well as correspondences between individual data-points. The proposed method requires only a few aligned exam…
New method aligns brain data across individuals for better brain decoding.
Proposes ACCA for better alignment of multiple data perspectives.
SrvfNet aligns multiple functional data to templates without supervision.
Congealing is a flexible nonparametric data-driven framework for the joint alignment of data. It has been successfully applied to the joint alignment of binary images of digits, binary images of object silhouettes, grayscale MRI images, color images of cars and faces, and 3D brain volumes. This research enhances congea…
MALI aligns distinct domains using labeled data.
This paper improves cross-domain learning using random forests for manifold alignment.
Kernel alignment measures the degree of similarity between two kernels. In this paper, inspired from kernel alignment, we propose a new Linear Discriminant Analysis (LDA) formulation, kernel alignment LDA (kaLDA). We first define two kernels, data kernel and class indicator kernel. The problem is to find a subspace to …
Unsupervised domain adaptation aims to transfer and adapt knowledge learned from a labeled source domain to an unlabeled target domain. Key components of unsupervised domain adaptation include: (a) maximizing performance on the target, and (b) aligning the source and target domains. Traditionally, these tasks have eith…
Deep neural networks show some layers better align with data than others.
DFA trains deep networks by aligning weights then memorizing data.
Direct Density Ratio Optimization aligns LLMs with human preferences without assuming specific models.
Hölder-DPO aligns models robustly with noisy human feedback.
Two semi-supervised manifold alignment methods improve cross-domain classification.
New method aligns brain surfaces based on functional signatures.
We propose a novel framework for combining datasets via alignment of their intrinsic geometry. This alignment can be used to fuse data originating from disparate modalities, or to correct batch effects while preserving intrinsic data structure. Importantly, we do not assume any pointwise correspondence between datasets…
We present ChromAlignNet, a deep learning model for alignment of peaks in Gas Chromatography-Mass Spectrometry (GC-MS) data. In GC-MS data, a compound's retention time (RT) may not stay fixed across multiple chromatograms. To use GC-MS data for biomarker discovery requires alignment of identical analyte's RT from diffe…
We introduce a kernel method for manifold alignment (KEMA) and domain adaptation that can match an arbitrary number of data sources without needing corresponding pairs, just few labeled examples in all domains. KEMA has interesting properties: 1) it generalizes other manifold alignment methods, 2) it can align manifold…
Alignment of neural network representations is influenced by SNR and sample size.
A novel OT-based method for aligning hyperbolic representations.
NTW aligns multiple time-series data efficiently using neural networks.
This study presents the results of a series of simulation experiments that evaluate and compare four different manifold alignment methods under the influence of noise. The data was created by simulating the dynamics of two slightly different double pendulums in three-dimensional space. The method of semi-supervised fea…
Objective: This paper targets a major challenge in developing practical EEG-based brain-computer interfaces (BCIs): how to cope with individual differences so that better learning performance can be obtained for a new subject, with minimum or even no subject-specific data? Methods: We propose a novel approach to align …
Robustly aligns datasets with partial GW distance to handle contamination.
AI models aligned with human vision perform well on few data tasks.
A new method reduces preference distortion in LLM alignment.
The application of machine learning to bioinformatics problems is well established. Less well understood is the application of bioinformatics techniques to machine learning and, in particular, the representation of non-biological data as biosequences. The aim of this paper is to explore the effects of giving amino acid…
Paper tackles entity matching over multi-source data, optimizing alignment and mitigating negative transfer.
FedCVT improves VFL models with limited aligned samples.
MKA incorporates manifold geometry into kernel alignment for more robust representation comparison.
Proposes a regularization method for unsupervised domain adaptation that aligns predictions with target data's top singular vectors.
We introduce BilBOWA (Bilingual Bag-of-Words without Alignments), a simple and computationally-efficient model for learning bilingual distributed representations of words which can scale to large monolingual datasets and does not require word-aligned parallel training data. Instead it trains directly on monolingual dat…
The paper proposes a method to align AI models using conformal risk control.
We present a probabilistic model for unsupervised alignment of high-dimensional time-warped sequences based on the Dirichlet Process Mixture Model (DPMM). We follow the approach introduced in (Kazlauskaite, 2018) of simultaneously representing each data sequence as a composition of a true underlying function and a time…
Unbalanced COOT improves feature alignment robustly to outliers.
In many machine learning applications, it is necessary to meaningfully aggregate, through alignment, different but related datasets. Optimal transport (OT)-based approaches pose alignment as a divergence minimization problem: the aim is to transform a source dataset to match a target dataset using the Wasserstein dista…
The paper examines how spike strengths and alignments affect overfitting in linear regression models.
New model optimizes feature alignment, improving statistical and computational efficiency.
Geometric approach improves motion alignment accuracy and efficiency.
The study explains how neural networks align their kernels to target functions during training.
We present a model that can automatically learn alignments between high-dimensional data in an unsupervised manner. Our proposed method casts alignment learning in a framework where both alignment and data are modelled simultaneously. Further, we automatically infer groupings of different types of sequences within the …
Multi-subject fMRI data analysis is an interesting and challenging problem in human brain decoding studies. The inherent anatomical and functional variability across subjects make it necessary to do both anatomical and functional alignment before classification analysis. Besides, when it comes to big data, time complex…
We examine the influence of input data representations on learning complexity. For learning, we posit that each model implicitly uses a candidate model distribution for unexplained variations in the data, its noise model. If the model distribution is not well aligned to the true distribution, then even relevant variati…
Deep neural networks, trained with large amount of labeled data, can fail to generalize well when tested with examples from a \emph{target domain} whose distribution differs from the training data distribution, referred as the \emph{source domain}. It can be expensive or even infeasible to obtain required amount of lab…
NTKs explain GNNs' alignment for graph prediction.