Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,181 papers · 148 categories

Trend · papers per month

18355370 · Jun 202019922001200920182026
48 results for tandem assessment

New metric t-DCF improves ASVspoof challenge results by considering spoofing attack prior.

problem Shortcomings of EER metric in assessing CMs and ASV together.
method Developed t-DCF metric with 6 parameters to assess ASV and CMs together.
result t-DCF shows different rankings for higher spoofing attack priors.

SPRT-TANDEM improves sequential classification accuracy with fewer samples.

problem Efficiently classifying sequential data with high accuracy and low sampling cost.
method Deep neural network-based SPRT algorithm that estimates log-likelihood ratio of two hypotheses.
result SPRT-TANDEM achieves statistically significantly better classification accuracy than other classifiers with fewer samples.

Study introduces new financial ratios for better predicting company performance.

problem Lack of progress in predicting company performance and assessing financial risks.
method Developed new financial and macroeconomic ratios, supervised learning models, and Bayesian models.
result New proposed variables improve model accuracy and FNN performs best across multiple tasks.

STanHop predicts multivariate time series with memory-enhanced capabilities.

problem Predicting multivariate time series with memory-enhanced capabilities.
method Sparse Tandem Hopfield Network (STanHop) with two external memory modules.
result STanHop outperforms dense Hopfield models in memory retrieval error.

This study combines ASV and CM systems for better performance using reinforcement learning.

problem Improving the combined performance of ASV and CM systems for better t-DCF measure.
method Training ASV and CM components together using reinforcement learning.
result Training ASV and CM components together improves the performance of the combined system.

Improved protein identification in mass spectrometry data.

problem Expanding peptide scoring capabilities in tandem mass spectrometry.
method Deriving concave emission distributions for dynamic Bayesian networks.
result Efficiently learned scoring function outperforms state-of-the-art.

The paper shows how data augmentation and regularization can enforce group equivariance in machine learning models.

problem Improving model performance by leveraging known symmetries in machine learning tasks.
method Training with data augmentation and regularization to enforce group equivariance.
result Equivariance of the trained model can be achieved through training on augmented data in tandem with regularization.

Unified approach for sequence design combining likelihood-free inference and black-box optimization.

problem Designing biological sequences efficiently and accurately.
method Unified probabilistic framework integrating likelihood-free inference and black-box optimization.
result Previous optimization methods can be adapted and new algorithms proposed within this framework.

This work improves Gaussian process regression for large, non-stationary data.

problem Scalability issues and performance degradation for non-stationary data.
method Combines variational free energy approximations with online expectation propagation and local splitting steps.
result Incremental adaptation to locality, heterogeneity, and non-stationarity in training data.

Data augmentation doesn't improve robustness, contrary to belief.

problem The effectiveness of data augmentation in improving model robustness is questioned.
method Taking a Domain Generalization viewpoint, the study examines the robustness of augmented representations.
result Augmented representations are not robust to distortions used during training.

New research shows alternative linear connections can outperform identity shortcuts in deep networks.

problem Explaining the effectiveness of shortcut connections in deep neural networks.
method Used variations of the standard residual block with different types of linear connections to build image classification networks.
result Alternative linear connections can be more effective than identity shortcuts in deep networks.

Rewriting history improves RL algorithms for solving multiple tasks.

problem Improving sample efficiency in multi-task reinforcement learning.
method Introducing hindsight relabeling as inverse RL to generalize goal-relabeling techniques.
result Relabeling data using inverse RL accelerates learning in multi-task settings.

Study learns mixtures of smooth product distributions from samples.

problem Learning mixtures of non-parametric product distributions.
method Two-stage approach using identifiability properties of tensor decomposition and signal processing techniques.
result Recovery of component distributions under a smoothness condition.

Automates feature extraction from JSON data for machine learning.

problem Manual feature engineering for JSON data is laborious, lossy, and prone to bias.
method Automates feature extraction using Mill.jl and JsonGrinder.jl.
result Creates a differentiable machine learning model from raw JSON samples.

Adversarial learning approximates unknown quantum states on near-term quantum computers.

problem Approximating unknown quantum pure states on near-term quantum computers.
method Two parametrized circuits optimized adversarially, with resilient backpropagation and bipartite entanglement entropy.
result Resilient backpropagation algorithms perform well in optimizing the two circuits.

A new framework uses deep RL to aggregate expert advice for better portfolio management.

problem Improving portfolio management through expert advice and deep reinforcement learning.
method Convolutional networks for signal aggregation and historical price data, Proximal Policy Optimization algorithm.
result Our framework can achieve 90% of the best expert's profit on average.

Large learning rates enhance model robustness and compressibility.

problem Achieving robustness and resource-efficiency in machine learning models.
method Identifying and utilizing large learning rates as a facilitator for robustness and compressibility.
result Large learning rates produce desirable representation properties and compare favorably to other methods.

New method uses multi-task learning to improve molecule representations.

problem Cost, bias, and data requirements in chemical representation generation.
method Intelligent task selection in deep multitask networks with transfer learning.
result Deep representations capture more expressive task-based information.

New theory shows interpretable models can outperform black-box models in decision-making systems.

problem The importance of interpretability in machine learning models.
method Characterized performance of two-node data fusion systems using distributed detection theory.
result A human with an interpretable classifier outperforms one with a black-box classifier.

Self-attention models benefit equally from width and depth, but beyond a certain point, depth becomes less efficient.

problem Understanding the optimal balance between depth and width in self-attention models.
method Theoretical predictions and empirical ablations on networks of varying depths and widths.
result An optimal width of 30K is recommended for a 1-Trillion parameter network, marking a significant width for self-attention models.

Protein structure prediction has been a grand challenge problem in the structure biology over the last few decades. Protein quality assessment plays a very important role in protein structure prediction. In the paper, we propose a new protein quality assessment method which can predict both local and global quality of …

2016-02-13abs ↗pdf ↗

Paper introduces active Bayesian method for assessing black-box classifiers efficiently.

problem Need to assess performance of black-box classifiers reliably with limited labels.
method Develops inference strategies and proposes active Bayesian framework for efficient instance selection.
result Significant gains in performance assessment with fewer labels compared to traditional methods.

AI enhances refinery optimization by detecting data errors and improving decision-making.

problem Interpreting and applying LP solutions for refinery optimization is challenging due to simplifications and data errors.
method Transformed ECOD methodology, Anomaly Detection tools, and high-dimensional data analysis.
result Identifies data supply errors and reveals business opportunities in refinery scheduling and planning.

Study examines how different assessment formats affect student learning in a data communications course.

problem Understanding how various assessment formats impact student learning outcomes.
method Comparing student learning outcomes across multiple assessment formats in a core data communications course at George Mason University.
result Collective assessment formats enhance student knowledge demonstration.

Paper explores physics-informed deep learning for system reliability assessment.

problem Limited study on deep learning for system reliability assessment.
method Physics-informed deep learning approach for system reliability assessment.
result Physics-informed deep learning can alleviate computational challenges and combine measurement data and mathematical models.

Locally sparse neural networks improve interpretability for biomedical tabular data.

problem Overfitting and lack of interpretability in neural networks for tabular biomedical data.
method Locally sparse neural network with a gating network to select relevant features.
result The method outperforms state-of-the-art models in synthetic and real-world biomedical datasets.

chemmodlab simplifies fitting and assessing machine learning models in cheminformatics.

problem Comparing the utility of new machine learning models in cheminformatics.
method Streamlines model fitting and assessment pipeline, using k-fold cross-validation and multiplicity adjustments.
result Ease of presenting statistically significant performance differences among models.

Bayesian networks improve product risk assessment by handling uncertainty and causality.

problem Limited handling of uncertainty and inability to incorporate causal explanations in existing methods.
method Bayesian Networks (BNs) for improved systematic product risk assessment.
result BN approach provides more powerful and flexible risk assessments.

Enhances early risk assessments for pediatric outcomes using contrastive learning.

problem Improving risk assessments in early stages of pediatric development.
method Contrastive multi-modal framework that treats each time window as a distinct modality, training on all available data.
result Consistent improvements in early-stage risk assessments validated on real-world tasks.

Improved assessment of knee osteoarthritis using geodesic B-score.

problem Need for automatic, reader-independent measures of osteoarthritis clinical outcomes.
method Derive a geodesic B-score for Riemannian shape spaces, develop efficient algorithm for large shape populations.
result Geodesic B-score exhibits improved discrimination ability over Euclidean B-score.

Paper proposes EEIPU, a memoization-aware BO algorithm to reduce hyperparameter tuning costs.

problem High costs in GPU-days for training and fine-tuning language models.
method Memoization-aware Bayesian Optimization (EEIPU) algorithm in tandem with pipeline caching.
result EEIPU produces 103% more hyperparameter candidates and 108% more validation metric improvement.