Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

6121,2241,8352,447 · Jun 202019922001200920172026
48 results for bias in bios

Study shows gender bias in occupation classification tasks.

problem Gender bias in machine learning for occupation classification.
method Analyzed impact of explicit gender indicators in semantic representations of biographies.
result True positive rates differ between genders, correlating with existing gender imbalances.

This paper describes a time-series-based classification approach to identify similarities between bio-medical-based situations. The proposed approach allows classifying collections of time-series representing bio-medical measurements, i.e., situations, regardless of the type, the length and the quantity of the time-ser…

2013-03-01abs ↗pdf ↗

This paper shows how learning the phase-amplitude coupling improves bio-signal classification.

problem Discarding phase component in bio-signal feature extraction leads to poor generalization.
method Introducing a novel self-supervised learning task called Phase-Swap to detect phase-amplitude coupling.
result Neural networks trained on Phase-Swap task generalize better across subjects and recording sessions.

Improved bio-surveillance through automated document classification.

problem Tracking infectious diseases across global news alerts.
method Recurrent neural networks, TF-IDF, Naive Bayes, logistic regression.
result 97% recall and 93.3% accuracy in bio-surveillance event classification.

Two case studies reveal hidden biases and confounders in machine learning models of biomedical data.

problem Hidden biases and confounders in machine learning models of biomedical data.
method Two case studies examining biases and confounders in machine learning models of biomedical data.
result Prediction models performed well but hidden biases and confounders were revealed.

Method reduces bias in occupation classification without protected attribute data.

problem Mitigating bias in occupation classification without access to protected attributes.
method Uses word embeddings to discourage correlation between predicted occupation probability and name.
result Reduces race and gender biases without significant loss in true positive rate.

This work applies deep learning to bio-sensing and video data for affective computing.

problem Lack of deep learning integration in bio-sensing for affective computing.
method Novel deep-learning-based methods applied to EEG, ECG, and video data.
result Outperforms other studies in emotion/valence/arousal/liking classification.

This paper analyzes the impact of loops on bilevel optimization efficiency.

problem The impact of loops on the efficiency of bilevel optimization algorithms.
method Unified convergence analysis and computational complexity characterization for AID-BiO and ITD-BiO with and without loops.
result Loops in bilevel optimization can improve overall efficiency but increase per-step complexity.

The purpose of this paper is to study the shapes and stabilities of bio-membranes within the framework of exterior differential forms. After a brief review of the current status in theoretical and experimental studies on the shapes of bio-membranes, a geometric scheme is proposed to discuss the shape equation of closed…

2004-03-12abs ↗pdf ↗

Novel bio-inspired masking for robust speech emotion recognition.

problem Noise degradation in speech emotion recognition.
method Cochlear cepstrogram-based contrastive learning with temporal and frequency masking.
result Improved speech emotion recognition performance on K-EmoCon benchmark.

Optimal transport strategy reduces gender bias in job recommendation systems.

problem Mitigating gender biases in AI-driven job recommendation systems.
method Model agnostic optimal transport strategy applied to multi-class neural networks.
result Reduced undesirable algorithmic biases in job recommendation tasks.

Study proposes a new approval policy for ML-based medical devices to prevent gradual performance degradation.

problem Gradual deterioration in machine learning model performance over time in medical devices.
method Formulated an automatic algorithmic change protocol (aACP) as an online hypothesis testing problem, considering both error-rate guarantees and non-guaranteed policies.
result Controlled the rate of gradual deterioration (biocreep) in machine learning models without significantly impacting approval of beneficial modifications.

The paper proposes a bio-inspired framework for better compression and adversarial robustness in machine learning models.

problem Machine learning models are vulnerable to adversarial examples.
method The paper introduces a bio-inspired classification framework that conditions model inference on label hypothesis and uses an information bottleneck regularizer.
result The framework enables better compression and adversarial robustness without loss of natural accuracy.

Paper improves Tm prediction of protein fragments using sparsity and probabilistic models.

problem Improving accuracy of melting temperature prediction for protein fragments.
method Promoting sparsity in pre-trained transformer models and adopting probabilistic frameworks.
result Mean absolute error of 0.23C for predicting melting temperature.

Optimizes biomanufacturing processes with a new digital twin calibration method.

problem Lack of interpretability and sample efficiency in traditional DoE methods.
method Developed a computational approach to calibrate Bio-SoS digital twin model.
result Guides sample-efficient and interpretable DoEs by quantifying sub-model parameter estimation errors.

Graph neural network predicts protonation energies of oxygen atoms in bio-oil molecules.

problem Predicting protonation energies of oxygen atoms in bio-oil molecules for chemical upgrading.
method Site-specific graph neural network approach using iterative local nonlinear embedding.
result Effective prediction of protonation energies of individual oxygen atoms in bio-oil molecules.

Derives a biologically plausible neural network for Slow Feature Analysis.

problem Learning latent features from time series data.
method Starting from an SFA objective, derives Bio-SFA with a biologically plausible neural network implementation.
result Validates Bio-SFA on naturalistic stimuli, reproducing interesting properties of brain cells.

CHANI learns classification tasks with local transformations inspired by biology.

problem Proving neural networks can learn classification tasks with local transformations.
method CHANI uses spiking neurons modeled by Hawkes processes with expert aggregation for local learning.
result CHANI can learn and encode multiple classes, forming assemblies of neurons.

A new method trains a smaller model from a larger one without needing the actual training data.

problem Training a smaller model from a larger one without access to the training data.
method Synthesizes data impressions from the Teacher model to train the Student model.
result Zero-Shot Knowledge Distillation achieves competitive generalization performance.

New training algorithm enhances SNNs for temporal signal processing.

problem Lack of robust training algorithms for large-scale SNNs.
method Formulated SNN as IIR filters, proposed training algorithm for optimal synapse filter kernels and weights.
result Model and training algorithm outperform state-of-the-art approaches in accuracy.

Bio-inspired neural networks use predictive coding for efficient weight updates.

problem Training artificial neural networks efficiently and biologically plausibly.
method Predictive Coding (PC) updates weights locally using only local information.
result PC provides theoretical advantages like automatic gradient scaling.

Structural Causal Models (SCMs) provide a popular causal modeling framework. In this work, we show that SCMs are not flexible enough to give a complete causal representation of dynamical systems at equilibrium. Instead, we propose a generalization of the notion of an SCM, that we call Causal Constraints Model (CCM), an…

2018-05-16abs ↗pdf ↗

Modern bio-technologies have produced a vast amount of high-throughput data with the number of predictors far greater than the sample size. In order to identify more novel biomarkers and understand biological mechanisms, it is vital to detect signals weakly associated with outcomes among ultrahigh-dimensional predictor…

2018-05-17abs ↗pdf ↗

The personalization of treatment via bio-markers and other risk categories has drawn increasing interest among clinical scientists. Personalized treatment strategies can be learned using data from clinical trials, but such trials are very costly to run. This paper explores the use of active learning techniques to desig…

2012-02-14abs ↗pdf ↗

The chapter improves deep learning models by interpreting and improving their performance.

problem Deep learning models often lack interpretability, leading to poor understanding of their predictions.
method The approach involves attributing importance to features and feature groups, including interactions, to improve model performance.
result The proposed attributions provide insights across various domains and can be used to improve model generalization.

Bio-inspired neuromorphic hardware is a research direction to approach brain's computational power and energy efficiency. Spiking neural networks (SNN) encode information as sparsely distributed spike trains and employ spike-timing-dependent plasticity (STDP) mechanism for learning. Existing hardware implementations of…

2018-09-18abs ↗pdf ↗

A new flow-based model for molecular graphs achieves better performance with fewer parameters.

problem Generating molecular graphs efficiently and accurately.
method Graph residual flow (GRF) based on residual flows for molecular graphs, with invertibility conditions derived.
result The GRF model achieves comparable performance to existing models with significantly fewer parameters.

It has been noticed that some external CVIs exhibit a preferential bias towards a larger or smaller number of clusters which is monotonic (directly or inversely) in the number of clusters in candidate partitions. This type of bias is caused by the functional form of the CVI model. For example, the popular Rand index (R…

2016-06-17abs ↗pdf ↗

Learning representation for graph classification turns a variable-size graph into a fixed-size vector (or matrix). Such a representation works nicely with algebraic manipulations. Here we introduce a simple method to augment an attributed graph with a virtual node that is bidirectionally connected to all existing nodes…

2017-08-14abs ↗pdf ↗