Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,291 papers · 148 categories

Trend · papers per month

9172634 · Jun 202019922001200920182026
48 results for sound interface

Research maps body movements to sound using advanced algorithms.

problem Detect and classify user commands with minimal latency.
method Vector Autoregressive Hierarchical Hidden Markov Models (VAR-HHMM) with wireless sensor nodes.
result VAR-HHMM algorithm outperforms K-nearest neighbors in detection and classification.

Paper introduces probabilistic module interface for complex models and inference.

problem Handling complex probabilistic models with latent variables and custom inference methods.
method Develops a platform-agnostic interface for encapsulating models and inference programs, allowing sound approximate inference algorithms for networks of modules.
result Sound approximate inference algorithms can be constructed for networks of probabilistic modules.

Deep learning maps tongue movements to speech sounds for voiceless individuals.

problem Developing silent speech interfaces for individuals without a larynx.
method Hybrid spatio-temporal 3D convolutions and feature shuffling for formant estimation and tracking from ultrasound tongue images.
result Best model achieves R-squared of 99.96% for vowel formant regression.

BC learning improves deep sound recognition performance.

problem Improving deep sound recognition using novel training data.
method BC learning: mixing sounds from different classes to generate between-class sounds and train models to recognize these.
result BC learning improves performance on various sound recognition networks, surpassing human level.

This paper improves universal sound separation using sound classification.

problem Separating acoustic sources from an open domain, regardless of their class.
method Utilizing semantic embeddings from a sound classifier to condition a separation network.
result Classifier embeddings provide nearly one dB of SNR gain, and iterative models achieve significant performance.

Insurance contracts for autonomous AI agents must be actuarially sound and resistant to gaming.

problem Designing insurance contracts for autonomous AI agents that are actuarially sound and resistant to gaming.
method Characterizing a five-attack space and proving the actuarial runtime is gaming-resistant.
result An incentive-compatible layer for actuarial control of autonomous-agent side effects.

Paper introduces ToyADMOS dataset for detecting anomalous machine sounds.

problem Lack of large-scale datasets for ADMOS anomaly detection.
method Collected anomalous sounds of miniature machines by deliberate damage.
result Released dataset includes over 180 hours of normal and 4,000 anomalous sounds.

Paper proposes a method to detect unknown anomalous sounds without training data using deep learning and Neyman-Pearson lemma.

problem Unsupervised detection of unknown anomalous sounds in audio data.
method Uses an autoencoder to minimize reconstruction error of normal sounds and Neyman-Pearson lemma to maximize true positive rate under low false positive rate conditions.
result The proposed method improves performance measures of unsupervised anomaly detection in audio data under low false positive rate conditions.

Batch uniformization improves anomaly detection in sound data.

problem Anomaly scores for rare and frequent normal sounds are not uniform.
method Propose batch uniformization to minimize anomaly scores by weighting samples based on their density.
result Improves performance of unsupervised anomaly detection in sound data.

Paper presents an audiovisual model to recognize sounds from weakly labeled video data.

problem Sound recognition from weakly labeled video data.
method Audiovisual fusion model with attention mechanism.
result The model achieves a mean Average Precision (mAP) of 46.16 on AudioSet, outperforming state-of-the-art models.

Neural network synthesizes percussive sounds with adjustable timbral features.

problem Control over high-level timbral characteristics of percussive sounds.
method Feedforward convolutional neural network mapping input parameters to waveform.
result Changing input parameters produces a waveform congruent with desired characteristics.

CNN improves whale sound detection in noisy environments.

problem Automatically detecting humpback whale sounds in complex background noises.
method Used Convolution Neural Network (CNN) for bi-class classification.
result CNN features outperformed traditional spectrogram methods in detecting whale sounds.

Deep learning models improve sound separation across various types of sounds.

problem Developing a universal method to separate arbitrary sounds of different types.
method Created a dataset of mixtures containing arbitrary sounds, investigated mask-based separation architectures, and tested different framewise analysis-synthesis bases.
result STFT outperformed learnable bases in universal sound separation tasks.

Improved neural network detects heart sounds with 87.5% accuracy from noisy recordings.

problem Detecting cardiac abnormalities from noisy heart sound recordings.
method Segmental Convolutional Neural Network (CNN) architecture trained on noisy recordings.
result Best model achieved 87.5% accuracy on PhysioNet/CinC Challenge dataset.

Paper proposes active learning for sound event detection with reduced annotation effort.

problem Reducing annotation effort for sound event detection.
method Change point detection for candidate selection, mismatch-first farthest-traversal for selection, training with context recordings.
result The proposed system achieves similar performance to full annotation with only 2% of data, reducing annotation effort.

This paper analyzes sound event detection in synthetic office audio, comparing different systems.

problem Comparing sound event detection systems in synthetic office audio.
method Analysis of systems submitted to DCASE 2016 task, using synthetic office sounds.
result Statistical analysis of results, highlighting system performance under controlled conditions.

The paper presents a method for sound event localization and detection using CRNN models.

problem Sound event localization and detection in complex environments.
method Consecutive ensemble of CRNN models for estimating event onset, offset, direction of arrival, and classification.
result The proposed method outperforms other participants in the DCASE2019 task3.

Extends diffuse interface methods to graphs and hypergraphs with non-smooth potentials.

problem Semi-supervised learning on graphs and hypergraphs.
method Generalizes diffuse interface methods using non-smooth potential functions and hypergraph Laplacians.
result The diffuse interface method can be applied to both graph and hypergraph data.

Paper discusses ASD challenge for machine condition monitoring.

problem Detecting unknown anomalous sounds without labeled data.
method Design and evaluation of a large-scale ASD dataset, novel approaches.
result Several novel approaches developed, evaluation results analyzed.

New interface explains contextual bandits to non-experts.

problem Interpreting and managing contextual bandits for non-expert operators.
method Developed a metric 'value gain' for off-policy evaluation and designed an interface to explain bandit behavior.
result Empowered non-experts to manage complex machine learning systems through accessible presentation.

A multi-head attention network improves ASC by recognizing overlapping sound patterns.

problem Challenging ASC due to overlapping sound patterns and complex event mixtures.
method Proposes a multi-head attention network to model complex temporal input structures.
result Achieved competitive performance on DCASE 2018 Task 5 dataset.

Paper detects adversarial attacks in sound classification models.

problem Adversarial attacks threaten data-driven models, especially in sound classification.
method Detects adversarial subspaces in unitary vector domain using chordal distance and generalized Schur decomposition.
result Regularized logistic regression detector outperforms other approaches on benchmark datasets.

Adaptive pooling operators improve sound event detection with weak labels.

problem Efficiently label audio recordings with weakly annotated sound sources.
method Developed adaptive pooling operators for multiple instance learning.
result Adaptive pooling operators outperform non-adaptive methods on static predictions and nearly match strong annotations.

Generative replay extends sound classification models to new classes without old data.

problem Incrementally refining a sound classifier with new data causes previously learned tasks to degrade.
method Developed a generative replay procedure to generate training data in place of older datasets.
result Generative replay with 4% of old data performs as well as keeping 20% of old data.

InverSynth automatically tunes synthesizer parameters from audio input.

problem Manual tuning of synthesizer parameters is time-consuming and requires expertise.
method Strided convolutional neural networks for inferring synthesizer parameters.
result InverSynth outperforms baselines in synthesizer parameter tuning.

We consider the regularity of an interface between two incompressible and inviscid fluids flows in the presence of surface tension. We obtain local in time estimates on the interface in H32k+1H^{\frac32k +1} and the velocity fields in H32kH^{\frac32k}. These estimates are obtained using geometric considerations which show th…

2006-09-20abs ↗pdf ↗

Study boundary behavior of limit interfaces in Riemannian manifolds without convexity assumptions.

problem Boundary behavior of limit interfaces in Riemannian manifolds.
method Proves limit-interface is a free boundary varifold, integer rectifiable up to boundary.
result No convexity assumption required; valid even when limit-interface clusters near boundary.