Novel ECG classification for AF using spectro-temporal Kalman filtering and deep CNN.
On-device research index
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
169,181 papers · 148 categories
Trend · papers per month
8 results for “spectro-temporal”
problem Atrial fibrillation (AF) detection in ECG signals.
method Spectro-temporal representation using Kalman filter and deep convolutional neural networks.
result Proposed method achieves an overall F1 score of 80.2% on the PhysioNet/Computing in Cardiology (CinC) 2017 dataset.
New WildMix dataset and Spectro-Temporal Transformer model for better monoaural audio source separation.
problem Challenging monoaural audio source separation.
method Introducing WildMix dataset and Spectro-Temporal Transformer model (STT) with Spectro-Temporal Encoder (STE).
result STT outperforms previous baselines on the WildMix dataset.
Enhances speech in noisy environments using neural networks and NMF.
problem Speaker-independent multichannel speech enhancement in unknown noisy conditions.
method Uses variational autoencoders for supervised speech modeling and NMF for unsupervised noise modeling.
result The proposed approach outperforms NMF-based methods in noisy environments.
Improved TDNNs enhance speech recognition with deep kernels and frequency-dependent processing.
problem Shallow TDNN models limit long-term context modeling.
method Deepened kernels with residual connections and spectro-temporal processing.
result Deep kernel TDNNs reduce WER by 6% and further by 9% with frequency-dependent Grid-RNN.
New neural network design improves speech enhancement metrics.
problem Improving speech enhancement metrics in noisy conditions.
method Combination of convolutional and recurrent layers in U-net architecture.
result Proposed solution outperforms current state-of-the-art in SDR, SIR, and STOI metrics.
The paper explores modifications to filter banks for speech recognition.
problem Improving speech recognition accuracy using modified filter banks.
method The authors investigate replacing triangular filters with Gabor or Gammatone filters, and rearranging filter bank computations to integrate features over smaller time scales.
result No significant improvements in phone error rate were observed with the modifications.
Develops state-space deep Gaussian processes for irregular signals.
problem Solving deep Gaussian process regression problems for irregular signals/functions.
method Represent DGPs as SDEs, solve using state-space filtering and smoothing methods.
result Rich class of priors compatible with irregular signals/functions.
Deep neural network learns robust acoustic models from speech waveforms.
problem Robustness in speech recognition systems using standard feature extraction techniques.
method Deep convolutional neural network with stochastic variational inference and cosine modulated filters.
result Superior performance compared to baseline waveform-based models and deep CNNs with FBANK features.