Acoustic scene classification and related tasks have been dominated by Convolutional Neural Networks (CNNs). Top-performing CNNs use mainly audio spectograms as input and borrow their architectural design primarily from computer vision. A recent study has shown that restricting the receptive field (RF) of CNNs in appro…
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Enhances ASC using time- and frequency-liked CNNs and bilinear pooling.
Paper improves sound event detection using semi-supervised learning.
Distribution mismatches between the data seen at training and at application time remain a major challenge in all application areas of machine learning. We study this problem in the context of machine listening (Task 1b of the DCASE 2019 Challenge). We propose a novel approach to learn domain-invariant classifiers in a…
Improved acoustic scene classification with factorized CNN.
As part of the 2016 public evaluation challenge on Detection and Classification of Acoustic Scenes and Events (DCASE 2016), the second task focused on evaluating sound event detection systems using synthetic mixtures of office sounds. This task, which follows the `Event Detection - Office Synthetic' task of DCASE 2013,…
DCASE 2022 Task 2 tackles domain shifts in ASD for machine condition monitoring.
Paper proposes MTL for weakly labelled SED, improving performance with 2-step attention.
DCASE 2021 ASD task tackles domain-shifted anomalous sound detection.
To address Task 5 in the Detection and Classification of Acoustic Scenes and Events (DCASE) 2018 challenge, in this paper, we propose an ensemble learning system. The proposed system consists of three different models, based on convolutional neural network and long short memory recurrent neural network. With extracted …
CNNs improve generalization to unseen audio devices with increased width, not depth.
A new method detects anomalous sounds using self-supervised learning.
New dataset for industrial machine malfunction detection with domain shifts.
This paper describes Task 2 of the DCASE 2018 Challenge, titled "General-purpose audio tagging of Freesound content with AudioSet labels". This task was hosted on the Kaggle platform as "Freesound General-Purpose Audio Tagging Challenge". The goal of the task is to build an audio tagging system that can recognize the c…
Replication confirms CGD's effectiveness in competitive games.
Abstracts index for ML4H workshop at NeurIPS 2019.
VoxCeleb 2019 challenge assesses speaker recognition in uncontrolled settings.
System tackles indeterminacies in automated audio captioning.
We give a twistorial interpretation of geometric structures on a Riemannian manifold, as sections of homogeneous fibre bundles, following an original insight by Wood (2003). The natural Dirichlet energy induces an abstract harmonicity condition, which gives rise to a geometric gradient flow. We establish a number of an…
Data set tracks real-time election results for 4 hours post-October 2019 Portuguese elections.
This paper describes our system submitted to SemEval 2019 Task 7: RumourEval 2019: Determining Rumour Veracity and Support for Rumours, Subtask A (Gorrell et al., 2019). The challenge focused on classifying whether posts from Twitter and Reddit support, deny, query, or comment a hidden rumour, truthfulness of which is …
This paper introduces a model of environmental acoustic scenes which adopts a morphological approach by ab-stracting temporal structures of acoustic scenes. To demonstrate its potential, this model is employed to evaluate the performance of a large set of acoustic events detection systems. This model allows us to expli…
We compare two recently proposed methods that combine ideas from conformal inference and quantile regression to produce locally adaptive and marginally valid prediction intervals under sample exchangeability (Romano et al., 2019; Kivaranovic et al., 2019). First, we prove that these two approaches are asymptotically ef…
Analyzed US firm data 1970-2019, identifying scale effects and distributional forms.
We describe a novel weakly labeled Audio Event Classification approach based on a self-supervised attention model. The weakly labeled framework is used to eliminate the need for expensive data labeling procedure and self-supervised attention is deployed to help a model distinguish between relevant and irrelevant parts …
Paper discusses ASD challenge for machine condition monitoring.
Investigates the relationship between US money supply and asset indices over 2001-2019.
Study confirms improved performance of Self-Critique and Adapt method.
Lectures on symplectic aspects of surface degenerations at KIAS.
Revisits causal inference identifiability with positivity assumption.
Recent work has shown that deep generative models can assign higher likelihood to out-of-distribution data sets than to their training data (Nalisnick et al., 2019; Choi et al., 2019). We posit that this phenomenon is caused by a mismatch between the model's typical set and its areas of high probability density. In-dis…
The study of projective varieties with nef anticanonical divisors and log terminal singularities.
Reply to Ogburn et al. on their critique of Wang and Blei's work.
Study the Mexican stock market's interdependency structure from 2000-2019.
New method calibrates eSSVI volatility surfaces without arbitrage.
Improved disentanglement in VAEs using aggregated feature maps.
Study optimizes trading strategies in markets with transaction costs and uncertain models.
The article describes our submission to SemEval 2019 Task 8 on Fact-Checking in Community Forums. The systems under discussion participated in Subtask A: decide whether a question asks for factual information, opinion/advice or is just socializing. Our primary submission was ranked as the second one among all participa…
Acoustic scene classification is the task of identifying the scene from which the audio signal is recorded. Convolutional neural network (CNN) models are widely adopted with proven successes in acoustic scene classification. However, there is little insight on how an audio scene is perceived in CNN, as what have been d…
New proof shows faster convergence rate for robust estimation with Lasso in adversarially contaminated outputs.
Acoustic Scene Classification (ASC) is a challenging task, as a single scene may involve multiple events that contain complex sound patterns. For example, a cooking scene may contain several sound sources including silverware clinking, chopping, frying, etc. What complicates ASC more is that classes of different activi…
In this paper we propose a novel environmental sound classification approach incorporating unsupervised feature learning from codebook via spherical -Means++ algorithm and a new architecture for high-level data augmentation. The audio signal is transformed into a 2D representation using a discrete wavelet transform …
This paper reports on Qwant Research contribution to tasks 2 and 3 of the DEFT 2019's challenge, focusing on French clinical cases analysis. Task 2 is a task on semantic similarity between clinical cases and discussions. For this task, we propose an approach based on language models and evaluate the impact on the resul…
In this paper, we establish the stochastic ordering of the Gini indexes for multivariate elliptical risks which generalized the corresponding results for multivariate normal risks. It is shown that several conditions on dispersion matrices and the components of dispersion matrices of multivariate normal risks for the m…
This report contains the details regarding our submission to the OffensEval 2019 (SemEval 2019 - Task 6). The competition was based on the Offensive Language Identification Dataset. We first discuss the details of the classifier implemented and the type of input data used and pre-processing performed. We then move onto…
We present seven myths commonly believed to be true in machine learning research, circa Feb 2019. This is an archival copy of the blog post at https://crazyoscarchang.github.io/2019/02/16/seven-myths-in-machine-learning-research/ Myth 1: TensorFlow is a Tensor manipulation library Myth 2: Image datasets are representat…
Researchers improved Minecraft game performance using imitation learning.
Question semantic similarity (Q2Q) is a challenging task that is very useful in many NLP applications, such as detecting duplicate questions and question answering systems. In this paper, we present the results and findings of the shared task (Semantic Question Similarity in Arabic). The task was organized as part of t…