Task focuses on fact checking in Q&A forums, improving over baseline systems.
problem Fact checking in community Q&A forums to distinguish factual from opinion.
method Two subtasks: distinguishing factual vs. opinion/advice/socializing, predicting answer truthfulness.
result Improved over baseline systems for both subtasks, but not for Subtask B.
Improved offensive language detection in tweets with multiple deep learning models.
problem Detecting offensive language in tweets using machine learning.
method Combination of multiple deep learning architectures for classification.
result Achieved macro-average F1-scores of 0.76, 0.68, 0.54 for different tasks.
System classifies Twitter and Reddit posts' stance towards hidden rumour threads.
problem Classifying posts' stance towards hidden rumour threads.
method Used pre-trained deep bidirectional transformers (BERT) for stance classification.
result Reached F1 score of 61.67% on test data, 2nd place in competition.
Ranked second in fact-checking task, using DRR NN with embeddings.
problem Fact-checking questions in community forums.
method Deeply Regularized Residual Neural Network (DRR NN) with Universal Sentence Encoder embeddings, ensemble methods.
result Ranked second in fact-checking task.
Team QCRI-MIT detects hyperpartisan news with 72.9% accuracy.
problem Detecting hyperpartisan news from biased political content.
method Logistic regression model using engineered features from propaganda detection.
result Significant performance improvements with better feature pre-processing.
RoBERTa model detects counterfactual statements in text.
problem Detecting and extracting counterfactual statements from text.
method Used RoBERTa language representation model for both subtasks.
result RoBERTa achieved top performance in both subtasks at SemEval-2020.
This paper describes the Amobee sentiment analysis system, adapted to compete in SemEval 2017 task 4. The system consists of two parts: a supervised training of RNN models based on a Twitter sentiment treebank, and the use of feedforward NN, Naive Bayes and logistic regression classifiers to produce predictions for the…
We describe the SemEval task of extracting keyphrases and relations between them from scientific documents, which is crucial for understanding which publications describe which processes, tasks and materials. Although this was a new task, we had a total of 26 submissions across 3 evaluation scenarios. We expect the tas…
Over 50 million scholarly articles have been published: they constitute a unique repository of knowledge. In particular, one may infer from them relations between scientific concepts, such as synonyms and hyponyms. Artificial neural networks have been recently explored for relation extraction. In this work, we continue…
This paper describes the participation of Amobee in the shared sentiment analysis task at SemEval 2018. We participated in all the English sub-tasks and the Spanish valence tasks. Our system consists of three parts: training task-specific word embeddings, training a model consisting of gated-recurrent-units (GRU) with …
In this paper we describe our attempt at producing a state-of-the-art Twitter sentiment classifier using Convolutional Neural Networks (CNNs) and Long Short Term Memory (LSTMs) networks. Our system leverages a large amount of unlabeled data to pre-train word embeddings. We then use a subset of the unlabeled data to fin…
Paper tackles counterfactual sentence detection and evaluation.
problem Detect and evaluate counterfactual sentences in natural language.
method Used a BERT base model for classification and a hybrid BERT Multi-Layer Perceptron for sequence identification. Introduced cascaded linear inputs to improve performance.
result Achieved an F1 score of 85.00% in Task 1 and 83.90% in Task 2.
We describe our language-independent unsupervised word sense induction system. This system only uses topic features to cluster different word senses in their global context topic space. Using unlabeled data, this system trains a latent Dirichlet allocation (LDA) topic model then uses it to infer the topics distribution…
Study categorizes and analyzes emotions in sexist tweets.
problem Lack of defined categories for sexism in NLP.
method Used a new dataset from SemEval-2018 to classify and analyze emotions in sexist tweets.
result Demonstrated the mental state and affectual state of users who tweet in different categories of sexism.
Replication confirms CGD's effectiveness in competitive games.
problem Reproducibility of a novel Nash equilibrium algorithm.
method Replicated experiments and provided Python implementation.
result CGD avoids oscillatory and divergent behaviours.
Abstracts index for ML4H workshop at NeurIPS 2019.
problem No specific problem stated; index of accepted abstracts.
method Not specified; index of accepted abstracts.
result No specific result stated.
VoxCeleb 2019 challenge assesses speaker recognition in uncontrolled settings.
problem Evaluate speaker recognition technology in unconstrained data.
method Public dataset, challenge, and workshop at Interspeech 2019.
result Baseline results and discussions provided.
Data set tracks real-time election results for 4 hours post-October 2019 Portuguese elections.
problem Real-time tracking of election results for predictive modeling.
method Real-time data collection and interval-based analysis.
result Data set supports various predictive modeling tasks including numerical forecasting.
Abstract geometric structures flow harmonically.
problem Geometric structures on Riemannian manifolds.
method Twistorial interpretation and abstract harmonicity condition.
result Established analytic properties of geometric gradient flow.
Interpretable semantic textual similarity (iSTS) task adds a crucial explanatory layer to pairwise sentence similarity. We address various components of this task: chunk level semantic alignment along with assignment of similarity type and score for aligned chunks with a novel system presented in this paper. We propose…
Recently, sentiment analysis has received a lot of attention due to the interest in mining opinions of social media users. Sentiment analysis consists in determining the polarity of a given text, i.e., its degree of positiveness or negativeness. Traditionally, Sentiment Analysis algorithms have been tailored to a speci…
Analyzed US firm data 1970-2019, identifying scale effects and distributional forms.
problem Understanding differences between small and large firms over time.
method Examined all public US firms, used stylized facts and DLN distribution analysis.
result Small firms are systematically different from large firms, with scale-dependent heteroskedasticity.
Two methods for quantile regression are compared and found to produce tighter intervals.
problem Comparing methods for producing prediction intervals in quantile regression.
method Two recently proposed methods combining conformal inference and quantile regression.
result Romano et al.'s method typically yields tighter prediction intervals in finite samples.
Investigates the relationship between US money supply and asset indices over 2001-2019.
problem Determining the relationship between US money supply and asset indices growth.
method Information entropy methodology applied to US asset indices (Property, Russell 2000, S&P 500, NASDAQ) over 2001-2019.
result Growth in US broad money supply is the main determinant of US asset indices growth, especially the NASDAQ and Russell 2000.
Study confirms improved performance of Self-Critique and Adapt method.
problem Improving performance of MAML++ method.
method Self-Critique and Adapt (SCA) method.
result SCA method improves performance of MAML++.
Abstract notes on robust statistical learning theory.
problem Developing robust estimators for statistical learning.
method Stressing principles of robust estimators construction and analysis.
result Emphasizes main principles of robust estimators construction and analysis.
Lectures on symplectic aspects of surface degenerations at KIAS.
problem Exploring symplectic structures in surface degenerations.
method Expository account of symplectic aspects of cyclic quotient surface singularities.
result Discussion of symplectic structures in surface degenerations.
Revisits causal inference identifiability with positivity assumption.
problem General identifiability in causal inference without positivity assumption.
method Introduces new algorithm sound and complete under positivity assumption.
result New algorithm connects general identifiability to classical identifiability.
Improved speech emotion recognition using pre-trained language models.
problem Challenging task of speech emotion recognition for natural human-machine interaction.
method Fine-tuning pre-trained language models for text emotion recognition, combining with speech emotion recognition.
result 73.5% accuracy in speech emotion recognition on a subset of IEMOCAP dataset.
The study of projective varieties with nef anticanonical divisors and log terminal singularities.
problem Understanding the structure and properties of projective varieties with specific divisor conditions.
method Analyzing the Albanese map and MRC fibration for klt projective varieties, showing locally constant fibrations and product decompositions.
result Generalization of results for smooth projective varieties to the klt case, including decomposition into rationally connected and projective varieties with trivial canonical divisor.
Reply to Ogburn et al. on their critique of Wang and Blei's work.
problem Critique of Wang and Blei's work on the blessings of multiple causes.
method Discussion and clarification of Wang and Blei's claims and findings.
result Wang and Blei's premise is correct and there are no foundational errors.
Study the Mexican stock market's interdependency structure from 2000-2019.
problem Characterize the interdependency structure of the Mexican Stock Exchange.
method Estimate correlation/concentration matrices from different models and compute network theory metrics.
result Visualizations provide a comprehensive overview of the stock market's interdependency structure.
New method calibrates eSSVI volatility surfaces without arbitrage.
problem Sequential calibration of eSSVI surfaces lacks global view and guarantees no arbitrage.
method Global and arbitrage-free parametrization of eSSVI surfaces.
result Faster calibration always guarantees an arbitrage-free fit.
Improved disentanglement in VAEs using aggregated feature maps.
problem Improving disentanglement in Variational Autoencoders (VAEs).
method Regionally aggregated feature maps extracted from pre-trained CNNs on ImageNet.
result 2nd place in NeurIPS 2019 disentanglement challenge.
Study optimizes trading strategies in markets with transaction costs and uncertain models.
problem Optimizing trading strategies in markets with transaction costs and model uncertainty.
method Maximizing worst-case expected utility over a class of models on a filtered probability space.
result Existence of optimal trading strategies for general càdlàg price processes and incomplete filtrations.
This paper covers the two approaches for sentiment analysis: i) lexicon based method; ii) machine learning method. We describe several techniques to implement these approaches and discuss how they can be adopted for sentiment classification of Twitter messages. We present a comparative study of different lexicon combin…
Detects out-of-distribution inputs in deep generative models.
problem Mismatch between model's typical set and high probability density areas.
method Statistically principled test using likelihood distribution.
result Successfully detects out-of-distribution sets in challenging cases.
New proof shows faster convergence rate for robust estimation with Lasso in adversarially contaminated outputs.
problem Robust estimation of parameters in the presence of adversarial output contamination.
method Extended Lasso with Huber loss function and L1 penalty, focusing on specific properties of the Huber function. result Same convergence rate as Dalalyan and Thompson (2019), but with a different proof.
Study improves CNNs for audio scene classification by restricting receptive fields and adding frequency awareness.
problem Improving CNNs for robust acoustic scene classification.
method Investigated different receptive field configurations for various CNN architectures and introduced Frequency Aware CNNs.
result Several well-performing submissions to DCASE 2019 Challenge were achieved.
AVEC 2019 challenges AI in detecting depression and cross-cultural emotions.
problem Detecting depression and cross-cultural emotions from audiovisual data.
method Comparison of machine learning methods under standardized conditions.
result Baseline system performance on state-of-mind, depression, and cross-cultural tasks.
Researchers improved Minecraft game performance using imitation learning.
problem Achieving state-of-the-art performance in immersive environments like Minecraft.
method Applied imitation learning to Minecraft, optimizing network architecture, loss function, and data augmentation.
result Reported stronger results than previous experiments, reaching second place in a competition.
A framework for multi-label sentiment analysis in 100 languages with dynamic weighting.
problem Cross-lingual sentiment analysis in multi-label settings with label imbalance.
method Dynamic weighting method, focal loss adaptation, optimal class-specific thresholds.
result State-of-the-art performance in 7 out of 9 metrics across 3 languages.
Paper proposes Experts Model for better emotion detection in tweets.
problem Estimating intensity of emotion in tweets.
method Inspired by Mixture of Experts (MoE) model, each expert learns different features.
result Our Experts Model stands at top-5 results in emotion detection.
NeurIPS 2019 program improves reproducibility in machine learning.
problem Ensuring machine learning research results are reproducible and reliable.
method Code submission policy, reproducibility challenge, and checklist integration.
result Improved reproducibility standards across the machine learning community.
Private classification and online prediction are shown to be equivalent.
problem Learning with differential privacy and online prediction equivalence.
method Introducing global stability and proving equivalence between online learnability and private PAC learnability.
result Every concept class with finite Littlestone dimension can be learned by a differentially-private algorithm.
The study compares different game-theoretic attribution methods and finds that interventional Shapley values yield less consistent results than Aumann-Shapley due to path symmetry.
problem Investigating the influence of path choice on game-theoretic attribution algorithms.
method Comparative analysis of interventional Shapley values and Generalized Integrated Gradients (GIG) methods.
result Interventional Shapley values yield less consistent attributions than Aumann-Shapley due to path symmetry and extended away from the training data manifold.
Polish hate speech detection team second in PolEval 2019.
problem Low-resource hate speech detection in Polish.
method Fine-tuning ULMFiT and BERT models, automated feature engineering with TPOT.
result TPOT's shallow model achieved second place in PolEval 2019.
Qwant Research improves clinical case matching and information retrieval.
problem Matching and retrieving relevant clinical cases and discussions.
method Approach based on language models and preprocessings, information extraction system using neural networks and linguistic analysis.
result Very encouraging results in information extraction accuracy.