AI tested on 10 math questions from research.
problem Assessing AI's ability to solve research-level math problems.
method Shared 10 math questions not previously publicly available.
result Answers to questions are known to authors but encrypted.
This study analyzes financial equity research reports to identify frequently asked questions and automates 80% of them.
problem Insufficient empirical analysis of questions answered in financial equity research reports.
method Analyzed 72 financial equity research reports, classifying sentences into 169 unique question archetypes. Used public corporate reports to classify questions' potential for automation.
result Approximately 80% of financial equity research reports can be automated, with 78.7% of questions automatable.
Reduce survey questions to scale market research without annoying customers.
problem Performing market research by surveying customers with many questions is inefficient and annoying.
method Used Bayesian networks to model and reduce the number of questions asked to customers.
result Demonstrated the effectiveness of the approach using an example of broadband customer segmentation.
20 questions to improve AI research transparency, replicability, ethics, and effectiveness.
problem Lack of transparency, replicability, ethical concerns, and effectiveness in AI research.
method Presenting 20 questions to guide project planning and post-hoc evaluation.
result Facilitating a discussion to develop an international consensus framework.
NukeBERT improves performance on nuclear domain Q&A with less training data.
problem Lack of annotated data for nuclear domain Q&A.
method Developed NQuAD dataset and NukeBERT model incorporating novel BERT vocabulary technique.
result NukeBERT outperformed BERT significantly on NQuAD.
Which topics of machine learning are most commonly addressed in research? This question was initially answered in 2007 by doing a qualitative survey among distinguished researchers. In our study, we revisit this question from a quantitative perspective. Concretely, we collect 54K abstracts of papers published between 2…
Unified framework for simple question answering using subgraph ranking and joint-scoring.
problem Simple question answering with knowledge graphs is challenging.
method Unified framework focusing on subgraph selection and fact selection, with novel ranking and joint-scoring methods.
result Achieved state-of-the-art accuracy of 85.44% on SimpleQuestions dataset.
CLEAR dataset for acoustic reasoning tasks.
problem Acoustic reasoning and question answering.
method Data generation from elementary sounds, functional programs for question composition.
result Validation of current state-of-the-art visual Q&A models on AQA task.
Machine learning uses crowdworkers; determining their status as human subjects is tricky.
problem Determining the appropriate status of ML crowdworkers as human subjects.
method Investigation of natural language processing studies to expose challenges and propose solutions.
result Potential loophole in the U.S. Common Rule for ML research oversight.
Paper evaluates Arabic question similarity, 9 teams participated.
problem Predicting semantic similarity between Arabic questions.
method 9 teams participated in a shared task to predict semantic similarity.
result 9 teams participated in the task, results made publicly available.
Researchers provide counterexamples to Pogorelov and Toponogov's questions about saddle surfaces.
problem Existence of closed asymptotic curves in saddle surfaces with nonpositive curvature.
method Explicit counterexamples and corrected statements of the original questions.
result Proved a corrected version of the original questions.
The primary goal of this study is doing a meta-analysis research on two groups of published studies. First, the ones that focus on the evaluation of the United States Department of Agriculture (USDA) forecasts and second, the ones that evaluate the market reactions to the USDA forecasts. We investigate four questions. …
Paper tackles medical question similarity using domain-relevant embeddings.
problem Identifying same-question pairs in medical contexts.
method Semi-supervised pre-training of a neural network on medical question-answer pairs.
result Our model achieves 82.6% accuracy on medical question similarity task.
One long-term goal of machine learning research is to produce methods that are applicable to reasoning and natural language, in particular building an intelligent dialogue agent. To measure progress towards that goal, we argue for the usefulness of a set of proxy tasks that evaluate reading comprehension via question a…
Researchers found non-smoothable surfaces in a 4-sphere, solving K3 problems.
problem Non-smoothable surfaces in the 4-sphere.
method Constructed non-orientable surfaces with specific knot groups.
result Found surfaces that are non-smoothable and answered K3 problems.
The present research aims to highlight the main factors influencing the development of entrepreneurial innovation in a rural environment and to perform an empirical study with the purpose of assessing the main problems in rural development. The research performed is mostly of a quantitative nature, being based on the u…
Paper analyzes biases in video QA datasets, showing models can answer 37-48% questions correctly without multimodal context.
problem Question answering biases in video QA datasets can lead to model overfitting and poor generalization.
method Analyzed popular video question answering datasets, conducted ablation studies on biases from annotators and question types.
result Pretrained language models can answer 37-48% questions correctly without multimodal context, far exceeding random guess baseline.
New task AQA tackles acoustic reasoning from sound scenes.
problem Promote research in acoustic reasoning.
method Generate acoustic scenes from elementary sounds and formulate questions.
result Preliminary results with models FiLM and MAC show promise.
Research explores how interconnected systems synchronize and how to control their behavior.
problem Understanding and controlling the behavior of interconnected dynamical systems.
method Mean field games approach applied to controlled coupled oscillators.
result Developed methods to predict and influence emergent phenomena in interconnected systems.
Synthetic data can be used to ask more questions and accelerate discovery with provable validity guarantees.
problem Valid inference with synthetic data
method Task exchangeability
result Provable validity guarantees for synthetic data inference
Survey of open problems linking integrable systems and Nijenhuis geometry.
problem Interplay between integrable systems and Nijenhuis geometry.
method Discussion of open problems and questions.
result Challenges and simple questions in the field.
In this paper we sketch some reflections on the pitfalls and inconsistencies of the research program - currently dominant among the profession - aimed at providing microfoundations to macroeconomics along a Walrasian perspective. We argue that such a methodological approach constitutes an unsatisfactory answer to a wel…
Paper tackles Arabic question similarity, outperforming state-of-the-art.
problem Detecting semantically similar questions in Arabic is challenging.
method Utilizes contextualized word representations (ELMo embeddings) trained on MSA and dialectic sentences, combined with a pairwise similarity layer.
result Achieves 93% F1-score on Modern Standard Arabic benchmark and 82% on dialectical benchmark.
Survey of PU learning methods for positive and unlabeled data.
problem Learning from only positive and unlabeled data.
method Analyzes various approaches to PU learning.
result Identifies seven key research questions in PU learning.
Expands robust profit opportunities to include distributional uncertainty.
problem Distributional uncertainty in financial markets.
method Formulates infinite dimensional primal problems, simplifies to finite dimensional dual problems using Wasserstein distance.
result Distributional uncertainty can enhance robustness of profit opportunities.
LSTM network aids intent classification in QA.
problem Classifying intent in question-answering.
method Used LSTM architecture for intent classification.
result Effective and efficient intent classification achieved.
A new model answers questions about medical images.
problem Lack of transparency in deep learning models for medical imaging.
method A question-centric model that queries image models directly.
result The model achieves equal or higher accuracy than existing methods.
New approach to fairness in machine learning using contrastive questions.
problem Ensuring fairness in algorithmic decision-making.
method Causal inference to address contrastive fairness.
result Mathematical tools for contrastive fairness in machine learning.
Survival analysis models research reproducibility, offering new insights.
problem Reproducibility crisis in machine learning research.
method Survival analysis to model reproducibility as a continuous process.
result Survival analysis provides deeper insights into research reproducibility.
Researchers analyze tagging patterns on Stack Exchange communities.
problem Understanding the structure and evolution of tags in Q&A platforms.
method Empirical analysis and development of a generative model for tag co-occurrence.
result The model can reproduce statistical properties of co-tagging graphs.
Current XAI research lacks solid foundations and clear goals.
problem Inadequate conceptual, ethical, and methodological foundations in XAI research.
method Discussion of misconceptions and suggestions for improvement.
result Current XAI research needs to address conceptual, ethical, and methodological issues.
WiseMove framework for safe deep RL in autonomous driving.
problem Ensuring safety in deep reinforcement learning for autonomous driving.
method Modular learning architecture for motion planning.
result Demonstrated on a common traffic scenario, WiseMove supports safe learning.
Traditionally it had been a problem that researchers did not have access to enough spatial data to answer pressing research questions or build compelling visualizations. Today, however, the problem is often that we have too much data. Spatially redundant or approximately redundant points may refer to a single feature (…
The paper explores arbitrage in financial markets under uncertainty using Wasserstein distance.
problem Investigating arbitrage in financial markets with distributional uncertainty.
method Using Wasserstein distance, the paper considers weak and strong forms of arbitrage conditions and introduces a relaxation called statistical arbitrage.
result The paper derives dual formulations of robust arbitrage conditions and conducts computational experiments to answer questions about ambiguity and statistical arbitrage.
EasyTime simplifies time series forecasting for researchers and practitioners.
problem Ease of use and accuracy in time series forecasting.
method One-click evaluation, automated ensemble, natural language Q&A.
result Superior forecasting accuracy compared to individual methods.
Comprehensive review of robust portfolio selection models.
problem Addressing uncertainty in financial portfolio optimization.
method Classification and analysis of various models and approaches.
result Identification of open research questions.
FinanceHarness automates financial deep research with specialized tools and benchmarks.
problem Lack of specialized financial deep research for financial analysis.
method FinanceHarness uses a layered harness and FinanceGym for specialized research questions and validation.
result FinanceHarness improves rubric scores by 7.1% and achieves 82% pass rate with expert validation.
This paper is a survey on the {\em Zimmer program}. In it's broadest form, this program seeks an understanding of actions of large groups on compact manifolds. The goals of this survey are (1) to put in context the original questions and conjectures of Zimmer and Gromov that motivated the program, (2) to indicate t…
Most research in reading comprehension has focused on answering questions based on individual documents or even single paragraphs. We introduce a neural model which integrates and reasons relying on information spread within documents and across multiple documents. We frame it as an inference problem on a graph. Mentio…
Digital personas improve survey results for stable attributes but fail for subjective responses.
problem When can digital personas reliably approximate human survey findings?
method Using LISS panel, constructed personas from background variables and survey histories, tested against held-out post-cutoff answers.
result Digital personas improve alignment with human response distributions for stable attributes but fail for subjective responses.
New math for deep learning tackles key questions about neural networks.
problem Understanding the exceptional performance of deep learning models.
method Analyzing overparametrized neural networks, depth, optimization, feature learning, and architecture effects.
result Partial answers to deep learning's outstanding generalization and optimization performance.
Maps of degree 1 and critical points on manifolds are studied.
problem Evaluate the minimal number of critical points on closed smooth manifolds.
method Examined maps of degree 1 and considered high-dimensional examples.
result In dimension 3 or less, CritM≥CritN for maps of degree 1. AI models help solve a symplectic topology problem.
problem Lagrangian smoothability question from Abouzaid et al.
method Used large language models (LLM) to approach the problem.
result Suggested new constructions in polyhedral symplectic topology.
This review explores ChatGPT in accounting and finance.
problem Understanding the current state of research on ChatGPT in accounting and finance.
method A scoping review of recent publications and working papers.
result Identifies three themes: applications, research tools, and implications.
Predicts student performance in interactive online question pools using GNNs.
problem Predicting student performance in interactive online question pools with evolving knowledge.
method Proposes R^2GCN, a GNN model for heterogeneous networks to predict student performance.
result Achieves higher accuracy in student performance prediction than traditional methods.
Many Machine Reading and Natural Language Understanding tasks require reading supporting text in order to answer questions. For example, in Question Answering, the supporting text can be newswire or Wikipedia articles; in Natural Language Inference, premises can be seen as the supporting text and hypotheses as question…
Probabilistic models can handle causal inference without special tools.
problem Confusion over necessary tools for causal inference.
method Demonstrated through concrete examples that causal questions can be answered using standard probabilistic models.
result Causal questions can be addressed using standard probabilistic modelling and inference.
The Johansen-Ledoit-Sornette (JLS) model of rational expectation bubbles with finite-time singular crash hazard rates has been developed to describe the dynamics of financial bubbles and crashes. It has been applied successfully to a large variety of financial bubbles in many different markets. Having been developed fo…