Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

25.0%50.0%75.0%100.0% · Feb 199419922001200920182026
48 results for research questions

AI tested on 10 math questions from research.

problem Assessing AI's ability to solve research-level math problems.
method Shared 10 math questions not previously publicly available.
result Answers to questions are known to authors but encrypted.

This study analyzes financial equity research reports to identify frequently asked questions and automates 80% of them.

problem Insufficient empirical analysis of questions answered in financial equity research reports.
method Analyzed 72 financial equity research reports, classifying sentences into 169 unique question archetypes. Used public corporate reports to classify questions' potential for automation.
result Approximately 80% of financial equity research reports can be automated, with 78.7% of questions automatable.

Reduce survey questions to scale market research without annoying customers.

problem Performing market research by surveying customers with many questions is inefficient and annoying.
method Used Bayesian networks to model and reduce the number of questions asked to customers.
result Demonstrated the effectiveness of the approach using an example of broadband customer segmentation.

20 questions to improve AI research transparency, replicability, ethics, and effectiveness.

problem Lack of transparency, replicability, ethical concerns, and effectiveness in AI research.
method Presenting 20 questions to guide project planning and post-hoc evaluation.
result Facilitating a discussion to develop an international consensus framework.

Unified framework for simple question answering using subgraph ranking and joint-scoring.

problem Simple question answering with knowledge graphs is challenging.
method Unified framework focusing on subgraph selection and fact selection, with novel ranking and joint-scoring methods.
result Achieved state-of-the-art accuracy of 85.44% on SimpleQuestions dataset.

Machine learning uses crowdworkers; determining their status as human subjects is tricky.

problem Determining the appropriate status of ML crowdworkers as human subjects.
method Investigation of natural language processing studies to expose challenges and propose solutions.
result Potential loophole in the U.S. Common Rule for ML research oversight.

The primary goal of this study is doing a meta-analysis research on two groups of published studies. First, the ones that focus on the evaluation of the United States Department of Agriculture (USDA) forecasts and second, the ones that evaluate the market reactions to the USDA forecasts. We investigate four questions. …

2018-01-19abs ↗pdf ↗

Paper analyzes biases in video QA datasets, showing models can answer 37-48% questions correctly without multimodal context.

problem Question answering biases in video QA datasets can lead to model overfitting and poor generalization.
method Analyzed popular video question answering datasets, conducted ablation studies on biases from annotators and question types.
result Pretrained language models can answer 37-48% questions correctly without multimodal context, far exceeding random guess baseline.

Research explores how interconnected systems synchronize and how to control their behavior.

problem Understanding and controlling the behavior of interconnected dynamical systems.
method Mean field games approach applied to controlled coupled oscillators.
result Developed methods to predict and influence emergent phenomena in interconnected systems.

In this paper we sketch some reflections on the pitfalls and inconsistencies of the research program - currently dominant among the profession - aimed at providing microfoundations to macroeconomics along a Walrasian perspective. We argue that such a methodological approach constitutes an unsatisfactory answer to a wel…

2006-08-14abs ↗pdf ↗

Paper tackles Arabic question similarity, outperforming state-of-the-art.

problem Detecting semantically similar questions in Arabic is challenging.
method Utilizes contextualized word representations (ELMo embeddings) trained on MSA and dialectic sentences, combined with a pairwise similarity layer.
result Achieves 93% F1-score on Modern Standard Arabic benchmark and 82% on dialectical benchmark.

Expands robust profit opportunities to include distributional uncertainty.

problem Distributional uncertainty in financial markets.
method Formulates infinite dimensional primal problems, simplifies to finite dimensional dual problems using Wasserstein distance.
result Distributional uncertainty can enhance robustness of profit opportunities.

Traditionally it had been a problem that researchers did not have access to enough spatial data to answer pressing research questions or build compelling visualizations. Today, however, the problem is often that we have too much data. Spatially redundant or approximately redundant points may refer to a single feature (…

2018-03-21abs ↗pdf ↗

The paper explores arbitrage in financial markets under uncertainty using Wasserstein distance.

problem Investigating arbitrage in financial markets with distributional uncertainty.
method Using Wasserstein distance, the paper considers weak and strong forms of arbitrage conditions and introduces a relaxation called statistical arbitrage.
result The paper derives dual formulations of robust arbitrage conditions and conducts computational experiments to answer questions about ambiguity and statistical arbitrage.

FinanceHarness automates financial deep research with specialized tools and benchmarks.

problem Lack of specialized financial deep research for financial analysis.
method FinanceHarness uses a layered harness and FinanceGym for specialized research questions and validation.
result FinanceHarness improves rubric scores by 7.1% and achieves 82% pass rate with expert validation.

This paper is a survey on the {\em Zimmer program}. In it's broadest form, this program seeks an understanding of actions of large groups on compact manifolds. The goals of this survey are (1)(1) to put in context the original questions and conjectures of Zimmer and Gromov that motivated the program, (2)(2) to indicate t…

2008-09-28abs ↗pdf ↗

Digital personas improve survey results for stable attributes but fail for subjective responses.

problem When can digital personas reliably approximate human survey findings?
method Using LISS panel, constructed personas from background variables and survey histories, tested against held-out post-cutoff answers.
result Digital personas improve alignment with human response distributions for stable attributes but fail for subjective responses.

Predicts student performance in interactive online question pools using GNNs.

problem Predicting student performance in interactive online question pools with evolving knowledge.
method Proposes R^2GCN, a GNN model for heterogeneous networks to predict student performance.
result Achieves higher accuracy in student performance prediction than traditional methods.

Many Machine Reading and Natural Language Understanding tasks require reading supporting text in order to answer questions. For example, in Question Answering, the supporting text can be newswire or Wikipedia articles; in Natural Language Inference, premises can be seen as the supporting text and hypotheses as question…

2018-06-20abs ↗pdf ↗

Probabilistic models can handle causal inference without special tools.

problem Confusion over necessary tools for causal inference.
method Demonstrated through concrete examples that causal questions can be answered using standard probabilistic models.
result Causal questions can be addressed using standard probabilistic modelling and inference.