Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,181 papers · 148 categories

Trend · papers per month

4386129172 · Jun 202019922001200920182026
48 results for AI embeddings

AI pipeline simplifies AI deployment on embedded devices.

problem Training and deployment of custom AI solutions on embedded devices requires expertise and integration barriers.
method Modular AI pipeline integrating data, algorithms, and deployment tools.
result LPDNN consistently outperforms other deployment frameworks on embedded platforms.

Fixing a closed hyperbolic surface S, we define a moduli space AI(S) of unmarked hyperbolic 3-manifolds homotopy equivalent to S. This 3-dimensional analogue of the moduli space M(S) of unmarked hyperbolic surfaces homeomorphic to S has bizarre local topology, possessing many points that are not closed. There is, howev…

2009-06-30abs ↗pdf ↗

Deep learning model for AIS data improves vessel monitoring.

problem Improving maritime safety and efficiency with AIS data.
method Combining RNNs and latent variable modeling for multi-task learning.
result Demonstrated relevance on real AIS datasets for trajectory reconstruction, anomaly detection, and vessel type identification.

GAICF proposes a framework for governing generative AI in banking.

problem Generative AI's impact on financial decision-making and governance.
method SR 26-2-compatible governance framework for generative AI applications.
result GAICF aligns generative AI practices with SR 26-2 supervisory expectations.

GAICF proposes a framework for managing generative AI risks in banking.

problem Generative AI's impact on financial decision-making and governance.
method SR 26-2-compatible governance framework for generative AI.
result GAICF aligns generative AI practices with SR 26-2 supervisory expectations.

Benchmark evaluates AI-generated financial QA hallucinations, highlighting system vulnerabilities.

problem Ensuring factual accuracy of AI-generated financial QA outputs.
method Developed a benchmark dataset and evaluated six detection methods under clean and noisy conditions.
result LLM-based judges and embedding methods perform best, but degrade under noisy conditions.

TED framework teaches AI to explain decisions, improving accuracy.

problem Providing understandable explanations for AI predictions in high-stakes applications.
method Augmenting training data with explanations from domain users, using embeddings and multi-task learning.
result AI models can be taught to provide meaningful explanations, sometimes improving accuracy.

This work advances collaborative decision making by combining human and AI strengths in uncertainty quantification.

problem Current AI lacks robust decision-making capabilities under uncertainty, especially in high-stakes contexts.
method Introduces Human AI Collaborative Uncertainty Quantification (HACUQ) framework, formalizing AI-human collaboration and developing calibration algorithms.
result Optimal collaborative prediction sets follow a two-threshold structure, and online adaptation algorithms can adapt to evolving human behavior.

Paper introduces RiskEmbed, a finetuned model for financial risk management.

problem Improving retrieval accuracy in financial question-answering systems.
method Curated dataset and finetuned BERT model for financial domain.
result RiskEmbed significantly outperforms general-purpose and financial embedding models.

Paper forecasts commodity price spikes using AI and economic news.

problem Accurate forecasting of commodity price spikes for economic stability.
method Hybrid framework combining historical data and semantic signals from economic news.
result Model achieves high AUC and accuracy in detecting price shocks.

Develops a topology-based test for AI model alignment and interpretability.

problem Testing and interpreting opaque AI models is difficult due to their complexity and lack of interpretability.
method Introduces a topology-based multi-modal alignment test to make AI models more interpretable.
result Demonstrates the effectiveness of the topology-based test in making AI model deployment and comparison more intuitive.

Study finds price-based clustering outperforms AI and human methods in stock market analysis.

problem Investigates if AI can improve stock clustering compared to traditional methods.
method Compares price-based, human-informed, and AI-driven clustering methods using synthetic factor models.
result Price-based clustering reduces RMSE by 15.9% relative to GICS and 14.7% relative to LLM embeddings.

New framework learns complex AI attitudes from heterogeneous data.

problem Heterogeneous ordinal structure in AI attitudes, poorly captured by existing methods.
method Monotone Gaussian score embedding, BNP complexity discovery, confirmatory fixed-K estimation.
result Reduced holdout MSE by 25.8% over single-graph baseline.

Novel method diagnoses large language models' reasoning abilities.

problem Fine-grained evaluation of large language models' reasoning abilities.
method Adapting cognitive diagnosis models to LLMs, estimating mastery profiles and Q-matrix, incorporating textual information.
result Accurate parameter recovery and insights into LLMs' capabilities.

This work improves AI's ability to solve physical tasks by optimizing world models in abstracted spaces.

problem Developing AI agents capable of solving diverse physical tasks and generalizing to new environments.
method Investigates and optimizes a family of joint-embedding predictive world models (JEPA-WMs) for efficient planning in abstracted spaces.
result Proposes a model that outperforms two established baselines in both navigation and manipulation tasks.

SEMASIA provides a large dataset of latent representations for model comparison.

problem Difficulty in comparing semantic structures across different neural network models.
method Collection of latent representations from 1700 pretrained models across various benchmarks.
result Consistent semantic organization across models and datasets.

Model shows AI adoption amplifies financial market risk through prediction, herding, and cognitive dependency.

problem Systemic risk in financial markets due to AI adoption.
method Developed a unified model within an extended rational expectations framework, incorporating endogenous adoption, performative prediction, algorithmic herding, and cognitive dependency.
result Systemic risk multiplier grows superlinearly with AI penetration, implying tail-loss amplification of 18-54%.

Study copyright's impact on creative industries using AI-generated fonts.

problem Estimating supply and demand in creative industries with AI-generated content.
method Neural network embeddings, spatial regression, event-study analyses, structural model of supply and demand.
result Copyright can raise consumer welfare by encouraging product relocation.

Comma.ai's approach to Artificial Intelligence for self-driving cars is based on an agent that learns to clone driver behaviors and plans maneuvers by simulating future events in the road. This paper illustrates one of our research approaches for driving simulation. One where we learn to simulate. Here we investigate v…

2016-08-03abs ↗pdf ↗

LeJEPA provides a scalable, theory-driven approach to self-supervised learning.

problem Lack of practical guidance and theory in JEPAs.
method Identified optimal Gaussian distribution and introduced SIGReg objective.
result LeJEPA achieves state-of-the-art performance with minimal hyperparameters and heuristics.

ViCE uses superpixels to enhance self-supervised learning for better dense visual embeddings.

problem Lack of high-resolution feature maps from self-supervised models.
method Superpixels for dense representation learning, contrasting over regions.
result Improves unsupervised semantic segmentation on benchmarks like Cityscapes and COCO.

Randomized Geometric Algebra for Convex Neural Networks Optimizes Transfer Learning.

problem Training neural networks to global optimality via convex optimization.
method Randomized algorithms in Clifford's Geometric Algebra for hypercomplex vector spaces.
result Convex optimization and geometric algebra improve LLMs' robustness and reliability in transfer learning.

AI improves credit rating predictions over traditional methods.

problem Improving credit rating predictions for global corporate entities.
method Applying deep learning techniques, specifically neural networks with categorical embeddings, to a large dataset of corporate obligations.
result Deep learning models achieve adequate accuracy in predicting different credit rating classes.

AI-enabled precision medicine promises a transformational improvement in healthcare outcomes by enabling data-driven personalized diagnosis, prognosis, and treatment. However, the well-known "curse of dimensionality" and the clustered structure of biomedical data together interact to present a joint challenge in the hi…

2022-11-29abs ↗pdf ↗

Study analyzes AI's impact on firms, markets, and workers using large language model data.

problem Understanding AI's effect on firms, markets, and workers.
method Used 380 trillion tokens from 400+ large language models to analyze AI's impact.
result Firms with higher AI exposure earn higher returns, creating an AI premium.