Paper improves geographic location embeddings using Flickr tags and structured data.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Deep learning model extracts location references from tweets during emergencies.
SE-KGE embeds spatial data into KGs for better spatial reasoning.
We are interested in learning customers' video preferences from their historic viewing patterns and geographical location. We consider a Bayesian latent factor modeling approach for this task. In order to tune the complexity of the model to best represent the data, we make use of Bayesian nonparameteric techniques. We …
Cost overruns in transport infrastructure projects know no geographical limits, overruns are a global phenomenon. Nevertheless, the size of cost overruns varies with location. In the Netherlands, cost overruns appear to be smaller compared to the rest of the world. This paper tests whether Dutch projects perform signif…
Paper presents a method for geographic ratemaking using spatial embeddings.
Study improves motor insurance claim prediction using geographic data.
Study evaluates methods for improving model robustness to various real-world distribution shifts.
Geography effect is investigated for the Chinese stock market including the Shanghai and Shenzhen stock markets, based on the daily data of individual stocks. The Shanghai city and the Guangdong province can be identified in the stock geographical sector. By investigating a geographical correlation on a geographical pa…
Study uses trajectory embedding to measure place function similarity at fine spatial granularity.
Study reveals centralization in Bitcoin transactions involving retail users.
Automated valuation model uses diverse data sources for real estate appraisal.
GWRBoost improves GWR for better spatial relationship quantification.
Interpretable ML models predict recidivism as well as non-interpretable methods and are more fair.
This paper develops synthetic mobility datasets to protect privacy while maintaining realism.
This research investigated the potential for improving Peer-to-Peer (P2P) credit scoring by using "private information" about communications and travels of borrowers. We found that P2P borrowers' ego networks exhibit scale-free behavior driven by underlying preferential attachment mechanisms that connect borrowers in a…
Which song will Smith listen to next? Which restaurant will Alice go to tomorrow? Which product will John click next? These applications have in common the prediction of user trajectories that are in a constant state of flux over a hidden network (e.g. website links, geographic location). What users are doing now may b…
Space2Vec learns multi-scale spatial representations from grid cell insights.
Study predicts climate data at distant locations using machine learning.
Spatially-aware model improves earthquake hazard assessment accuracy.
Proposes D2D-LSTM for predicting mobile social network content diffusion paths.
We present an analysis of the credit market of Japan. The analysis is performed by investigating the bipartite network of banks and firms which is obtained by setting a link between a bank and a firm when a credit relationship is present in a given time window. In our investigation we focus on a community detection alg…
Dual random fields improve mineral potential predictions.
A model predicts solar irradiance without local data using satellite and weather forecasts.
We investigate the community structure of the global ownership network of transnational corporations. We find a pronounced organization in communities that cannot be explained by randomness. Despite the global character of this network, communities reflect first of all the geographical location of firms, while the indu…
Paper tackles pandemic resource allocation challenges.
GeoLifeCLEF 2020 dataset pairs species observations with environmental data.
Modern physics has demonstrated that matter behaves very differently as it approaches the speed of light. This paper explores the implications of modern physics to the operation and regulation of financial markets. Information cannot move faster than the speed of light. The geographic separation of market centers means…
How are economic activities linked to geographic locations? To answer this question, we use a data-driven approach that builds on the information about location, ownership and economic activities of the world's 3,000 largest firms and their almost one million subsidiaries. From this information we generate a bipartite …
This paper pretends to analyze the importance which the natural advantages and local resources are in the manufacturing industry location, in relation with the "spillovers" effects and industrial policies. To this, we estimate the Rybczynski equation matrix for the various manufacturing industries in Portugal, at regio…
Cloud-native simulator studies global trading latencies.
Graph-partitioning-based DCRNN improves traffic forecasting for large highways.
The problem of predicting the location of users on large social networks like Twitter has emerged from real-life applications such as social unrest detection and online marketing. Twitter user geolocation is a difficult and active research topic with a vast literature. Most of the proposed methods follow either a conte…
In this paper, we consider the problem of predicting demographics of geographic units given geotagged Tweets that are composed within these units. Traditional survey methods that offer demographics estimates are usually limited in terms of geographic resolution, geographic boundaries, and time intervals. Thus, it would…
We propose a latent self-exciting point process model that describes geographically distributed interactions between pairs of entities. In contrast to most existing approaches that assume fully observable interactions, here we consider a scenario where certain interaction events lack information about participants. Ins…
Mathematical analysis shows Delisle-Euler map methods are optimal.
Neural networks are capable of learning rich, nonlinear feature representations shown to be beneficial in many predictive tasks. In this work, we use such models to explore different geographical feature representations in the context of predicting colorectal cancer survival curves for patients in the state of Iowa, sp…
Active authentication is the problem of continuously verifying the identity of a person based on behavioral aspects of their interaction with a computing device. In this study, we collect and analyze behavioral biometrics data from 200subjects, each using their personal Android mobile device for a period of at least 30…
Weather derivatives help farmers hedge against crop yield risks.
In this paper we propose a Bayesian nonparametric model for clustering partial ranking data. We start by developing a Bayesian nonparametric extension of the popular Plackett-Luce choice model that can handle an infinite number of choice items. Our framework is based on the theory of random atomic measures, with the pr…
We investigate the tendency for financial instruments to form clusters when there are multiple factors influencing the correlation structure. Specifically, we consider a stock portfolio which contains companies from different industrial sectors, located in several different countries. Both sector membership and geograp…
A model for POI recommendation using relation embedding.
PREMA recovers detailed data from aggregated views.
Image compression techniques reveal network structure for shipping box optimization.
STICC clusters geographic objects considering both spatial contiguity and attributes.
Tropical cyclone wind-intensity prediction is a challenging task considering drastic changes climate patterns over the last few decades. In order to develop robust prediction models, one needs to consider different characteristics of cyclones in terms of spatial and temporal characteristics. Transfer learning incorpora…
Post-estimation smoothing improves prediction accuracy with structural indices.
Predict and explain service failures in supply-chain networks using data models.