Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,932 papers · 148 categories

Trend · papers per month

1122 · May 201919922001200920172026
19 results for geolocation

New algorithm predicts geolocation of fungi samples with high accuracy.

problem Identifying the origin of biological material at crime scenes.
method Ensemble of deep neural network classifiers trained on Voronoi partitions.
result More than half of geolocation errors under 100 kilometers for continental analysis and nearly 90% accuracy for global analysis.

The problem of predicting the location of users on large social networks like Twitter has emerged from real-life applications such as social unrest detection and online marketing. Twitter user geolocation is a difficult and active research topic with a vast literature. Most of the proposed methods follow either a conte…

2017-12-21abs ↗pdf ↗

New method generates geolocated synthetic populations from real data.

problem Generating synthetic populations with explicit geographic coordinates.
method Mapping coordinates into a latent space using Normalizing Flows (NF), then combining with other features in a Variational Autoencoder (VAE).
result NF+VAE architecture outperforms existing methods in generating geolocated synthetic populations.

Data on human spatial distribution and movement is essential for understanding and analyzing social systems. However existing sources for this data are lacking in various ways; difficult to access, biased, have poor geographical or temporal resolution, or are significantly delayed. In this paper, we describe how geoloc…

2016-06-20abs ↗pdf ↗

In this paper we tackle the issue of clustering trajectories of geolocalized observations. Using clustering technics based on the choice of a distance between the observations, we first provide a comprehensive review of the different distances used in the literature to compare trajectories. Then based on the limitation…

2015-08-20abs ↗pdf ↗

Users form information trails as they browse the web, checkin with a geolocation, rate items, or consume media. A common problem is to predict what a user might do next for the purposes of guidance, recommendation, or prefetching. First-order and higher-order Markov chains have been widely used methods to study such se…

2017-04-20abs ↗pdf ↗

Spatial variable selection is crucial for reliable spatial predictions in machine learning.

problem Spatial autocorrelation leads to overfitting and poor spatial predictions.
method Used Random Forests with non-spatial and spatial cross-validation strategies.
result Spatial variable selection is essential for reliable spatial predictions.

Chagas disease is a neglected disease, and information about its geographical spread is very scarse. We analyze here mobility and calling patterns in order to identify potential risk zones for the disease, by using public health information and mobile phone records. Geolocalized call records are rich in social and mobi…

2018-08-09abs ↗pdf ↗

The paper classifies U.S. crop types using hyperspectral satellite imagery.

problem Classifying crop types from hyperspectral satellite imagery.
method Gaussian Bayesian models and neural networks applied to NASA data.
result Bayesian methods outperform standard LDA and QDA.

This work improves local differential privacy by considering context to make it more effective.

problem Local differential privacy often sacrifices utility, especially for sensitive data.
method Introduces context-aware local differential privacy, optimizing privacy and utility.
result Contextual information can reduce the number of samples needed for privacy compared to classical LDP.

This study maps cycling risks and discomfort in Zurich, offering personalized route recommendations.

problem High cycling accidents and discomfort in Smart Cities.
method Geolocated bike accidents data, kernel density contours, weather, time, accident type and severity analysis.
result Empirical continuous spatial risk estimations and personalized route recommendations.

Paper tackles fairness in algorithms by predicting protected class from auxiliary data.

problem Protected class membership is often unobserved in data, leading to unfair algorithmic decisions.
method Use auxiliary datasets to predict protected class from proxy variables and provide characterizations of possible disparities.
result Common disparity measures are generally unidentifiable with auxiliary data, highlighting the need for robust assessments.

Study uses satellite and lidar data to map forest height and biomass in France.

problem Mapping forest resources and carbon in large areas.
method Machine learning approach using Sentinel-1, Sentinel-2, ALOS-2, and GEDI Lidar data.
result High-resolution maps of forest height and biomass produced with good accuracy.