Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

1.6%3.2%4.8%6.4% · Dec 201819922001200920172026
48 results for water quality

The study forecasts water quality from satellite data using machine learning.

problem Predicting future water quality from satellite data for coastal regions.
method Decomposed time series into components and used machine learning models (SARIMA, regression, neural network).
result Regression and neural network models are best at predicting Chl-a, SARIMA model best at FLH and SST.

LightGBM outperforms other models in predicting pH values in Georgia, USA.

problem Accurate water quality prediction for effective resource management and pollution mitigation.
method Five distinct predictive models (linear regression, Random Forest, XGBoost, LightGBM, MLP neural network) were assessed for pH value forecasting in Georgia, USA.
result LightGBM achieved the highest average precision in predicting pH values.

Study predicts coastal water quality using machine learning, identifying salinity as key factor.

problem Predicting and managing coastal water quality for public health and tourism.
method Machine learning models (Catboost, Xgboost, Random Forests, Support Vector Regression, Artificial Neural Networks) trained on environmental data.
result Catboost algorithm performed best, with R² values of 0.71 and 0.68 for E. Coli and enterococci predictions.

Machine learning models estimate nutrient concentrations from water quality surrogates.

problem Estimating high frequency nutrient concentrations from limited in-situ measurements.
method Used machine learning (Random Forests) to estimate nutrient concentrations using surrogate measures.
result Reduced RMSE by up to 60.1% compared to linear models, with additional sensors not providing significant benefits.

The study analyzes river water quality using statistical and machine learning methods.

problem Analyzing spatio-temporal dynamics of dissolved oxygen in the River Thames.
method Superstatistical methods and machine learning (e.g., Light Gradient Boosting Machine, Informer model).
result The Informer model outperforms others in long-term dissolved oxygen concentration forecasting.

Novel Orlicz regrets consistently bound environmental variable statistics.

problem Consistent evaluation of stochastic environmental variables like water quality indices.
method Proposed novel Orlicz regrets for upper and lower bounds.
result Explicit linkage between Orlicz regrets and divergence risk measures.

The paper presents anomaly detection in time series data using InfluxDB and Python.

problem Anomalous data points in time series data affect decision making in water and environmental systems.
method Data cleaning, cost-sensitive machine learning (Logistic Regression, Random Forest, SVM), feature selection, and InfluxDB integration.
result Random Forest outperformed other models in detecting anomalies.

Transformer-based diffusion models improve hydrological time series imputation and forecasting.

problem Limited observations in hydrometeorological time series.
method Transformer-based diffusion models applied to hydrological data.
result Transformer-based models efficiently sample realistic time series distributions under variable missing data.

Study predicts stream turbidity using surrogate data and meta-model.

problem Costly turbidity sensor deployment limits monitoring networks.
method Dynamic regression (ARIMA), LSTM, GAM models; surrogate covariates (rainfall, water level, temperature, solar exposure); meta-model combining strengths of individual models.
result ARIMA and GAM models with all covariates outperform single models; meta-model yields highest accuracy.

Predict water pipe failures using machine learning and survival analysis.

problem Difficulty in accessing water pipes for maintenance.
method Classical and modern classifiers for short-term prediction, survival analysis for long-term forecast, and oversampling technique for imbalanced data.
result Identifies important risk factors for water pipe failures.

Energy consumption for hot water production is a major draw in high efficiency buildings. Optimizing this has typically been approached from a thermodynamics perspective, decoupled from occupant influence. Furthermore, optimization usually presupposes existence of a detailed dynamics model for the hot water system. The…

2018-01-04abs ↗pdf ↗

New method clusters hydrological and sediment data for storm event analysis.

problem Analyzing storm events for water quality constituents like turbidity.
method Multivariate time series clustering of river discharge and sediment data.
result Clusters differ from 2-D hysteresis loop classifications.

CoCAI uses copulas for accurate multivariate time-series forecasting and anomaly detection.

problem Accurate multivariate time-series forecasting and robust anomaly detection.
method Copula-based conformal prediction for multivariate time-series analysis.
result CoCAI provides statistically valid predictive regions and robust anomaly scores.

Machine learning predicts liquid water properties from cluster data.

problem Accuracy of bulk properties from machine-learned potentials is limited by training data.
method Local, atom-centred descriptors enable prediction of bulk properties from cluster data.
result Excellent agreement with experimental and theoretical counterparts of liquid water properties.

Hybrid model predicts flow and pressure in water systems.

problem Predicting flow and pressure in water distribution systems with complex spatial-temporal correlations.
method Hybrid dual-stage spatial-temporal attention-based recurrent neural networks (hDS-RNN).
result Our model outperformed 9 baseline models in flow and pressure series prediction.

Method reduces model bias in water temperature prediction using physics-guided GNNs.

problem Model bias in traditional physics-based models across different income and education levels.
method Physics-guided GNNs with refined neighbor selection and weights.
result Preserves equitable performance across different sensitive groups in the Delaware River Basin.

Combines machine learning and convex limiting for accurate subgrid flux modeling in shallow-water equations.

problem Accurate subgrid flux modeling in shallow-water equations.
method Machine learning and flux limiting for property-preserving subgrid scale modeling.
result The proposed method produces meaningful closures even in untrained scenarios.

Low-cost water-level tracking using LTE power metrics and wavelet analysis.

problem Real-time water-level monitoring across many locations with fixed instruments.
method Extracts per-antenna RSRP, RSSI, and RSRQ, applies CWT to RSRP, and uses a neural network to track water-level changes.
result Achieves root-mean-square and mean-absolute errors of 0.8 cm and 0.5 cm, respectively, under line-of-sight conditions.

Generative model improves noise estimation in stochastic rotating shallow water models.

problem Improving noise estimation in stochastic partial differential equations for fluid dynamics.
method Replaced PCA with a generative model to avoid constraints on stochastic increments.
result Generative model produces better RMSE, CRPS score, and forecast rank histograms.

In the face of growing needs for water and energy, a fundamental understanding of the environmental impacts of human activities becomes critical for managing water and energy resources, remedying water pollution, and making regulatory policy wisely. Among activities that impact the environment, oil and gas production, …

2019-08-29abs ↗pdf ↗

Deep learning method improves myelin water fraction estimation.

problem Estimating myelin water fraction in the brain using magnetic resonance relaxometry.
method Combines input layer regularization with automated regularization hyperparameter tuning.
result Proposed method outperforms classical methods and multi-layer perceptrons on in vivo brain data.