Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,291 papers · 148 categories

Trend · papers per month

14274154 · Jun 202019922001200920182026
48 results for feature-based partitions

GRANITE unifies feature-based explanation methods to reduce disagreement.

problem Disagreement among feature-based explanation methods.
method GRANITE partitions feature space into regions minimizing interaction and distribution influences.
result Unified and consistent feature explanations.

The paper sets communication limits for distributed optimization with feature-based data partitions.

problem Understanding communication limits in distributed convex optimization with feature-based data partitions.
method Developed tight lower bounds on communication rounds for non-incremental and incremental algorithms.
result Established communication limits for a broad class of algorithms under feature-based data partitioning.

Enhances privacy-preserving logistic regression for diabetes prediction.

problem Maintaining privacy while improving prediction accuracy in machine learning.
method Proposes stacking to enhance privacy-preserving logistic regression, either sample-based or feature-based.
result Feature-based partitioning requires fewer samples than sample-based, potentially offering better performance.

This work proposes a robust ensemble method for decision trees that resists adversarial attacks.

problem Adversarial attacks on machine learning models, especially decision trees.
method Feature partitioning to train robust ensembles and approximate certification methods.
result The proposed ensemble method can resist evasion attacks by a majority of its models.

The paper proposes criteria and methods for evaluating and aggregating feature-based model explanations.

problem Lack of quantitative evaluation criteria for feature-based model explanations.
method Developed quantitative evaluation criteria (low sensitivity, high faithfulness, low complexity), devised a framework for aggregation, and derived a new aggregate Shapley value explanation function.
result A new aggregate Shapley value explanation function that minimizes sensitivity.

Unified framework for feature-based explanations using ANOVA and game theory.

problem Differences between feature-based explanations methods limit their applicability.
method Introduces a unified framework combining fANOVA and cooperative game theory.
result Uncovered similarities and differences between various explanation techniques.

Survey and new methods for feature-based aggregation in reinforcement learning.

problem Solving finite-state discounted Markov decision problems with feature-based aggregation.
method Feature construction and aggregation, combined with deep neural networks for policy improvement.
result Approximate policy iteration can be more effective when using nonlinear function of features.

Neural clustering learns time series affinity from statistical features.

problem Challenging time series clustering with unknown cluster shapes and structures.
method Amortized neural inference using statistical features.
result Competitive clustering accuracy without manual specification of cluster shapes.

Two algorithms achieve optimal logarithmic regret in feature-based dynamic pricing.

problem Optimal pricing for products with features based on online learning.
method Developed EMLP and ONSP algorithms for stochastic and adversarial settings, proving O(dlogT)O(d\log{T}) regret bounds.
result Achieved optimal O(dlogT)O(d\log{T}) regret, improving over existing bounds.

Study agnostic feature-based dynamic pricing models with linear policies and noisy valuations.

problem Tackles dynamic pricing with unknown noise and no assumptions on data.
method Studies two agnostic models: linear policy and linear noisy valuation, presenting algorithms and regret bounds.
result Demonstrates no-regret learning is possible under weak assumptions, but noisy feedback is not significantly more useful than bandit feedback.

New model combines shape and feature-based measures for better time series classification.

problem Limited approaches in time series classification lead to poor results for some classes.
method Proposes a new model that automatically decides between shape and feature-based measures.
result Improves classification accuracy statistically significantly on real-world datasets.

A new method uses Gaussian Processes for feature-based nonrigid image registration.

problem Estimating dense displacement fields for nonrigid image registration.
method Using Gaussian Processes to estimate both dense displacement field and uncertainty map.
result GP-based interpolation performs similarly to state-of-the-art B-spline interpolation.

Unified metric c-Eval evaluates feature-based explanations by perturbation.

problem Lack of consensus on evaluating feature-based local explanations.
method Introduces c-Eval metric and framework to quantify explanation quality.
result c-Eval captures the importance of input features and is applicable in adversarial-robust models.

The paper presents a method for analyzing shape graphs using specific features.

problem Analyzing geometric and topological variations in shape graphs.
method Curated set of topological, geometric, and directional features for shape graph analysis.
result The feature representation is effective for tasks like group comparison and classification.

Rating Prediction is a basic problem in Recommender System, and one of the most widely used method is Factorization Machines(FM). However, traditional matrix factorization methods fail to utilize the benefit of implicit feedback, which has been proved to be important in Rating Prediction problem. In this work, we consi…

2014-10-29abs ↗pdf ↗

TREP learns pedestrian trajectories efficiently without needing full datasets.

problem Learning fixed-length vector representations of variable-length trajectories.
method Actor-critic sequence-to-sequence autoencoder with spatial-aware objective function.
result TREP efficiently learns trajectory representations without needing full datasets.

Model explanations can leak sensitive training data information, posing privacy risks.

problem Privacy risks of model explanations that expose training data information.
method Membership inference attacks on feature-based model explanations.
result Backpropagation-based explanations reveal statistical information about decision boundaries, leaking training data membership.

A new weighted dissimilarity measure reduces positioning errors in feature-based systems.

problem Reducing errors in feature-based positioning systems, especially in areas with high variability.
method Iterative scheme using location-dependent standard deviations as weights.
result Maximum radial positioning error reduced by 40% using the weighted dissimilarity measure.

Paper introduces methods to integrate external knowledge into RNNs using attention mechanisms.

problem Incorporating external knowledge into RNNs for improved performance.
method Proposes three methods: attentional concatenation, feature-based gating, and affine transformation.
result Attentional feature-based gating consistently improves performance across tasks.

A new method for dynamic feature selection outperforms existing approaches.

problem Sequentially selecting features based on current information in machine learning.
method Greedy selection of features based on conditional mutual information, combined with a learning approach for optimization.
result The method outperforms existing feature selection methods in experiments.

Improves machine learning models by incorporating physical laws into feature maps.

problem Lack of model interpretability in classical machine learning approaches.
method Physics-informed feature maps constructed from physical laws and dimensional analysis.
result Enhanced model interpretability and potential discovery of new physical equations.

Study shows increased precipitation variability in Paris area over years.

problem Evaluating the evolution of precipitation variability over time.
method Shape-based Dynamic Time Warping (IMS-DTW) for clustering rainfall time series.
result Precipitation variability increased in Paris area over years.

Energy-efficient detection of natural errors in deep networks.

problem Deep networks lack error detection capability without additional energy costs.
method Append RACs at hidden layers to detect natural errors with early classification termination.
result Early classification termination reduces energy consumption.

Deep CNN architectures improve neonatal seizure detection accuracy.

problem Improving EEG-based neonatal seizure detection accuracy.
method Design and test of deep convolutional networks of varying depths compared to a shallow SVM-based detector.
result A deep 11-layer CNN architecture significantly outperforms shallow architectures, improving AUC90 from 82.6% to 86.8%.

Catch22 reduces time series feature space to 22 canonical characteristics for efficient analysis.

problem Efficiently capturing and comparing time series properties for diverse applications.
method Inference of minimal sets of time-series features from a comprehensive library.
result Catch22 (22 canonical characteristics) reduces computation time and complexity.

Paper introduces privacy-preserving inventory policy learning for feature-based newsvendor with unknown demand.

problem Privacy-preserving inventory policy learning for feature-based newsvendor with unknown demand distribution and nonsmooth loss function.
method Developed a clipped noisy gradient descent algorithm based on convolution smoothing for optimal inventory estimation within f-differential privacy framework.
result Achieved privacy-preserving optimal inventory policy with provable privacy guarantees and desirable statistical precision.

Paper compares graph and set partition measures for graph clustering.

problem Comparing graph clustering methods using different similarity measures.
method Introduces graph-aware partition similarity measures and compares them with set partition measures.
result Graph-aware measures provide complementary information to set partition measures.

flacco simplifies feature-based landscape analysis for optimization problems.

problem Choosing the best optimizer from a portfolio of algorithms.
method Developed an R-package for feature-based landscape analysis.
result Makes landscape analysis accessible and comprehensible.

DFM binarizes feature embeddings for fast, accurate recommendation.

problem Expensive storage and computational cost due to large feature dimensions.
method DFM binarizes real-valued model parameters into binary codes for efficient storage and computation.
result DFM outperforms state-of-the-art binarized recommendation models and shows competitive performance compared to its real-valued version.

Rectangular Bounding Process (RBP) improves partitioning efficiency in multi-dimensional spaces.

problem Creating many unnecessary divisions in sparse regions when describing dense regions.
method Introduces Rectangular Bounding Process (RBP) to efficiently partition multi-dimensional spaces using a bounding strategy.
result The RBP is self-consistent and can be extended to infinite space, offering rich yet parsimonious expressiveness.

The study examines the balancedness of random partition models and finds the rich-get-richer characteristic is a result of model assumptions.

problem The balancedness of random partition models is largely neglected in the literature.
method Formulated a framework to define and study the balancedness of exchangeable random partition models, analyzed using product-form exchangeability and projectivity assumptions.
result The 'rich-get-richer' characteristic is an inevitable consequence of the model assumptions.

Automatically extracts features from time series data for improved forecasting.

problem Manual feature selection for time series forecasting is inefficient and prone to errors.
method Extracts features from time series using recurrence plots and computer vision algorithms.
result Automatically extracted features lead to highly comparable and sometimes superior forecasting performance.

The paper develops mixed-integer formulations for neural networks using partitioning.

problem Optimizing trained ReLU neural networks with balanced model size and tightness.
method Partitioning node inputs into groups, forming the convex hull via disjunctive programming.
result The proposed formulations outperform existing ones, especially with fewer partitions.

New partition designs reduce star discrepancy in high-dimensional sampling.

problem Improving the expected star discrepancy in high-dimensional sampling.
method Developed non-equal volume partitions to achieve lower expected star discrepancy.
result Explicit upper bounds for expected star discrepancy under non-equal volume partitions.

Study shows safely discarding features based on aggregate SHAP values is sound.

problem Discarding features based on aggregate SHAP values without proper justification.
method Investigated the soundness of discarding features based on aggregate SHAP values, proposing to aggregate SHAP values over the extended support.
result A small aggregate SHAP value implies safely discarding the corresponding feature.