TopoFisher learns topological summaries by maximizing Fisher information, improving parameter efficiency and inference quality.
problem Simulation-based inference misses key information in low-order statistics, especially for non-Gaussian fields.
method TopoFisher uses a differentiable persistent-homology pipeline that learns topological summaries by maximizing local Gaussian Fisher information.
result TopoFisher recovers much of the available information and outperforms fixed topological vectorizations in weak gravitational lensing.
Researchers use topological summaries in regression and classification tasks.
problem Lack of positive semi-definite property in topological summaries.
method Defined a topological exponential kernel and showed its effectiveness in regression and classification.
result Topological summaries can be successfully used in supervised learning tasks.
Parallel Mapper algorithm for efficient topological data analysis.
problem Efficient parallel processing of Mapper for topological data analysis.
method Provable correct parallel algorithm for Mapper execution on multiple processors.
result Demonstrates the efficiency of parallel Mapper compared to sequential implementations.
TCDA separates observation space and causal assumptions for stable summaries.
problem Undefined or inadequate outcomes in modern data.
method Separates observation space, causal-model class, and topological representation.
result Identification and stability of causal effects through topology.
Topological methods improve neuron analysis and tracer injection summary.
problem Traditional methods fail to capture the tree-like structure of neurons.
method Discrete Morse (DM) Theory for neuron skeletonization and consensus tree summarization.
result Significant performance improvements over non-topological methods.
A new method compares persistent cycles in topological data.
problem Comparing persistent homology representations of two spaces.
method Direct comparison of individual persistent cycles based on persistence intervals and spatial placement.
result Demonstrated the effectiveness of the method in topological inference.
New methods compare neural network models using geometric and topological summaries.
problem Comparing deep representations of complex networks in models and brains.
method Develops inference methods based on topological data analysis (TDA) and graph-based techniques.
result New statistical methods enable better model comparison and inference.
Two approaches use TDA and graph theory for tennis match prediction.
problem Predicting tennis match outcomes using network features.
method Lower-star filtration on player competitive networks, Random Forest model, modified Katz similarity index.
result TDA features alone can achieve above-chance prediction in tennis match outcomes.
A new method uses vectorized summaries of persistence diagrams for efficient hypothesis testing.
problem Efficient hypothesis testing for large and complex persistence diagrams.
method Vectorized summaries of Betti functions and a new shuffling technique.
result The vectorized Betti function leads to competitive results compared to baseline methods.
New tool for summarizing time-varying data shapes.
problem Understanding dynamic data shapes.
method Introducing crocker stacks for time-varying metric spaces.
result Demonstrated utility in parameter identification task.
Estimates and quantizes expected persistence diagrams for efficient analysis.
problem Statistical summary of the topology of structured data.
method Expected Persistence Diagram (EPD) and its quantization.
result Optimal estimation of EPD with near-optimal quantization.
Stable topological summary captures evolving dependency structure in dynamic Bayesian networks.
problem Missing larger-scale patterns in evolving dependency structures in dynamic Bayesian networks.
method Topological approach using Dynamic Bayesian Graphs and persistent homology.
result Stable topological summary (barcodes) captures evolving dependency structure in DBNs.
Study examines persistence diagrams in machine learning, proposing permutation tests.
problem Understanding the power and limitations of persistence diagrams in machine learning.
method Carried out experiments on graph and shape data, proposed permutation tests for persistence diagrams.
result Persistence pairing shows significant improvement in various tasks, but the most critical values are most discriminative.
Topological data analysis offers a rich source of valuable information to study vision problems. Yet, so far we lack a theoretically sound connection to popular kernel-based learning techniques, such as kernel SVMs or kernel PCA. In this work, we establish such a connection by designing a multi-scale kernel for persist…
We give a short introduction to the theory of twisted Alexander polynomials of a 3--manifold associated to a representation of its fundamental group. We summarize their formal properties and we explain their relationship to twisted Reidemeister torsion. We then give a survey of the many applications of twisted invarian…
New mechanism for pure differential privacy on functional summaries using Laplace-like process.
problem Challenges in achieving differential privacy for complex, structured functional summaries.
method Independent Component Laplace Process (ICLP) mechanism for infinite-dimensional Hilbert space.
result Effective enhancement of utility of private summaries through oversmoothing.
Approximate Bayesian Computation (ABC) methods are used to approximate posterior distributions in models with unknown or computationally intractable likelihoods. Both the accuracy and computational efficiency of ABC depend on the choice of summary statistic, but outside of special cases where the optimal summary statis…
LIDS assesses LLM summaries with interpretable key words.
problem Challenges in evaluating the quality of LLM summaries.
method BERT-SVD-based direction metric and SOFARI for key word extraction.
result LIDS provides interpretable key words for layered themes.
We use barcodes to analyze neural networks' loss surfaces, revealing important properties.
problem Understanding the topology of neural networks' loss surfaces.
method Topological data analysis using Morse complexes and barcodes.
result Barcodes of local minima are located in a small part of the loss function's range and decrease with network depth and width.
A new method uses bandits to select summary statistics for Bayesian inference.
problem Dynamic selection of summary statistics for likelihood-free inference.
method Treats summary statistic selection as a multi-armed bandit problem.
result Improves efficiency and scalability of approximate Bayesian computation.
Unified approach for selecting summary statistics in ABC.
problem Efficient inference from large datasets in likelihood-free methods.
method Characterizing and unifying three classes of summary statistics, minimizing expected posterior entropy.
result EPE-minimizing summaries lead to competitive posterior inference.
Enhances weak lensing inference with neural summaries.
problem Extracting additional information from weak lensing convergence maps.
method Hybrid approach combining physics-based and neural summaries.
result Neural summaries extract up to 8 times more information than angular power spectra.
New method detects uncertainty in neural networks for out-of-distribution detection.
problem Detecting out-of-distribution inputs to ensure model reliability.
method Predictive topological uncertainty (pTU) based on persistent homology.
result pTU provides a statistical framework for OOD detection.
A very interesting problem in the classical theory of minimal surfaces consists of the classification of such surfaces under some geometrical and topological constraints. In this short paper, we give a brief summary of the known classification results for properly embedded minimal surfaces with genus zero in $\mathbb{R…
Approaches for approximating persistent homology for large datasets.
problem Inability to compute persistent homology for large datasets.
method Multiple subsampling framework for statistical approximation of persistent homology.
result Derivation of finite sample convergence rates for empirical means of persistent homology.
Novel TRI-GNN framework improves graph classification robustness.
problem Graph neural networks suffer from over-smoothing and vulnerability to graph perturbations.
method Integrates higher-order graph information via persistent homology and local graph structure learning.
result TRI-GNN outperforms state-of-the-art baselines on node classification tasks.
TDA-based portfolios show better risk-adjusted returns than classical methods.
problem Traditional portfolio selection methods fail to capture complex asset dynamics.
method Topological Data Analysis (TDA) using persistence landscapes to quantify portfolio risk.
result TDA-based portfolios outperform classical models in excess mean return and financial ratios.
Often, high dimensional data lie close to a low-dimensional submanifold and it is of interest to understand the geometry of these submanifolds. The homology groups of a manifold are important topological invariants that provide an algebraic summary of the manifold. These groups contain rich topological information, for…
Automatically learns summary features from time series data for likelihood-free inference.
problem Necessity of hand-tailored summary features for time series data in likelihood-free inference.
method Data-driven approach to automatically learn summary features.
result Learning summary features from data can outperform hand-crafted values in likelihood-free inference.
Improves inference from sparse data with hybrid summary statistics.
problem Robust simulation-based inference from limited data.
method Augment traditional summary statistics with neural network outputs to maximize mutual information.
result Improves information extraction and makes inference robust in low-data settings.
Text clustering method replaces centroids with summaries for interpretability and scalability.
problem Efficiently clustering text data while maintaining interpretability and scalability.
method k-NLPmeans and k-LLMmeans, which periodically replace numeric centroids with textual summaries.
result Consistently outperforms classical baselines and recent LLM-based clustering methods.
Real world systems typically feature a variety of different dependency types and topologies that complicate model selection for probabilistic graphical models. We introduce the ensemble-of-forests model, a generalization of the ensemble-of-trees model. Our model enables structure learning of Markov random fields (MRF) …
Bayesian neural networks improve with summary information and Dirichlet process.
problem Lack of prior knowledge in BNNs for complex architectures.
method Incorporates external summary information about predicted probabilities using a Dirichlet process.
result Improves model accuracy, uncertainty calibration, and robustness.
A new method for stable vector representation of persistence diagrams.
problem Finding a stable vector representation of persistence diagrams for ML tasks.
method Persistence B-spline Grid (PBSG) based on data fitting.
result The PBSG method is stable with respect to the 1-Wasserstein distance metric.
Plug-in robust NPE method adapts summaries independently of pretrained NPE.
problem Misspecification of neural posterior estimators under test data distribution.
method Minimum-distance summaries using maximum mean discrepancy (MMD).
result Substantial robustness gains with minimal additional overhead.
Z-GCNETs uses topological data to improve time series forecasting.
problem Improving time series forecasting accuracy.
method Integrates topological data into graph convolutional networks (GCNs) using zigzag persistence.
result Z-GCNETs outperforms 13 state-of-the-art methods in traffic forecasting and Ethereum price prediction.
Few summaries enable automatic summarization of product reviews.
problem Lack of large labeled datasets for training supervised models in opinion summarization.
method Conditional Transformer model trained to generate summaries given other reviews, fine-tuned to predict summary properties.
result Few summaries (5-10) are sufficient to generate fluent, informative, and sentiment-preserving summaries.
PENs learn summary statistics for ABC using invariant neural architectures.
problem Learning summary statistics for approximate Bayesian computation (ABC).
method Partially exchangeable networks (PENs) that are invariant to block-switch transformations.
result PENs provide more reliable posterior samples with less training data.
Improved likelihood-free inference by localizing and refining low-dimensional approximations.
problem Poor performance of common likelihood-free methods in high-dimensional models.
method Localisation followed by refinement of low-dimensional summaries.
result Improved accuracy in marginal posteriors through localized and refined approximations.
New summary measures reveal geometric structure in weighted measures on manifolds.
problem Lack of geometric information in standard weight-only summaries.
method Heat-kernel entropy profiles, tracking nonuniformity across scales.
result Geometric effective sample size discounts nearby or duplicate particles.
A new method improves likelihood-free Bayesian inference by transforming summary statistics and using efficient Variational Bayes.
problem Incorrectly assuming normally distributed summary statistics in likelihood-free Bayesian inference.
method Wasserstein Gaussianization transformation combined with robust BSL and efficient Variational Bayes.
result Highly efficient and reliable approximate Bayesian inference for likelihood-free problems.
New algorithm creates small summaries for various clustering problems.
problem Scaling clustering algorithms to massive data sets.
method One-shot coreset construction for k-clustering problems.
result Constructs small data summaries for a wide range of clustering problems simultaneously.
Paper uses learned summary statistics for Bayesian inference with difficult likelihood functions.
problem Difficult to obtain exact likelihood function for observation data and simulation model.
method Simulation-based inference with learned summary statistics, using Cressie-Read discrepancy criterion.
result Effective inference performed over selected sample sets of observation data.
Flexible multi-task learning framework using summary statistics.
problem Data-sharing constraints in healthcare settings.
method Proposes a flexible multi-task learning framework utilizing summary statistics and adaptive parameter selection.
result Systematic non-asymptotic analysis and simulations demonstrate the method's performance.
Develops methods to learn centre groupings from summary statistics in multi-centre studies.
problem Violation of homogeneity of parameters across centres in multi-centre studies.
method Clusters-of-Centres (CoC) algorithm that merges centres based on multivariate Cochran-type tests.
result Golden-partition recovery as the number of rounds grows with sample size.
Unsupervised summarization generates novel reviews reflecting consensus opinions.
problem Creating summaries that reflect subjective information in multiple documents.
method Generative model with hierarchical variational autoencoder, pointer-generator mechanism.
result Model produces fluent and coherent summaries reflecting common opinions.
The paper proposes using Autoencoders to learn summary statistics for Bayesian inference.
problem Approximating posterior distributions for models with intractable likelihood functions.
method Using Autoencoders to extract summary statistics that retain parameter information and cancel noise.
result The approach effectively learns summary statistics that improve posterior approximation.
Paper introduces methods to automatically generate SOAP notes from patient-physician conversations.
problem Burden of creating digital SOAP notes by physicians.
method Cluster2Sent algorithm for summarizing patient-physician conversations.
result Cluster2Sent algorithm outperforms existing methods by 8 ROUGE-1 points.