CausalMix generates synthetic data with causal controls for mixed-type tables.
problem Synthetic data for causal inference with mixed-type and multimodal tabular data.
method CausalMix combines Gaussian latent priors with data-type-specific decoders for control over overlap, confounding, and treatment effect heterogeneity.
result CausalMix achieves state-of-the-art distributional metrics and stable causal control.
A new method clusters mixed-type data tables effectively.
problem Clustering data with mixed types (numerical and categorical).
method Two-step approach: binarize mixed data, then co-cluster.
result Shows improved clustering of mixed-type data compared to MCA.
We propose a MAP Bayesian approach to perform and evaluate a co-clustering of mixed-type data tables. The proposed model infers an optimal segmentation of all variables then performs a co-clustering by minimizing a Bayesian model selection cost function. One advantage of this approach is that it is user parameter-free.…
Extends co-clustering to mixed numerical and binary data.
problem Co-clustering of mixed data types (numerical and binary).
method Latent block models for mixed data types.
result Effectiveness of the proposed approach on simulated data.
New examples of mixed-type zero-curvature graphs found.
problem Finding new examples of zero-curvature graphs in Lorentz-Minkowski space.
method Using Konderak's representation formula to construct entire zero-curvature graphs over specific planes.
result Existence of new types of entire zero-curvature graphs in mixed-type in Lorentz-Minkowski space.
A connected regular surface in Lorentz-Minkowski 3-space is called a mixed type surface if the spacelike, timelike and lightlike point sets are all non-empty. Lightlike points on mixed type surfaces may be regarded as singular points of the induced metrics. In this paper, we introduce the L-Gauss map around non-degener…
New model clusters mixed-type data with missing values, improving air quality analysis.
problem Clustering mixed-type data with missing values and regime persistence.
method Statistical jump model incorporating regime persistence and handling missing data.
result Superior performance in inferring persistent air quality regimes compared to traditional methods.
New PDEs of mixed type emerge in fluid mechanics and geometry.
problem Analysis of nonlinear PDEs of mixed type.
method Through historical problems and recent trends.
result Many PDEs are of mixed type, requiring new analysis.
A mixed type surface is a connected regular surface in a Lorentzian 3-manifold with non-empty spacelike and timelike point sets. The induced metric of a mixed type surface is a signature-changing metric, and their lightlike points may be regarded as singular points of such metrics. In this paper, we investigate the beh…
It is classically known that the only zero mean curvature entire graphs in the Euclidean 3-space are planes, by Bernstein's theorem. A surface in Lorentz-Minkowski 3-space R13 is called of mixed type if it changes causal type from space-like to time-like. In R13, Osamu Kobayashi found …
Bayesian test assesses dependence between mixed data types.
problem Assessing dependence between text, image, and sound data.
method Bayesian kernelised correlation test using Dirichlet process model.
result Demonstrated effectiveness compared to other methods.
Study the geometry of lightlike loci on mixed type surfaces in Lorentz-Minkowski 3-space.
problem Characterize the differential geometric properties of lightlike loci on mixed type surfaces.
method Define a frame field and lightlike ruled surfaces along the lightlike locus, analyze their singularities and intersections.
result Establish a relationship between the singularities of lightlike ruled surfaces and the differential geometric properties of the lightlike locus.
Proposes CDTD, a diffusion model for mixed-type tabular data.
problem Adapting diffusion models to mixed-type tabular data.
method Score matching and score interpolation for continuous features, adaptive noise schedules for categorical features.
result Consistently outperforms state-of-the-art models in mixed-type tabular data.
This paper presents novel mixed-type Bayesian optimization (BO) algorithms to accelerate the optimization of a target objective function by exploiting correlated auxiliary information of binary type that can be more cheaply obtained, such as in policy search for reinforcement learning and hyperparameter tuning of machi…
Study on Bertrand lightcone framed curves in Lorentz-Minkowski 3-space.
problem Analyzing mixed types of curves with singular points in Lorentz-Minkowski 3-space.
method Using lightcone frame to consider Bertrand types for lightcone framed curves.
result Existence conditions of Bertrand lightcone framed curves in all cases.
Cohomogeneity-one actions on symmetric spaces of mixed type
problem Classifying cohomogeneity-one actions on symmetric spaces
method Using a new family of diagonal cohomogeneity-one actions
result Reducing the classification problem to symmetric spaces of a single type
We construct embedded triply periodic zero mean curvature surfaces of mixed type in the Lorentz-Minkowski 3-space with the same topology as the Schwarz D surface in the Euclidean 3-space.
We prove the existence of C^{\infty} local solutions to a class of mixed type Monge-Ampere equations in the plane. More precisely, the equation changes type to finite order across two smooth curves intersecting transversely at a point. Existence of C^{\infty} global solutions to a corresponding class of linear mixed ty…
A new method clusters mixed-type data efficiently.
problem Clustering mixed-type data with continuous and categorical variables.
method Extends Information Bottleneck principle to heterogeneous data using generalised product kernels.
result DIBmix outperforms four established methods in various scenarios.
A new method generates mixed-type features in tabular data with improved realism and accuracy.
problem Generating mixed-type features combining discrete and continuous data is challenging.
method A cascaded approach: first generates low-resolution categorical and coarse numerical features, then uses these in a high-resolution flow matching model.
result The model significantly improves detection scores, generating more realistic samples and capturing distributional details.
In this paper we outline a general method for finding well-posed boundary value problems for linear equations of mixed elliptic and hyperbolic type, which extends previous techniques of Berezanskii, Didenko, and Friedrichs. This method is then used to study a particular class of fully nonlinear mixed type equations whi…
The article proposes modified Gower's coefficients for handling mixed type variables in nearest neighbor methods.
problem Handling mixed type variables in nearest neighbor methods, especially imputation and statistical matching.
method Suggests modifications to the Gower's distance for interval and ratio scaled variables to address unbalanced contributions and outlier sensitivity.
result Improved distance calculations reduce the unbalanced contribution of different variable types and attenuate outlier effects.
MMM model clusters mixed-type longitudinal data efficiently.
problem Challenges in clustering multivariate longitudinal mixed-type data.
method MMM model reorganizes data into a three-way structure, using a mixture of matrix-variate normal distributions.
result MMM model handles various data types (continuous, ordinal, binary, nominal, count) and temporal dependence.
Outlier detection amounts to finding data points that differ significantly from the norm. Classic outlier detection methods are largely designed for single data type such as continuous or discrete. However, real world data is increasingly heterogeneous, where a data point can have both discrete and continuous attribute…
We introduce the DP-auto-GAN framework for synthetic data generation, which combines the low dimensional representation of autoencoders with the flexibility of Generative Adversarial Networks (GANs). This framework can be used to take in raw sensitive data and privately train a model for generating synthetic data that …
Study compares clustering methods for mixed-type data.
problem Challenges in clustering mixed-type data.
method Distance-based (k-prototypes, PDQ, convex k-means), probabilistic (KAY-means, MBNs, LCM).
result KAMILA, LCM, and k-prototypes perform best.
We focus on the problem of unsupervised cell outlier detection and repair in mixed-type tabular data. Traditional methods are concerned only with detecting which rows in the dataset are outliers. However, identifying which cells are corrupted in a specific row is an important problem in practice, and the very first ste…
SNI framework for mixed-type data imputation interprets and explains missing values.
problem Missing data in mixed-type databases skew analysis results.
method SNI couples statistical priors with neural attention to impute and explain missing values.
result SNI provides interpretable feature dependency diagnostics and soft regularization of attention.
Revisit Fenn's table theorem from a differential-topological perspective.
problem Prove zero-existence theorem on a cylinder and horizontal square-table theorem under Fenn's boundary conditions.
method Differential-topological approach.
result Prove horizontal square-table theorem under more general boundary conditions.
The study introduces a holdout-based framework to assess synthetic data fidelity and privacy.
problem Evaluating the quality and privacy of synthetic data solutions for mixed-type tabular data.
method Holdout-based empirical assessment framework measuring fidelity and privacy risk.
result Synthetic data samples are as close to the training as to the holdout data, indicating generalization and independence from individual records.
VAEM extends VAEs to handle mixed-type data heterogeneity.
problem Heterogeneous data with different types and marginal distributions.
method Two-stage training approach to handle mixed-type data.
result VAEM improves deep generative model performance on diverse tasks.
Improves Gower's similarity for mixed-type variables with automatic weighting.
problem Handling missing values and unbalanced variable contributions in Gower's similarity for mixed-type data.
method Automatic weighting scheme minimizing differences in correlation between contributing dissimilarities and weighted Gower's dissimilarity.
result Improved performance in classification and imputation of missing values.
Study detects synthetic tabular data across different tables.
problem Detecting synthetic tabular data in varied tables.
method Four table-agnostic detectors combined with preprocessing schemes.
result Cross-table learning possible with naive preprocessing, but cross-table transfer challenging.
A novel graph spectral method for mixed categorical and numerical data.
problem Feature learning for mixed data types (numerical and categorical).
method Graph spectral decomposition of the graph Laplacian to model probabilistic dependence structure.
result Increased separability and clusterability of observations in the transformed feature space.
Proves a generalized table theorem for odd Euler characteristic surfaces.
problem Proving a generalized table theorem for surfaces with odd Euler characteristic.
method Using the square peg problem for smooth curves, the result is generalized to real valued functions on Riemannian surfaces with odd Euler characteristic.
result Proves the table conjecture for even functions on the two sphere.
CTSyn generates high-quality synthetic tabular data.
problem Challenges in generating high-quality synthetic tabular data.
method Diffusion-based generative foundation model with autoencoder and conditional latent diffusion.
result CTSyn outperforms existing table synthesizers on standard benchmarks.
Upper bounds for surface-links in the Yoshikawa table are estimated.
problem Estimating Kirby-Thompson invariants of surface-links.
method Using tri-plane diagrams and L-, L*-invariants.
result Upper bounds for surface-links in the Yoshikawa table are obtained.
This paper compiles and calculates triple point numbers for surface-links in Yoshikawa's table.
problem Determining the triple point number of surface-links in Yoshikawa's table.
method Using broken sheet diagrams, the paper compiles known triple point numbers and calculates or bounds the remaining ones.
result Compilation and calculation of triple point numbers for surface-links in Yoshikawa's table.
The paper proves geometric properties of square tables and saddle surfaces.
problem The mathematical table problem from a geometric-topological perspective.
method Geometric-topological proofs on cylinder, saddle surfaces, and level sets of Fenn graphs.
result Zero-existence theorem on a cylinder, proving Fenn's square-table theorem under different boundary conditions.
Given data over the joint distribution of two random variables X and Y, we consider the problem of inferring the most likely causal direction between X and Y. In particular, we consider the general case where both X and Y may be univariate or multivariate, and of the same or mixed data types. We take an inf…
Machine learning speeds up search procedures for sorted tables.
problem Improving the speed of sorted table search procedures.
method Systematic experimental comparison of efficient implementations with learned counterparts.
result Learned data structures can significantly speed up search procedures.
Novelty detection is the unsupervised problem of identifying anomalies in test data which significantly differ from the training set. Novelty detection is one of the classic challenges in Machine Learning and a core component of several research areas such as fraud detection, intrusion detection, medical diagnosis, dat…
In this paper, we present smooth examples of degenerate hyperbolic and mixed type Monge-Ampere equations in the plane, which do not admit a local C^3 solution.
Unified multitask learning framework for mixed-type outcomes.
problem Difficulty in formulating a unified objective for tasks with different outcomes.
method Multitask transformation framework with shared sparsity, using deep neural networks and rank-based optimization.
result Improved prediction and variable selection across continuous, binary, and mixed outcomes.
DCRL learns causal relationships from mixed-type discrete data.
problem Challenges in learning causal relationships from discrete, mixed-type data.
method Generative framework modeling directed acyclic graph and sparse bipartite graph, flexible measurement models for different types of data.
result Consistent recovery of latent causal structure from observed data distribution.
Study on focal surfaces of lightcone framed surfaces in Lorentz-Minkowski 3-space.
problem Investigate differential geometry properties of focal surfaces of lightcone framed surfaces.
method Introduced lightcone frame to define lightcone framed surfaces, then investigated their differential geometry properties.
result Investigated differential geometry properties of focal surfaces of lightcone framed surfaces.
Method finds differential equations for integrable billiard tables.
problem Finding differential equations for integrable billiard tables.
method Introducing a method to find differential equations for functions defining tables.
result Illustrated method in three billiard systems.
We construct triply periodic zero mean curvature surfaces of mixed type in the Lorentz-Minkowski 3-space, with the same topology as the triply periodic minimal surfaces in the Euclidean 3-space, called Schwarz rPD surfaces.