Method generates multiple 3D poses from 2D joint detections, addressing ambiguity and uncertainty.
problem Ambiguity and uncertainty in 3D human pose estimation from 2D joint detections.
method Generative model, compositional, anatomical constraints, removing model bias.
result Generates multiple valid 3D poses consistent with 2D joint detections.
Paper tackles in-bed pressure-based pose estimation, improving accuracy.
problem Pose estimation models fail to generalize with in-bed pressure data.
method End-to-end framework with a deep neural network pre-processing pressure data.
result Model accurately reconstructs unclear body parts for improved pose estimation.
The paper learns pose variations within shape populations using constrained mixtures of factor analyzers.
problem Learning pose variations within a shape population with articulated parts and relative rotations.
method Formulated as mixtures of factor analyzers, segmentation by component posterior probabilities, and constraints on factor loading matrices for rotation matrices.
result Automatic learning of pose variations from shape populations, resulting in smooth and realistic animations.
Generates realistic person images for re-id, overcoming pose variations.
problem Lack of cross-view paired training data and pose variations in person re-identification.
method Pose-normalization GAN (PN-GAN) for generating images conditioned on pose.
result Synthesized images enable learning invariant features free of pose variations.
Generative models improve pose transfer between people.
problem Transferring actions from one person to another.
method Used nearest neighbor and generative models (pix2pix) for pose transfer.
result Generative models outperform k-NN in generating corresponding frames and generalizing outside the action set.
The paper analyzes the observability of relative pose estimation using dual quaternions.
problem Estimating relative pose in robotics applications.
method Lie algebraic nonlinear observability analysis on a dual quaternion system.
result Dual quaternion representation yields an observability matrix with a simple block triangular structure and full rank.
New models disentangle body pose and appearance for flexible human body analysis.
problem Interpretable latent space for human body analysis.
method Conditional-DGPose and Semi-DGPose models for disentangled pose and appearance.
result Models enable independent manipulation of pose and appearance.
Capsule models enforce object pose relationships for robustness, explored with probabilistic generative and variational methods.
problem Enforcing object pose relationships for robustness to viewpoint changes.
method Probabilistic generative model with variational bound, exploring capsule assumptions and inference mechanisms.
result Unified objective and test time optimisation demonstrated for capsule models.
Unsupervised mesh disentanglement separates identity and pose.
problem Geometric disentanglement for 3D deformable models.
method CFAN-VAE architecture using conformal factor and normal features.
result CFAN-VAE achieves state-of-the-art performance on unsupervised geometric disentanglement.
Paper introduces EnDKF for more accurate pose tracking.
problem Accurate pose tracking with directional uncertainty.
method EnDKF integrates unit-quaternion attitude representation for better directional uncertainty capture.
result Significant reduction in error compared to traditional methods.
Paper generates high-resolution fashion images based on body pose.
problem Limited variety of outfit images for shopping.
method Generative model trained on body pose and style transfer.
result Realistic high-resolution images of models in custom outfits.
Generative model creates realistic dance poses from music.
problem Generating human-like dance poses from music.
method Music feature encoder, pose generator, music genre classifier integrated.
result Generative autoregressive model synthesizes dance sequences up to 5,000 frames.
DPNs learn pose-invariant object representations.
problem Pose-invariant 2D object recognition.
method Deformable Part Networks (DPNs) as sequences of LDPM units.
result 17-layer DPN outperforms CapsNets and STNs significantly on affNIST.
Optimizes ligand binding poses using CNNs and atomic grids.
problem Improving the accuracy of docking predictions for drug discovery.
method Differentiable atomic grid representation, CNN for scoring and optimization.
result Iteratively-trained CNNs outperform single CNNs in optimizing poses.
Extracts controllable models from videos of real-world activities.
problem Creating realistic and controllable character models from video data.
method Two networks: one for pose and control signal to next pose, and another for pose, new pose, and background to output frame.
result High-quality, controllable character models can be generated from arbitrary videos.
TIP model improves POSE prediction with less resources.
problem Predicting polypharmacy side effects from drug-protein interactions.
method TIP model operates on three subgraphs for progressive representation learning.
result Improves accuracy by 7%+, time efficiency by 83imes, and space efficiency by 3imes. Geometric Capsule Autoencoders group 3D points into parts and objects.
problem Learning object representations from 3D point clouds.
method Geometric capsules with pose and feature components, Multi-View Agreement voting mechanism.
result Learned representations enable object identification and canonical pose recovery.
Proposes a new model to capture joint influence of correlated events on user search behavior.
problem Real-world events influence each other and pose joint influence on user search behavior, not independent.
method Joint Influence Model based on Multivariate Hawkes Process.
result The model captures the temporal dynamics of joint influence and outperforms baseline methods.
Tensor neural network improves human pose classification from 3D skeleton data.
problem Efficiently processing spatiotemporal data for human pose classification.
method Proposes a tensor-based neural network with three components: spatiotemporal feature construction, tensor fusion, and tensor-based neural network processing.
result Achieves state-of-the-art performance in human pose classification.
KeypointNet learns 3D keypoints for object pose estimation without ground-truth.
problem Learning 3D keypoints for object pose estimation without manual annotations.
method End-to-end geometric reasoning framework to discover keypoints.
result End-to-end framework outperforms fully supervised baseline.
We integrate camera pose correlations into deep models using Gaussian processes.
problem Lack of inter-frame reasoning in deep neural networks.
method Derive a principled framework combining camera pose information with deep models using a novel view kernel.
result Soft-prior knowledge aids pose-related vision tasks like novel view synthesis.
Improves deep pose estimation for extreme motions using pre-trained models.
problem Accuracy and training time issues for extreme/wild motions.
method Post-data augmentation to improve pre-trained models.
result Enhanced accuracy in pose estimation for extreme motions.
Paper introduces a new method for generating diverse human motion predictions.
problem Stochastic human motion prediction with limited flexibility.
method Stochastically combines root variations with previous pose information in a recurrent network.
result Model generates more diverse motion sequences than existing techniques.
This thesis models and approximates pose distributions in robot perception using quaternion and Gaussian methods.
problem Modeling and approximating probability distributions of poses in robot perception.
method Uses dual quaternions, unit quaternions, and Gaussian distributions to represent and approximate pose distributions.
result A framework for probabilistic modeling of poses in S3imesR3 that approximates various probability distributions. DepthNets learns 3D face geometry and transformations without supervision.
problem Learning 3D face geometry and transformations from a single image.
method Unsupervised learning of facial keypoints depth, using backpropable loss for 3D transformations.
result DepthNets can predict 3D transformations and re-target faces to new poses or geometries.
Augment small datasets with synthetic backgrounds to train lightweight CNNs for human pose estimation.
problem Training CNNs from limited real-world data for human pose estimation.
method Synthetic background substitution for data augmentation.
result Improves generalization to unseen environments.
This paper separates static and dynamic features in video data.
problem Combining static and dynamic features in video data.
method Hierarchical Variational Auto-encoders with factored prior distributions.
result The model successfully separates static and dynamic features.
Max pooling selects frames with similar poses for better face recognition.
problem Measuring similarity of faces with varying poses in videos.
method Pose-Selective Max Pooling: Select frames closest to centroid of K-means cluster.
result Max correlation among selected features yields better performance than VGG-face.
Paper optimizes estimation of quadratic functionals in nonparametric IV models.
problem Optimal estimation of a nonlinear functional in ill-posed inverse regression.
method Adaptive, minimax estimation using leave-one-out, sieve NPIV estimator with data-driven sieve dimension selection.
result Adaptive estimator achieves minimax optimal rate in various ill-posed cases.
We introduce SARR for symmetric object pose estimation, improving CNN performance.
problem Ambiguities in symmetric object orientations hinder deep learning pose estimation.
method Numeric rotation representation using symmetry-derived trigonometric identities.
result SARR enables standard CNNs to achieve state-of-the-art performance.
MCGDiff uses SGM to guide SMC for solving ill-posed linear inverse problems.
problem Solving ill-posed linear inverse problems in Bayesian settings.
method Exploiting SGM structure, defining a sequence of intermediate problems, and using SMC methods.
result MCGDiff outperforms competing methods in Bayesian ill-posed inverse problems.
New model resolves signal ambiguities in ill-posed systems.
problem Signal retrieval from indirect measurements with known models.
method Variational generative model that captures signal distribution.
result Retrieves consistent signals with high fidelity.
Paper introduces a Gibbs sampler for Bayesian inversion of ill-posed problems.
problem Bayesian inversion of ill-posed problems with linear transformation and additive noise.
method Gibbs algorithm based on prior diffusion model.
result Gibbs algorithm offers a guarantee of convergence in a specific situation.
Scalable approach for object pose estimation across domains.
problem Object pose estimation across different datasets and models.
method Multi-path learning: shared encoder, object-specific decoders.
result Generalizes well from synthetic to real data and across various instances.
Score-based models improve diffuse optical tomography accuracy.
problem Improving accuracy in diffuse optical tomography with uncertainty quantification.
method Score-based diffusion models with a mixed score function to prevent overfitting.
result Data-driven prior distribution results in posterior samples with low variance and centred around the ground truth.
CSI-Net learns WiFi signals for body characterization and pose recognition.
problem Unified learning of body characteristics and pose recognition.
method Unified Deep Neural Network (DNN) for WiFi signal representation and multi-task learning.
result CSI-Net solves biometrics estimation and person recognition.
Robust visual tracking for long video sequences is a research area that has many important applications. The main challenges include how the target image can be modeled and how this model can be updated. In this paper, we model the target using a covariance descriptor, as this descriptor is robust to problems such as p…
New CNN architecture improves pediatric image segmentation by homogenizing pose and size.
problem Challenges in segmenting pediatric images due to pose and size heterogeneity.
method Spatial Transformer Network (STN) for pose and scale invariance, combined with UNet for segmentation.
result Improved pediatric segmentation, especially renal tumor delineation, with accelerated processing.
Aerial robot estimates human pose and path using dynamic classifier selection.
problem Estimating human pose and trajectory from aerial video.
method Dynamic classifier selection architecture; perspective correction; HOG and CNN features; 64 pose-viewpoint classes.
result Dynamic classifier selection improves efficiency and accuracy.
New algorithm optimizes pose graphs for SFM and SLAM.
problem Optimizing pose graphs for structure from motion and simultaneous localization and mapping.
method Tempered Geodesic Markov Chain Monte Carlo (TG-MCMC) algorithm.
result Robust initialization and uncertainty estimates for reliable solutions.
Maximum likelihood estimation fails to be well-posed in Gaussian process regression.
problem Establishing well-posedness of maximum likelihood estimation in Gaussian process regression.
method Analyzing the conditions under which maximum likelihood estimation is not Lipschitz in the data with respect to the Hellinger distance.
result Maximum likelihood estimation is not well-posed in the noiseless data setting for any Gaussian process with a stationary covariance function whose lengthscale parameter is estimated using maximum likelihood.
CP2 uses geometric information to improve conformal prediction robustness.
problem CP fails under geometric data shifts, losing coverage guarantees.
method Integrates geometric pose information into CP via canonicalization.
result Integrating geometric information with CP ensures robustness under geometric shifts.
Improved UAV navigation and landing using deep learning.
problem Autonomous navigation and landing of UAVs with high accuracy.
method Multimodal fusion of visual and inertial sensor data using deep neural networks.
result 25% improvement in pose estimation accuracy compared to traditional methods.
This paper formulates and studies a general continuous-time behavioral portfolio selection model under Kahneman and Tversky's (cumulative) prospect theory, featuring S-shaped utility (value) functions and probability distortions. Unlike the conventional expected utility maximization model, such a behavioral model could…
We consider the problem of learning object arrangements in a 3D scene. The key idea here is to learn how objects relate to human poses based on their affordances, ease of use and reachability. In contrast to modeling object-object relationships, modeling human-object relationships scales linearly in the number of objec…
P3I learns holistic scene representations from a single image.
problem Inferring camera poses, object locations, and global scene structures from a single image.
method Combines search-based and gradient-based algorithms.
result P3I outperforms baselines on various image manipulation tasks.
ES-VAE models skeletal pose trajectories by removing nuisance factors.
problem Handling camera orientation, subject scale, viewpoint, and execution speed in skeletal data.
method ES-VAE uses TSRVF representation on Kendall's shape manifold to isolate shape dynamics.
result ES-VAE outperforms standard VAEs and sequence modeling baselines in gait cycle prediction and action recognition.
Study Tikhonov regularization for ill-posed surface equations, analyzing perturbations and applications.
problem Solving ill-posed operator equations with solutions on surfaces.
method Error analysis of Tikhonov regularization considering surface perturbations and vector bundle solutions.
result Error analysis and practical applications demonstrated for functions on surfaces.