Optimizes master faces for 2D and 3D face verification using evolutionary algorithms and neural networks.
problem Impersonation attacks using master faces for face-based identity authentication.
method Evolutionary algorithm in latent space of StyleGAN, neural network to direct search, 2D and 3D face reconstruction.
result Obtains high impersonation rates with fewer master faces for 2D and 3D face verification.
DepthNets learns 3D face geometry and transformations without supervision.
problem Learning 3D face geometry and transformations from a single image.
method Unsupervised learning of facial keypoints depth, using backpropable loss for 3D transformations.
result DepthNets can predict 3D transformations and re-target faces to new poses or geometries.
Enhances 2D face recognition with 3D features using active illumination.
problem Improving robustness of 2D face recognition to spoofing attacks and low-light conditions.
method Projecting a high spatial frequency pattern onto the face to recover 3D information and a 2D image simultaneously.
result Significantly boosts face recognition performance and dramatically improves robustness to spoofing attacks.
This work generates synthetic 3D thermal facial data using 2D facial data and deep learning.
problem Creating large datasets for deep learning in computer vision.
method 3D facial modelling techniques and deep learning methodologies.
result Synthetic 3D thermal facial data created for deep learning applications.
PolyGen models 3D meshes directly, predicting vertices and faces sequentially.
problem Efficiently modeling 3D geometry for computer graphics, robotics, and games.
method Transformer-based autoregressive model for predicting mesh vertices and faces.
result PolyGen produces high-quality, usable 3D meshes and competitive conditional performance.
Energy-based models can generate complex images by combining simpler concepts.
problem Generating natural images that satisfy complex logical combinations of concepts.
method Energy-based models combine probability distributions of simpler concepts to generate compositions.
result Energy-based models can generate images that satisfy conjunctions, disjunctions, and negations of concepts.
Model manipulates facial expressions without affecting other attributes.
problem Manipulating specific visual attributes in real scenes without altering others.
method Trains model on nonphotorealistic 3D renders to manipulate facial expressions, preserving other attributes.
result Model can manipulate facial expressions without affecting other attributes like head orientation.
3D quantum trace map connects 3-manifold quantizations.
problem Quantization of 3-manifold character varieties.
method Study of stated skein modules and face suspensions.
result Existence of 3D quantum trace map proved.
Propose a new 3d quantum trace map that agrees with Garoufalidis and Yu's construction and extends to certain manifolds with ideal triangulated boundaries.
problem Relationship between two constructions of 3d quantum trace maps.
method Propose a new 3d quantum trace map.
result Proposed 3d quantum trace map agrees with Garoufalidis and Yu's construction and extends to certain manifolds with ideal triangulated boundaries.
We study geometric consistency relations between angles on 3-dimensional (3D) circular quadrilateral lattices -- lattices whose faces are planar quadrilaterals inscribable into a circle. We show that these relations generate canonical transformations of a remarkable ``ultra-local'' Poisson bracket algebra defined on di…
Researchers visualize all surfaces from tesseract faces.
problem Visualizing all surfaces from tesseract faces.
method Generated 3D models of all closed surfaces and exhibited the Möbius strip.
result Generated and visualized all surfaces from tesseract faces.
Study uses stacked hourglass networks to improve facial landmark detection for medical diagnosis.
problem Improving accuracy of facial landmark detection for medical diagnosis.
method Conducted a study on landmark localisation methods using stacked hourglass networks.
result State-of-the-art stacked hourglass architecture outperforms traditional methods.
The paper uses facial keypoints to estimate post-surgical pain intensity.
problem Accurately assessing pain levels from self-reported ratings is challenging.
method The approach analyzes 2D and 3D facial keypoints to estimate pain intensity.
result The pain estimation model uses multiple instance learning.
ED-NeRF efficiently edits 3D scenes using latent space NeRF and improved loss functions.
problem Slow training speeds and inadequate editing loss functions in existing NeRF editing techniques.
method Embedding real-world scenes into latent space of LDM, using a unique refinement layer and an improved loss function.
result ED-NeRF achieves faster editing speed and improved output quality compared to state-of-the-art models.
Unified description of tetrahedra in various spacetimes.
problem Characterizing tetrahedra with lightlike faces in different spacetimes.
method Using a generalized cross-ratio and edge lengths/dihedral angles to describe tetrahedra and their duals.
result Generalized ideal tetrahedra are the duals of tetrahedra with lightlike faces.
Constructs examples of complex 3D shapes with specific properties.
problem Creating fibered three-manifolds with certain characteristics.
method Builds examples using handlebody bundles and polytopes.
result Examples of fibered three-manifolds with specific properties.
Paper introduces a method for generating interlocutor-aware facial gestures in dyadic settings.
problem Generating appropriate non-verbal behavior for conversational agents in dyadic settings.
method Probabilistic method using multi-modal cues from the interlocutor to synthesize facial gestures.
result The model successfully leverages multi-modal input from the interlocutor to generate more appropriate behavior.
Enhanced 3D shape analysis using information geometry.
problem Challenges in comparing 3D point clouds due to their unstructured nature and complex geometry.
method Information geometric framework for 3D point cloud shape analysis using Gaussian Mixture Models (GMMs) on a statistical manifold. Proposed MSKL divergence with upper and lower bounds.
result MSKL provides stable and monotonically varying values that directly reflect geometric variation, outperforming traditional distances and existing KL approximations.
Generative model learns to autoencode and generate sets of images.
problem Learning to represent and generate sets of images with unknown number of sets.
method Set Distribution Networks (SDNs) learn set encoder, discriminator, generator, and prior.
result SDNs can reconstruct and generate sets of images with preserved attributes.
Paper proposes Roweisposes for 3D action recognition using generalized eigenvalue problem.
problem Need for basic methods in 3D action recognition.
method Roweisposes uses Roweis discriminant analysis for generalized subspace learning.
result Roweisposes is effective for 3D action recognition.
Elastic-InfoGAN learns object identity in class-imbalanced data.
problem Learning disentangled representations in class-imbalanced data.
method Invariance to identity-preserving transformations to learn object identity.
result Effectiveness in disentangling object identity in imbalanced datasets.
Enhances curve alignment for diverse data types.
problem Aligning curve data effectively.
method Developed nonlinear transformations for curve data.
result Successfully aligned synthetic and real curve data.
Researchers classify and visualize 5-cube cubical surfaces.
problem Classifying and visualizing surfaces in a 5-dimensional cube.
method Exhaustive search, classification by genus and demigenus, 3D visualization, reinforcement learning for optimization.
result 2690 connected closed cubical surfaces in the 5-cube, visualized and optimized for 3D printing.
Proposes polynomial neural networks for improved function approximation in various tasks.
problem Improving function approximation in various tasks like image generation, face verification, and 3D mesh representation learning.
method Introduces polynomial neural networks (Π-Nets) and three tensor decompositions to reduce parameter count and enhance expressiveness. result Demonstrates that Π-Nets can produce state-of-the-art results in challenging tasks without non-linear activation functions. FEAFA dataset annotates facial expressions with high detail.
problem Lack of detailed facial expression annotations in existing datasets.
method Manual annotation of 122 participants' facial expressions.
result FEAFA dataset provides detailed annotations for facial expressions.
Develops a fast non-invasive tool for diagnosing pediatric sleep apnea.
problem Diagnosing pediatric obstructive sleep apnea using an overnight sleep study is often impractical.
method Combines persistent homology, geometric shape analysis, and convolutional neural networks to classify facial images.
result Facial features associated with obstructive sleep apnea can be recognized for diagnosis.
3D Adversarial Autoencoder learns compact binary descriptors from 3D point clouds.
problem Learning meaningful representations of 3D shapes for various tasks.
method End-to-end Adversarial Autoencoder model trained on 3D input and output.
result 3D Adversarial Autoencoder (3dAAE) generates state-of-the-art results for 3D points clustering and retrieval.
Automates detection of electric devices in 3D x-ray images of luggage.
problem Detecting electric devices in cluttered 3D baggage images.
method Unpack, Predict, eXtract, Repack (UXPR) algorithm using segmentation and ensemble learning.
result System can accurately detect electric devices in 3D baggage images.
3D dual field theories for Virasoro minimal models constructed using Seifert fiber spaces.
problem Constructing 3D dual field theories for Virasoro minimal models.
method 3D-3D correspondence and Seifert fiber spaces.
result 3D dual field theories constructed for Virasoro minimal models.
DreamFusion uses text-to-image diffusion models to create 3D images efficiently.
problem Lack of large-scale 3D datasets and efficient architectures for 3D synthesis.
method Adapting a 2D diffusion model to 3D synthesis using a loss based on probability density distillation.
result A 3D model can be optimized from a 2D diffusion model, allowing for text-to-3D synthesis.
By using two different invariants for the Rubik's Magic puzzle, one of metric type, the other of topological type, we can dramatically reduce the universe of constructible configurations of the puzzle. Finding the set of actually constructible shapes remains however a challenging task, that we tackle by first reducing …
Flexible pipeline for 3D vehicle detection from 2D images.
problem Current methods lack 3D perception of vehicles and other objects.
method Adopt any 2D detection network, fuse with 3D point cloud, develop model fitting algorithm, refine with CNN.
result 3D detection results rank second among algorithms, demonstrating competencies.
One of the key challenges of visual perception is to extract abstract models of 3D objects and object categories from visual measurements, which are affected by complex nuisance factors such as viewpoint, occlusion, motion, and deformations. Starting from the recent idea of viewpoint factorization, we propose a new app…
Generative model disentangles 3D shapes into independent factors.
problem Learning rich representations of deformable 3D shapes.
method Supervised 3D mesh-convolutional Variational AutoEncoder with latent feature disentanglement.
result Explicit disentanglement of latent factors improves shape generation and downstream tasks.
The field of multiple view geometry has seen tremendous progress in reconstruction and calibration due to methods for extracting reliable point features and key developments in projective geometry. Point features, however, are not available in certain applications and result in unstructured point cloud reconstructions.…
Generates coherent 3D scenes from monocular videos without supervision.
problem Lack of 3D scene modeling in video generation models.
method Trains a model to generate 3D scenes with moving objects and a background from monocular videos.
result Trained model generates coherent 3D scenes with multiple moving objects and a background.
Paper develops a differentiable approach for 3D imaging models using Fourier slice theorem.
problem Uncertainty in 3D structure modeling and pose estimation in scientific imaging.
method Differentiable probabilistic models in Fourier space with backpropagation through projection.
result Validates approach on 3D protein reconstruction and extends to probabilistic models.
GCDM generates valid large 3D molecules and optimizes existing molecules.
problem Lack of geometric properties in 3D molecule generation models.
method Introduces Geometry-Complete Diffusion Model (GCDM) using equivariant GNNs.
result Significantly outperforms existing models in 3D molecule generation and optimization.
Generative model creates detailed 3D shapes from text descriptions.
problem Creating high-resolution 3D models from natural language descriptions.
method Two-step process: first generating low-resolution shapes, then high-resolution shapes using Conditional Wasserstein GAN framework.
result Improved method generates 3D shapes more faithful to natural language.
3D object detection improved using energy-based models.
problem Accurate 3D object detection in cluttered environments from sparse LiDAR data.
method Designing a differentiable pooling operator for 3D bounding boxes integrated into a state-of-the-art 3D object detector.
result Our approach consistently outperforms the SA-SSD baseline across all 3DOD metrics on the KITTI dataset.
Equivariant diffusion model generates 3D molecules efficiently.
problem Generating high-quality 3D molecules efficiently.
method Equivariant Diffusion Model (EDM) that operates on atom coordinates and types.
result Significantly outperforms previous methods in molecule quality and training efficiency.
3D Axial-Attention improves lung nodule classification accuracy.
problem Limited 3D attention in existing methods.
method Proposes 3D Axial-Attention network with 3D positional encoding.
result 3D Axial-Attention achieves state-of-the-art performance.
ROOTS learns to represent and render 3D scenes with object-centric models.
problem Learning to represent and render 3D scenes with object-centric compositionality.
method Probabilistic generative model for learning object representations and scene rendering from partial observations.
result The model can infer 3D object representations and render scenes from arbitrary viewpoints.
Paper presents a consistent discretization for Hodge decomposition on volumetric meshes.
problem Discretization of Hodge decomposition for vector fields on volumetric meshes.
method Edge-based Nedelec elements and face-based Crouzeix-Raviart elements interplay.
result Stable and efficient method for large-sized models with good performance.
Paper tackles unsupervised learning of 3D shapes from single images.
problem Learning 3D shapes from single images without supervision.
method Generative models, variational auto-encoders, adversarial methods.
result Model learns 3D shapes and poses from single images, showing potential for various datasets.
GIBLy adds geometric priors to 3D segmentation models, improving performance with minimal overhead.
problem Lack of explicit geometric information in 3D semantic segmentation models.
method Introduces GIBLy, a lightweight geometric inductive bias layer that integrates learnable geometric priors into existing 3D segmentation pipelines.
result Consistent performance gains across multiple benchmarks, including up to +11.5% mIoU on TS40K with PTV3.
Paper extends 2D ZSAD to 3D MRI without training, achieving robust anomaly detection.
problem Challenges in extending zero-shot anomaly detection to 3D medical images.
method Constructs localized volumetric tokens by aggregating 2D slices processed by 2D foundation models.
result Training-free, batch-based ZSAD effectively extends from 2D encoders to full 3D MRI volumes.
3D topological models link to HOMFLYPT homology via braids.
problem Understanding topological invariants of links and braids.
method 3D topological B-models with Hilbert schemes of points.
result Hilbert space of braids corresponds to HOMFLYPT homology.