The paper analyzes relationships between various image models for restoration.
problem Unclear relationships among popular image models for restoration.
method Theoretical analysis and experimental study of image models.
result Improved denoising performance by combining multiple image models.
Generative Adversarial Networks optimize model parameters for image matching.
problem Optimizing model parameters for accurate image matching.
method Model-Assisted Generative Adversarial Network (GAN) to produce fake images matching true images.
result Best match model parameter values can minimize bias in image recognition.
Latent feature models are attractive for image modeling, since images generally contain multiple objects. However, many latent feature models ignore that objects can appear at different locations or require pre-segmentation of images. While the transformed Indian buffet process (tIBP) provides a method for modeling tra…
Paper proposes a CNN-LSTM model for image denoising and reconstruction.
problem Challenging task of image denoising and reconstruction in computer vision.
method Proposes an encoder-decoder model with direct attention, using CNN for encoding and LSTM for decoding.
result Model can reconstruct clean images from highly corrupted ones, even when human understanding is difficult.
New model improves image captioning's ability to describe unseen concepts.
problem Image captioning models struggle with describing unseen combinations of concepts.
method Proposes a multi-task model combining caption generation and image-sentence ranking, with a decoding mechanism to re-rank captions based on image similarity.
result The model significantly outperforms state-of-the-art models in compositional generalization.
Develops statistical guarantees for image-to-image regression models.
problem Current image-to-image regression models lack statistical guarantees for model mistakes and hallucinations.
method Uncertainty quantification techniques with rigorous statistical guarantees for image-to-image regression problems.
result Derives uncertainty intervals around each pixel with formal mathematical guarantees.
Two-layer model sparsifies image residuals for CT image reconstruction.
problem Image reconstruction from limited and corrupted data.
method Pre-learning a two-layer sparsifying transform model with block coordinate descent optimization.
result Preliminary experiments show the two-layer model improves CT image reconstruction from low-dose measurements.
StrokeCoder uses Transformers to generate images from single examples.
problem Creating diverse images from a single example.
method Transformer Neural Network learns from a single path-based example to generate a set of images.
result The model can generate a large set of deviated images that still represent the original image's style and concept.
Proposes a model to generate 3D-aware images from 2D images.
problem Generating 3D-aware images from 2D images.
method Likelihood-based top-down model using Neural Radiance Fields and energy-based latent variables.
result Model can infer 3D object structures from 2D images and generate novel views.
This paper tackles generating manifold-valued images using WGAN.
problem Generating manifold-valued images over natural images.
method Formulated a theorem of optimal transport for Wasserstein distance on manifolds, introduced a new WGAN framework.
result Proposed model generates more plausible manifold-valued images than competitors.
CAFLOW uses auto-regressive flows to translate images efficiently.
problem Image-to-image translation tasks.
method Transforms conditioning image into latent encodings using normalizing flows, models conditional distribution with auto-regressive distributions.
result Outperforms former conditional flow designs.
Project analyzes images' impact on sentiment analysis.
problem Understanding how images contribute to sentiment classification.
method Compared models using only images, only text, or both.
result Combined models improved sentiment classification accuracy.
Improved image translation using asymmetric gradient guidance.
problem Trade-off between style transformation and content preservation in diffusion models.
method Asymmetric gradient guidance to guide reverse diffusion sampling.
result Our method outperforms state-of-the-art models in image translation tasks.
Generates coherent storybooks from plain text using diffusion models.
problem Ensuring coherency in a sequence of images for storytelling applications.
method Combines pre-trained LLM and text-guided Latent Diffusion Model for zero-shot generation.
result Outperforms state-of-the-art image editing baselines in generating coherent storybooks.
Generative model uses captions to generate images, improving semantic understanding.
problem Complex image generation models require large datasets and intricate learning.
method Adapts captioning models to generate images, using learned sentence and frame vectors.
result Images generated from multiple captions better capture semantic meaning.
InVA models image outcomes from multiple modalities, outperforming standard VAEs.
problem Understanding relationships across multiple imaging modalities in neuroimaging.
method Integrative Variational Autoencoder (InVA) framework for image-on-image regression.
result InVA accurately predicts PET scans from structural MRI, outperforming conventional models.
Image-to-image networks speed up SAR model parameter estimation.
problem Computational infeasibility of MLE for large, non-stationary spatial fields.
method Used image-to-image networks to estimate SAR model parameters.
result Image-to-image networks enable faster and more accurate parameter estimation.
Generative models improve image probability estimation but lack interpretability.
problem Lack of interpretability in generative models for natural image distributions.
method Extracted explicit probability density estimates from GANs and analyzed latent representations.
result Natural image density functions are difficult to interpret.
New method estimates image appearance models for segmentation.
problem Estimating appearance models for image segmentation.
method Tensor factorization-based estimator for latent variable models.
result Automatic estimation of image regions and proportions.
Framework generates realistic crop images for growth modeling.
problem Modeling crop growth over time with precision and detail.
method Two-stage framework: image prediction and growth estimation models.
result Framework accurately predicts crop images with varying conditions.
A new model answers questions about medical images.
problem Lack of transparency in deep learning models for medical imaging.
method A question-centric model that queries image models directly.
result The model achieves equal or higher accuracy than existing methods.
Gaudy images help train deep neural networks with less data.
problem Training deep neural networks with limited real data from visual cortex neurons.
method Used high-contrast binarized natural images (gaudy images) to train DNNs.
result Reduced training data needed for accurate DNN predictions of visual cortex neuron responses.
Self-guidance controls image generation by extracting properties from diffusion model representations.
problem Generating images from text descriptions is challenging due to the complexity of visual details.
method Self-guidance uses internal representations of diffusion models to control image generation.
result Properties like object shape, location, and appearance can be extracted and used to steer image generation.
Modeling the distribution of natural images is challenging, partly because of strong statistical dependencies which can extend over hundreds of pixels. Recurrent neural networks have been successful in capturing long-range dependencies in a number of problems but only recently have found their way into generative image…
A novel generative encoder model for imaging and image processing.
problem Efficiently processing and recovering images with noise.
method A pre-training phase with a GAN and an AE, followed by an optimization phase.
result The GE model outperforms state-of-the-art algorithms in image recovery.
Generative models improve image restoration from unknown transformations.
problem Restoring images distorted by unknown transformations.
method Combining maximum a-posteriori probability with maximum likelihood estimation.
result Restores images without requiring exact knowledge of transformations.
HW2MP-GAN tackles ancient handwritten text recognition.
problem Automatic text recognition from ancient handwritten records.
method Conditional Generative Adversarial Network (HW2MP-GAN) with Sliced Wasserstein distance and U-Net architectures.
result HW2MP-GAN outperforms state-of-the-art models in image-to-image translation and handwritten recognition.
Paper proposes efficient method for evaluating Bayesian models in imaging.
problem Evaluation of Bayesian models in imaging when ground truth is unavailable.
method Novel combination of Bayesian cross-validation and data fission for unsupervised model selection and misspecification detection.
result Achieved excellent selection and detection accuracy with low computational cost.
New model generates images by reversing heat equation, revealing disentanglement.
problem Image generation without considering image structure.
method Stochastically reverses the heat equation to generate images, using variational approximation.
result Emergent disentanglement of overall colour and shape in images.
Deep image clustering improved with STN and DAC.
problem Challenges in clustering images, especially with spatial transformations.
method Combining DAC with STN to reduce spatial transformation issues.
result The combined model outperformed baseline models on MNIST and FashionMNIST.
Generative model creates realistic images with 3D understanding.
problem Lack of 3D understanding in existing image generation models.
method Disentangled 3D representation using shape, viewpoint, and texture.
result Generates more realistic images and enables 3D operations.
Paper explores generalization of GAN image forensics methods.
problem Ensuring forensics models detect GAN-generated images across new types.
method Preprocessed images training for a forensic CNN model.
result Proposed method effectively detects GAN-generated images.
Generative models solve medical imaging inverse problems without needing paired data.
problem Reconstructing medical images from partial measurements.
method Score-based generative models trained on medical images, then sampling to reconstruct images consistent with measurements and physical model.
result Comparable or better performance in CT and MRI tasks, with improved generalization to unknown measurement processes.
This paper tackles blind image denoising with unknown noise models.
problem Real noisy images have complex noise models that are unknown beforehand.
method Proposes a novel Bayesian nonparametric prior called Dependent Dirichlet Process Tree to model the noise and a variational inference algorithm to recover clean patches.
result Achieves better performance compared to previous approaches on synthesis and real noisy images.
Zero-shot contrastive loss improves text-guided image style transfer without extra training.
problem Stochastic nature of diffusion models leads to trade-offs between style transformation and content preservation.
method Proposes a zero-shot contrastive loss for diffusion models that doesn't require additional fine-tuning or auxiliary networks.
result Method outperforms existing methods while preserving content and requiring no additional training.
Generative model creates meal images from ingredient descriptions.
problem Synthesize photo-realistic meal images from ingredient descriptions.
method Attention-based ingredients-image association model, cycle-consistent constraint.
result Model generates meal images corresponding to ingredient descriptions.
A new method improves robustness in image translation by modeling uncertainty.
problem Performance degradation in image translation models due to lack of robustness to outliers and uncertainty.
method UGAC method based on Uncertainty-aware Generalized Adaptive Cycle Consistency, modeling per-pixel residual with generalized Gaussian distribution.
result Our method exhibits stronger robustness towards unseen perturbations in test data.
I2SB learns nonlinear diffusion processes between images.
problem Image restoration tasks, especially with limited structural information.
method Conditional diffusion models, Schrödinger bridge approach.
result I2SB outperforms standard models in various image restoration tasks. Reduces GAN image priors' representation error using a Deep Decoder.
problem Representation error in GAN priors for in-distribution and out-of-distribution images.
method Hybrid model combining GAN prior and Deep Decoder.
result Consistently higher PSNRs on in-distribution and out-of-distribution images.
Unified model for image-to-image translation explained with new geometrical perspective.
problem Lack of solid theoretical interpretations for image-to-image translation models.
method Reformulated adversarial learning model from a geometrical perspective and extended generalization definition.
result Derived a condition to control the generalization capability of the model.
Generative model improves zero-shot sketch-based image retrieval.
problem Existing SBIR methods struggle with novel classes.
method Generative model learns to generate images conditioned on novel sketches.
result Significantly outperforms baselines on two challenging datasets.
Wavelets improve VAE image quality.
problem VAEs produce blurry images due to lack of high-frequency detail emphasis.
method Wavelet space VAE that emphasizes high-frequency components.
result Wavelet-based VAE generates higher quality images.
Develops a fast method to analyze image similarities without pre-trained models.
problem Lack of efficient methods to analyze fine-level similarities in large image datasets.
method Wavelet decomposition and numerical analysis tools.
result Identifies similar images in standard datasets (like CIFAR) in a few seconds.
ProAGAN stabilizes GANs for learning SOMs from noisy medical imaging data.
problem Learning stochastic object models from noisy and indirect medical imaging measurements.
method Developed Progressive Growing of AmbientGANs (ProAGAN) to stabilize GANs training.
result Signal detection performance improved using ProAGAN-generated images.
Bayesian algorithm detects image matches and fraud.
problem Detecting identity matches and fraud in image databases.
method Generative model of image graph trained with matching algorithm.
result Bayesian approach improves detection accuracy.
Develops a variational autoencoder for image, label, and caption modeling.
problem Deep learning of images, labels, and captions.
method Uses a Deep Generative Deconvolutional Network (DGDN) and a Convolutional Neural Network (CNN) for image and latent feature encoding.
result Able to predict labels or captions for new images using latent code distributions.
Develops a deep model for joint image-text learning.
problem Bidirectional joint image-text modeling.
method Variational hetero-encoder randomized GAN (VHE-GAN).
result Achieves state-of-the-art performance in image-text learning and generation.
Generative models create image perturbations to fool AI models.
problem Creating adversarial examples that fool pre-trained models.
method Trainable deep neural networks for image perturbation generation.
result High fooling rates with small perturbation norms, faster than current methods.