Paper proposes a method to enhance low-quality retinal images using optimal transport.
problem Artifacts and imperfections in retinal images lead to diagnostic inaccuracies.
method Leveraging optimal transport theory, an unpaired image-to-image translation scheme is proposed.
result The method improves perceptually and quantitatively the quality of low-quality retinal images.
EdgeFool generates adversarial images to mislead classifiers.
problem Misleading classifiers with adversarial images.
method Trains a fully convolutional neural network to generate perturbations that enhance image details and mislead classifiers.
result EdgeFool outperforms other adversarial methods on various classifiers and datasets.
Enhances image quality to improve test-time adaptation accuracy.
problem Reducing accuracy loss due to distribution shift in deep networks.
method Integrates image enhancement with TTA methods to reduce prediction uncertainty.
result TECA method increases accuracy of TTA methods without hyperparameters.
OTRE uses OT to improve retinal images, outperforming existing methods.
problem Improving quality of non-mydriatic retinal images for accurate diagnoses.
method OT theory for image-to-image translation, regularization by enhancing.
result OTRE outperforms state-of-the-art methods on various retinal image tasks.
CURL uses neural curves to enhance global image properties.
problem Global image enhancement using neural networks.
method CURL is a multi-colour space neural retouching block trained in HSV, CIELab, and RGB color spaces.
result CURL produces state-of-the-art image quality in RGB-to-RGB and RAW-to-RGB transformations.
Enhanced rotation prediction improves SSL models by capturing both shape and texture information.
problem Rotation prediction misses texture information, limiting model performance.
method Introduces image enhanced rotation prediction (IE-Rot) that combines rotation and image enhancement tasks.
result IE-Rot models outperform Rotation on various benchmarks.
Enhances high-throughput imaging of microtubule networks, improving clarity and consistency.
problem Fluorescence noise obscures microtubule structures in high-throughput imaging.
method CycleGAN learning to enhance low-resolution images of microtubule networks.
result CycleGAN effectively identifies microtubules with high accuracy (0.93+ AUC-ROC).
Paper introduces DACAL for high-resolution photo and video enhancement.
problem Photo and video enhancement with weak supervision.
method Divide-and-conquer adversarial learning approach with hierarchical decomposition.
result State-of-the-art performance in high-resolution photo and video enhancement.
Style generators can invertibly generate images, useful for image enhancement and animation.
problem Generating all possible images and maintaining invertibility are challenging for GANs.
method Proposed style generators are invertible and can generate all possible images.
result Style generators outperform other GANs and Deep Image Prior as image priors.
Enhances trading signals using image analysis and weighted moving averages.
problem Improving price trend trading strategies in financial markets.
method Image-induced importance weights applied to weighted moving averages of trading signals.
result Significant enhancement of price trend trading signals with improved portfolio selection.
Enhances neural networks' robustness against adversarial samples without sacrificing clean sample generalization.
problem Limited generalization and time complexity of adversarial training.
method Feature Pyramid Decoder (FPD) framework that integrates denoising and image restoration modules into CNNs and constrains the Lipschitz constant.
result FPD-enhanced CNNs achieve sufficient robustness against general adversarial samples on various datasets.
Enhances image-to-image translation using adversarial latent space.
problem Image-to-image translation task in computer vision.
method Introduces an adversarial discriminator on the latent representation to enforce similar latent space distributions.
result Significantly outperforms competing approaches on MNIST and USPS domain adaptation tasks.
Enhances facial emotion recognition with gradient and Laplacian images.
problem Improving the performance of facial emotion recognition systems.
method Proposes using gradient and Laplacian of input images with a CNN.
result Enhances FER systems by 3 to 5%.
Deep learning methods quantify uncertainty in neuroimage enhancement.
problem Uncertainty in deep learning models for medical image enhancement.
method Heteroscedastic noise model and approximate Bayesian inference for uncertainty quantification.
result Uncertainty quantification improves predictive performance and risk assessment.
Self-supervised method enhances ultrasound images without needing clean targets.
problem Multiplicative speckle, acquisition blur, and scanner artifacts hamper ultrasound interpretation.
method Physics-guided degradation model trained on rotated/cropped patches with synthesized inputs.
result Achieves highest PSNR/SSIM across Gaussian and speckle noise levels, with significant improvements in heavy noise conditions.
Enhances uncertainty estimation in medical image segmentation.
problem Frequency-related noise in medical imaging leads to biased uncertainty estimates.
method Extends MC-Dropout to the frequency domain for better uncertainty estimation.
result MC-Frequency Dropout improves calibration and uncertainty in semantic segmentation.
Improved image generation through iterative flow matching to reduce hallucinations.
problem Hallucinations in image generation models.
method Iterative flow matching to refine and correct paths in generative models.
result Enhanced generative modeling with reduced unrealistic images.
Enhances curve alignment for diverse data types.
problem Aligning curve data effectively.
method Developed nonlinear transformations for curve data.
result Successfully aligned synthetic and real curve data.
IEBN normalizes noise by enhancing instance-specific information, improving deep learning performance.
problem Improving deep learning performance by regulating noise in batch normalization.
method Integrates self-attention mechanism to recalibrate channel information in BN.
result IEBN outperforms BN with improved generalization and stability.
Enhances MRI image quality with Conditional WGAN and adaptive balancing.
problem Struggles to reconstruct sharp images with fine detail.
method Conditional Wasserstein Generative Adversarial Network (WGAN) with Adaptive Gradient Balancing.
result Produces sharper images than other techniques.
Method matches noisy remote sensing images robustly.
problem Matching noisy remote sensing images.
method Combining attention mechanism with feature enhancement.
result More efficient and accurate matches achieved.
Deep learning has brought an unprecedented progress in computer vision and significant advances have been made in predicting subjective properties inherent to visual data (e.g., memorability, aesthetic quality, evoked emotions, etc.). Recently, some research works have even proposed deep learning approaches to modify i…
Enhances deep networks robustness with data mollification and label smoothing.
problem Improving deep neural networks' robustness against corruptions.
method Coupling data mollification (image noising and blurring) with label smoothing.
result Improved robustness and uncertainty quantification on corrupted image benchmarks.
CeCNN predicts SE and AL from UWF images, improving myopia screening.
problem Predicting axial length and spherical equivalence from UWF fundus images.
method Copula-enhanced Convolutional Neural Network (CeCNN) for multiresponse regression.
result CeCNN improves prediction of SE and AL compared to baseline CNNs.
Boomerang generates nonidentical images similar to input on image manifolds.
problem Generating nonidentical images similar to input on image manifolds.
method Adding noise to input image, moving closer to latent space, and mapping back through partial reverse diffusion.
result Boomerang generates nonidentical images similar to input on image manifolds.
Total variation and mean curvature flows on a Lie group quotient enhance and denoise crossing structures.
problem Preserving crossing curvilinear structures in image enhancement and denoising.
method Lifting images to the homogeneous space M=RdtimesSd−1, applying PDEs for TVF and MCF, and using locally optimal differential frames. result Better preservation of bundle boundaries and angular sharpness in fiber orientation densities at crossings compared to data-driven diffusions.
Enhances VAEs for sharper image synthesis.
problem Blurriness in generated images from VAEs.
method Integrates a downscaled version of the original image into the VAE framework and uses it as input to the decoder.
result Improves FID score in image synthesis while maintaining similar log-likelihood performance.
Deep learning improves 3D microscopy resolution without matched target images.
problem Anisotropic resolution in volumetric fluorescence microscopy.
method Cycle-consistent generative adversarial network trained on unpaired 2D images.
result Enhanced axial resolution and restored details between imaging planes.
A 3-stage method enhances hyperspectral image classification accuracy.
problem Classifying detailed classes in hyperspectral images with limited labeled data.
method Uses Nested Sliding Window and PCA for spatial consistency, SVM for spectral estimation, and TV model for spatial smoothing.
result Our method outperforms state-of-the-art algorithms, especially in scenarios with small training sets.
Enhances few-shot image classification using unlabelled examples.
problem Few-shot image classification with limited labeled data.
method Transductive meta-learning combining soft k-means clustering and neural feature extractor.
result State-of-the-art performance on Meta-Dataset, mini-ImageNet, and tiered-ImageNet benchmarks.
Enhanced deep learning model improves tumor segmentation in ultrasound images.
problem Challenges in integrating patient-specific medical priors into deep learning models.
method Integrates visual saliency into a U-Net architecture with attention blocks.
result Achieved a Dice similarity coefficient of 90.5 percent on a dataset of 510 images.
This master thesis focuses on practical application of Convolutional Neural Network models on the task of road labeling with bike attractivity score. We start with an abstraction of real world locations into nodes and scored edges in partially annotated dataset. We enhance information available about each edge with pho…
Obtaining magnetic resonance images (MRI) with high resolution and generating quantitative image-based biomarkers for assessing tissue biochemistry is crucial in clinical and research applications. How- ever, acquiring quantitative biomarkers requires high signal-to-noise ratio (SNR), which is at odds with high-resolut…
Scene text magnifier enhances readability for visually impaired.
problem Helps visually impaired read natural scene text.
method Four CNN-based networks: character erasing, extraction, magnify, synthesis.
result Effective text magnification without background alteration.
Extended RDS filtering for positions and orientations, improving crossing structure enhancement and inpainting.
problem Enhancing and inpainting images with crossing structures.
method Extended RDS filtering to M2 space, using gauge frames to mitigate issues. result RDS filtering outperforms existing techniques in denoising and inpainting images with crossing structures.
New polynomials defined for quandle structures, enhancing graph invariants.
problem Enhancing the counting invariant for spatial graphs and handlebody-links.
method Introducing quandle polynomials and G-family polynomials for quandles, defining enhancements for invariants.
result New enhancements of the G-family counting invariant for trivalent spatial graphs and handlebody-links.
Enhances visual localization using graph smoothing.
problem Inferring camera pose from a single image.
method Constructs a graph based on GPS coordinates and temporal information, then smooths feature representations.
result Significantly improves localization accuracy on large datasets.
DEUA detects diffusion-generated images by accounting for different types of uncertainty.
problem Detecting generated images with varying aleatoric and epistemic uncertainty.
method DEUA framework using Laplace approximation for DEU estimation and asymmetric loss function.
result DEUA achieves state-of-the-art performance on large-scale benchmarks.
Acoustic scene classification is the task of identifying the scene from which the audio signal is recorded. Convolutional neural network (CNN) models are widely adopted with proven successes in acoustic scene classification. However, there is little insight on how an audio scene is perceived in CNN, as what have been d…
Enhanced VQ-VAE generates high-fidelity images faster.
problem Generating high-fidelity images efficiently.
method Scaled VQ-VAE with fast autoregressive sampling.
result VQ-VAE generates samples with quality rivaling GANs.
Improving speech system performance in noisy environments remains a challenging task, and speech enhancement (SE) is one of the effective techniques to solve the problem. Motivated by the promising results of generative adversarial networks (GANs) in a variety of image processing tasks, we explore the potential of cond…
Improved VGG networks enhance image classification accuracy.
problem Enhancing image classification accuracy using modified VGG architectures.
method Two improved VGG architectures were created by freezing the first two blocks and applying different dilation rates in the last three blocks.
result Significant out-performance on image classification tasks on CIFAR-10 and CIFAR-100 datasets.
In this paper, we consider biquandle colorings for knotoids in R2 or S2 and we construct several coloring invariants for knotoids derived as enhancements of the biquandle counting invariant. We first enhance the biquandle counting invariant by using a matrix constructed by utilizing the orientation a kno…
TSSC images enhance chaotic signal classification using ConvNets.
problem Classifying chaotic signals accurately and robustly.
method Triad State Space Construction (TSSC) for image encoding, Convolutional Neural Network (ConvNet) for classification.
result TSSC-ConvNet achieves high accuracy and robustness in chaotic signal classification.
Enhances nighttime vehicle detection using style transfer and augmentation.
problem Nighttime object detection challenges due to lack of lighting and glare.
method Day-to-night style transfer and labeling-free augmentation with CARLA synthetic data.
result Significant improvements in nighttime vehicle detection with YOLO11 model.
Enhanced Neural ODEs outperform traditional models in image classification and video prediction.
problem Efficiently modeling time-varying dynamics in neural networks.
method Proposed a novel family of non-autonomous Neural ODEs with time-varying weights.
result Outperformed previous Neural ODE variants in speed and representational capacity.
Intratumor heterogeneity is often manifested by vascular compartments with distinct pharmacokinetics that cannot be resolved directly by in vivo dynamic imaging. We developed tissue-specific compartment modeling (TSCM), an unsupervised computational method of deconvolving dynamic imaging series from heterogeneous tumor…
Efficiently generates high-resolution images with reduced sampling time using LEGO bricks.
problem Efficiently generating high-resolution images with reduced sampling time.
method Introduces LEGO bricks that integrate Local-feature Enrichment and Global-content Orchestration to create a test-time reconfigurable diffusion backbone.
result Significantly reduces sampling time compared to other methods.