Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,291 papers · 148 categories

Trend · papers per month

140279419558 · Jun 202019922001200920182026
48 results for visual effects

A faster method for visualization recommendations on large datasets.

problem Infeasibility of state-of-the-art vis-rec models on large datasets due to high computational time.
method Reinforcement-learning (RL) framework that identifies optimal statistics within a time budget.
result Significantly reduces time-to-visualize with minimal error compared to baseline approaches.

Proposes a deep Auto-Encoder-like framework for visual-tactile fusion object clustering.

problem Combining visual and tactile information for better object clustering.
method Deep Auto-Encoder-like Non-negative Matrix Factorization framework, graph regularizer, modality-level consensus regularizer, alternating minimization strategy.
result Improves object clustering performance by leveraging both visual and tactile modalities.

This paper explores neural network loss landscapes and their effects on generalization.

problem Understanding the structure of neural network loss functions and their impact on generalization.
method Simple filter normalization and various visualization methods to explore loss landscape structure and network architecture effects.
result Visualizations reveal how network architecture and training parameters affect loss landscape curvature and minimizers.

AVDCNN combines audio and visual data for better speech enhancement.

problem Improving speech quality by reducing noise in audio signals.
method Proposes an AVDCNN model that integrates audio and visual streams into a unified deep CNN network for end-to-end training.
result AVDCNN outperforms audio-only and conventional SE methods in enhancing speech quality.

CDDN tackles visual relationship detection with context-dependent diffusion networks.

problem Combustion of combinatorial explosion in relation triplets detection.
method CDDN framework using semantic and visual scene graphs for adaptive information aggregation.
result CDDN achieves state-of-the-art performance on visual relationship detection datasets.

VTAB benchmarks diverse visual tasks to assess representation learning effectiveness.

problem Lack of a unified evaluation for general visual representations.
method Developed VTAB, a benchmark for diverse visual tasks, and evaluated many representation learning algorithms.
result VTAB revealed insights into the effectiveness of various representation learning methods.

Novel model for decoding visual stimuli in human brains.

problem Challenges in MVP techniques, including noise and sparsity, and the cost of brain studies.
method Automatic detection of active regions, new Gaussian smoothing method, combining fMRI data sets.
result Superior performance compared to state-of-the-art methods.

Consensus dimension reduction combines multiple visualizations to identify shared patterns.

problem Conflicting visualizations from different dimension reduction methods.
method Multi-view learning to identify stable patterns across multiple views.
result Consensus visualization effectively identifies shared low-dimensional data structure.

Proposes a spectral method to assess and combine multiple data visualizations.

problem Evaluating and combining the strengths of different data visualization algorithms.
method Spectral method for assessing and combining multiple visualizations.
result Proposes a visualization eigenscore to quantify relative performance and a consensus visualization.

This paper discusses the role of risk communication in macroprudential oversight and of visualization in risk communication. Beyond the soar in data availability and precision, the transition from firm-centric to system-wide supervision imposes vast data needs. Moreover, except for internal communication as in any orga…

2014-04-17abs ↗pdf ↗

Self-supervised learning of visual semantics in image games.

problem Learning visual semantics in referential emergent language games.
method Investigating the impact of feature extractor weights and tasks on visual semantics, using various image augmentations and additional tasks.
result Communication systems can learn visual semantics in a self-supervised manner by playing the right types of games.

DarkSight visualizes deep classifiers more effectively than t-SNE.

problem Interpreting black box classifiers like deep networks.
method DarkSight embeds data points into a low-dimensional space to compress deep classifiers, using dark knowledge for a new confidence measure.
result DarkSight visualizations are more informative and yield a new confidence measure.

Neural model predicts object states and physical parameters from visual observations.

problem Computational models struggle with physical reasoning and adapting to new environments.
method Visual prior predicts particle-based system from visual observations; inference module refines estimates subject to dynamics constraints.
result Model can infer physical properties within a few observations and adapt to unseen scenarios.

AdvReg improves VQA models but introduces instability and bias issues.

problem VQA models over-rely on linguistic biases, ignoring visual context.
method Adversarial regularization to encourage bias-free question representations.
result AdvReg yields side-effects like unstable gradients and reduced performance on in-domain examples.

RuleMatrix visualizes machine learning models for non-experts.

problem Making machine learning models transparent and interpretable for non-expert users.
method Extracts rule-based knowledge from model behavior and presents it in an interactive matrix visualization.
result RuleMatrix helps non-expert users understand and validate machine learning models.

FiLM layers improve visual reasoning tasks by modulating features.

problem Visual reasoning tasks that require multi-step, high-level processes.
method General-purpose FiLM layers that apply feature-wise linear transformations based on conditioning information.
result FiLM layers reduce error by half on the CLEVR benchmark and improve feature coherence.

FeatureEnVi aids in feature engineering with visual analytics.

problem Insufficient support for feature engineering in visual analytics tools.
method Stepwise selection and semi-automatic extraction approaches.
result Extracts heavily engineered features evaluated by multiple metrics.

Paper proposes a new black-box attack approach to minimize visual distortion.

problem Constructing adversarial examples that minimize visual distortion in a black-box threat model.
method Learning the noise distribution of adversarial examples to approximate the gradient of a non-differentiable loss function.
result The proposed attack results in much lower visual distortion compared to state-of-the-art black-box attacks.

This work tackles long-term visual planning by goal-conditioned hierarchical predictors.

problem Current learning approaches fail on long-horizon tasks due to lack of goal information and coarse-to-fine planning.
method Formulate goal-conditioned predictors (GCPs) and hierarchical models to predict trajectories between observations.
result GCPs enable effective long-term planning with much longer horizons than before.

Previously, we proposed a physically-inspired method to construct data points into an effective in-tree (IT) structure, in which the underlying cluster structure in the dataset is well revealed. Although there are some edges in the IT structure requiring to be removed, such undesired edges are generally distinguishable…

2015-07-29abs ↗pdf ↗

New methods for assessing and visualizing feature groups in machine learning models.

problem Lack of methods for interpreting feature groups in machine learning models.
method Permutation-based, refitting, and Shapley-based techniques for grouped feature importance. Introduced a sequential procedure for identifying stable feature combinations. Developed a combined features effect plot.
result Effective methods for assessing and visualizing the importance and effect of feature groups in machine learning models.

Enhanced PCA method highlights essential features of clusters in high-dimensional data.

problem Interpreting clusters in dimensionality reduction results is challenging.
method Contrastive Principal Component Analysis (cPCA) for identifying essential features.
result ccPCA method effectively highlights essential features of clusters in high-dimensional data.

Visual system compares and evaluates machine learning models for clinical data predictions.

problem Challenges in comparing and evaluating different machine learning models for medical predictions.
method Developed a visual analytics system to compare and evaluate multiple models' prediction criteria and consistency.
result Demonstrated the effectiveness of the visual analytics system in assisting clinicians and researchers.

Paper proposes a deep learning architecture for generating long stories from images.

problem Maintaining context in long event sequences for visual storytelling.
method Hierarchical deep learning architecture with encoder-decoder networks and natural language descriptions.
result Our method outperforms state-of-the-art techniques on automatic evaluation metrics.

This work examines how non-identical data distributions affect Federated Learning performance.

problem The impact of non-identical data distributions on Federated Learning performance.
method Synthesized datasets with varying degrees of data distribution similarity, evaluated Federated Averaging algorithm performance, proposed server momentum mitigation.
result Performance of Federated Learning degrades as data distributions differ more, and a mitigation strategy improves accuracy.