Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

5.0%10.0%15.0%20.0% · Aug 199419922001200920182026
48 results for motion cues

We present a model for the joint estimation of disparity and motion. The model is based on learning about the interrelations between images from multiple cameras, multiple frames in a video, or the combination of both. We show that learning depth and motion cues, as well as their combinations, from data is possible wit…

2013-12-12abs ↗pdf ↗

This paper presents a novel yet intuitive approach to unsupervised feature learning. Inspired by the human visual system, we explore whether low-level motion-based grouping cues can be used to learn an effective visual representation. Specifically, we use unsupervised motion-based segmentation on videos to obtain segme…

2016-12-19abs ↗pdf ↗

Proposes MSFCN and RFCN for video semantic segmentation improving accuracy by 9-15%.

problem Lack of temporal information in semantic segmentation algorithms for videos.
method Integrates motion cues and temporal consistency using recurrent and multi-stream FCN architectures.
result MSFCN-3 achieves significant accuracy improvements in various datasets.

Study reveals DNNs prefer easy-to-learn cues over essential ones in image recognition.

problem DNNs learn easy-to-learn features that aren't essential to the task.
method WCST-ML training setup with shortcut cues on synthetic and face datasets.
result DNNs converge to solutions focusing on preferred cues, leading to flat minima.

We propose a recurrent extension of the Ladder networks whose structure is motivated by the inference required in hierarchical latent variable models. We demonstrate that the recurrent Ladder is able to handle a wide variety of complex learning tasks that benefit from iterative inference and temporal modeling. The arch…

2017-07-28abs ↗pdf ↗

Automatically differentiable estimation for BLP model reduces bias in demand estimation.

problem Estimating the BLP model with reduced bias and improved performance.
method Phrasing BLP as an automatically differentiable moment function, using CUE for estimation, and incorporating MCMC credible intervals.
result CUE estimation shows lower bias but higher MAE compared to 2S-GMM, with MCMC providing closest empirical coverage.

Deep learning detects protective movement behavior in chronic pain patients.

problem Detecting protective movement behavior in chronic pain patients for intervention.
method End-to-end deep learning architecture named BodyAttentionNet (BANet) that learns temporal and bodily parts.
result Statistically significant improvements in detecting protective behavior using attention mechanisms.

Paper introduces a method for generating interlocutor-aware facial gestures in dyadic settings.

problem Generating appropriate non-verbal behavior for conversational agents in dyadic settings.
method Probabilistic method using multi-modal cues from the interlocutor to synthesize facial gestures.
result The model successfully leverages multi-modal input from the interlocutor to generate more appropriate behavior.

Visually predicting the stability of block towers is a popular task in the domain of intuitive physics. While previous work focusses on prediction accuracy, a one-dimensional performance measure, we provide a broader analysis of the learned physical understanding of the final model and how the learning process can be g…

2018-06-14abs ↗pdf ↗

Multimodal analysis assesses job interview performance and provides feedback.

problem Assessing candidate performance in interviews for professional roles.
method Multimodal analytical framework using video, audio, and text data.
result The proposed methodology achieved promising results in predicting behavioral cues.

AR-GANs learn depth and DoF from unlabeled images using aperture rendering and focus cues.

problem Learning depth and DoF from unlabeled natural images with diverse viewpoints and shapes.
method Aperture rendering and focus cues to learn depth and DoF from unlabeled images.
result AR-GANs effectively learn depth and DoF from various datasets, including flower, bird, and face images.

This work improves medical image segmentation with limited annotations using contrastive learning.

problem Lack of labeled data for medical image segmentation.
method Contrastive learning framework for semi-supervised segmentation with domain-specific and problem-specific cues.
result Significant improvements in segmentation performance compared to other methods.

Framework improves CT image segmentation robustness with domain-specific cues.

problem Challenges in CT image segmentation by deep learning models.
method Combines domain-specific preprocessing and augmentation with CNN architectures.
result Framework stabilizes prediction performance on varying CT volumes.

Co-PLNet combines point and line predictions to improve wireframe parsing accuracy and efficiency.

problem Separate line and point predictions lead to inconsistent wireframes.
method Co-PLNet uses a Point-Line Prompt Encoder to convert early point detections into spatial prompts, which guide line refinement.
result Co-PLNet achieves better accuracy and robustness in wireframe parsing compared to existing methods.

Paper addresses shortcomings in pointer generator networks for summarization.

problem Extractive summaries and factual inaccuracies in generated text.
method Appends traditional linguistic information to teach networks on text structure.
result Feasibility and potential of additional cues for improved generation.

A new algorithm efficiently partitions images without seeds or thresholds.

problem Efficiently partitioning images without explicit seeds or thresholds.
method Mutex Watershed algorithm that incorporates both attractive and repulsive cues.
result Determines optimal segments without seeds or thresholds, solving NP-hard problem.

A movie multilayer network model captures narration from script, subtitles, and content.

problem Discovering content and stories in movies using network models.
method Developed a multilayer network model using visual and textual semantic cues.
result Demonstrated the effectiveness of the model on the Star Wars saga.

Paper proposes a new method for learning compact representations of sequential data.

problem Learning compact representations of sequential data capturing spatio-temporal cues.
method Contrastive representation learning via adversarial optimal transport on the Grassmann manifold.
result Empirical results show competitive performance in human action recognition.

New corpus improves coreference resolution by removing gender and number cues.

problem Challenges in resolving ambiguous pronoun references.
method Developed a new annotated corpus, introduced antecedent switching technique.
result Models perform poorly on ambiguous pronoun references, but antecedent switching improves performance.

New method separates audio sources without needing known decompositions.

problem Difficulty in training source separation models on real-world mixtures.
method Generates estimated decompositions from stereo mixtures and trains a deep learning model.
result Trained model can separate single-channel audio sources effectively.

Introduces Motion Programs for better video analysis of human motion.

problem Current video analysis focuses on raw pixels or keypoints, missing higher-level motion primitives.
method Introduces Motion Programs as a neuro-symbolic representation of motions as a composition of high-level primitives.
result Motion Programs accurately describe diverse human motions and improve downstream tasks.

Unified framework for human motion generation on Riemannian manifolds.

problem Learning valid human motion in Euclidean spaces.
method Riemannian Motion Generation (RMG) on product manifolds, Riemannian flow matching.
result Achieves state-of-the-art FID (0.043) on HumanML3D and surpasses strong baselines on MotionMillion.

Study on determinants of unitary Brownian motion and their asymptotic laws.

problem Understanding determinants of unitary Brownian motion and their behavior over time.
method Using Stiefel fibration and skew-product decomposition of the Stiefel Brownian motion.
result Prove asymptotic laws for determinants of block entries of unitary Brownian motion.

New framework predicts diverse, contextually plausible 3D human motions.

problem Predicting multiple plausible future 3D poses given observed poses.
method Developed a new variational framework that conditions latent variable on past observation to encourage relevant information.
result Our approach generates motions of higher quality and preserves contextual information.

Neural network predicts vessel motions with high accuracy.

problem Real-time prediction of heave and surge motions for improved performance and safety.
method Developed an LSTM-based machine learning model trained on measured waves and motion data.
result The model predicts vessel motions up to 46.5 seconds into the future with an average accuracy of 90%.

Let EE be a closed set in the Riemann sphere C^\widehat{\mathbb{C}}. We consider a holomorphic motion φφ of EE over a complex manifold MM, that is, a holomorphic family of injections on EE parametrized by MM. It is known that if MM is the unit disk ΔΔ in the complex plane, then any holomorphic motion of EE ove…

2017-09-22abs ↗pdf ↗

Study cohomological equation for robotic screw motions on SE(3).

problem Understanding obstruction phenomena in robotic rigid-body motion.
method Combining Fourier analysis and Peter-Weyl theory, reduce to finite-dimensional linear transport systems.
result Explicit screw motion illustrates resonance conditions and finite-dimensional obstructions.

New approach for obstacle avoidance in robotics using learned representations.

problem Challenges in sensor-based motion planning for new and dynamic environments.
method Proposes a new obstacle representation using PointNet architecture trained jointly with policies for obstacle avoidance.
result Significant improvements in accuracy and efficiency compared to state of the art.

The paper presents a method to reduce arm motion complexity for prosthetics and robotics.

problem Reducing the complexity of human arm motions for robotic and prosthetic control.
method Data-driven techniques including DTW, DBA, Ward's distance, batch-DTW, and fPCA.
result Representative motion clusters and averages for different arm DOF levels.