Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

75151226301 · Jun 202019922001200920172026
48 results for Object Appearance

Deep learning animates objects from input images and videos.

problem Animating arbitrary objects from input images and videos.
method A deep learning framework with three modules: Keypoint Detector, Dense Motion prediction network, and Motion Transfer Network.
result Our method outperforms state-of-the-art image animation and video generation methods.

ST-STORM separates semantic and appearance features for robust representation learning.

problem Traditional SSL methods fail to capture appearance cues in critical applications.
method Hybrid SSL framework with two latent streams, Content and Style, disentangled through gating mechanisms.
result The Style branch effectively isolates complex appearance phenomena without degrading semantic performance.

End-to-end method learns geometry and appearance for multi-view object detection.

problem Challenges in multi-view object detection, including viewpoint, lighting, and scale variability.
method Jointly learns multi-view geometry and warping for robust cross-view object detection.
result Superior performance compared to baselines on a new street-level panorama data set.

Unified tensor model disentangles object appearance factors.

problem Representing hierarchical intrinsic and extrinsic causal factors of object appearance.
method Compositional hierarchical tensor factorization.
result Interpretable object representation robust to occlusion and reduced training data requirements.

Context-aware ZSL improves object recognition by considering object context.

problem Previous ZSL approaches ignore object context, limiting their effectiveness.
method Proposes a new approach that models the conditional likelihood of objects appearing in specific contexts.
result Contextual information significantly improves ZSL performance and is robust to class imbalance.

Self-guidance controls image generation by extracting properties from diffusion model representations.

problem Generating images from text descriptions is challenging due to the complexity of visual details.
method Self-guidance uses internal representations of diffusion models to control image generation.
result Properties like object shape, location, and appearance can be extracted and used to steer image generation.

Top 8 robotic vision systems tackled lifelong object recognition challenges.

problem Lifelong learning in robotic vision for varied, dynamic environments.
method Design of a dataset with diverse conditions and rules for evaluation.
result Robotic vision systems improved over time with dynamic object appearances.

This paper establishes the existence of observable footprints that reveal the "causal dispositions" of the object categories appearing in collections of images. We achieve this goal in two steps. First, we take a learning approach to observational causal discovery, and build a classifier that achieves state-of-the-art …

2016-05-26abs ↗pdf ↗

We give a mathematical computation of the number of solutions of Apollonius problem, by use of Lie Sphere Geometry. Unlike in higher dimensions, the number of solutions depends only on the topology of the configuration of the 3 objects. It appears that our classification is non redundant, and far simpler than those obt…

2013-07-21abs ↗pdf ↗

We present an explicit construction of the basic bundle gerbes with connection over all connected compact simple Lie groups. These are geometric objects that appear naturally in the Lagrangian approach to the WZW conformal field theories. Our work extends the recent construction of E. Meinrenken \cite{Meinr} restricted…

2003-07-01abs ↗pdf ↗

Develops a new method to improve performance in multi-objective learning problems.

problem Gradient bias in multi-objective learning leading to degraded performance.
method Stochastic Multi-objective gradient Correction (MoCo) method that guarantees convergence without increasing batch size.
result Demonstrates effectiveness of MoCo method in simulations on multi-task learning.

A general theory of quantum spinor structures on quantum spaces is presented, within the conceptual framework of the formalism of quantum principal bundles. Quantum analogs of all basic objects of the classical theory are constructed and analyzed. This includes Laplace and Dirac operators, quantum versions of Clifford …

2000-04-24abs ↗pdf ↗

We consider a multi-objective risk-averse two-stage stochastic programming problem with a multivariate convex risk measure. We suggest a convex vector optimization formulation with set-valued constraints and propose an extended version of Benson's algorithm to solve this problem. Using Lagrangian duality, we develop sc…

2017-11-17abs ↗pdf ↗

We define partial differential (PD in the following), i.e., field theoretic analogues of Hamiltonian systems on abstract symplectic manifolds and study their main properties, namely, PD Hamilton equations, PD Noether theorem, PD Poisson bracket, etc.. Unlike in standard multisymplectic approach to Hamiltonian field the…

2009-03-26abs ↗pdf ↗

Object ranking or "learning to rank" is an important problem in the realm of preference learning. On the basis of training data in the form of a set of rankings of objects represented as feature vectors, the goal is to learn a ranking function that predicts a linear order of any new set of objects. In this paper, we pr…

2017-11-28abs ↗pdf ↗

The abstract discusses parallels between Galois theory and Stone-Weierstrass theorem in various fields.

problem Connecting distinguishing power and expressive power in different fields.
method Elementary theorem connecting distinguishing power and expressive power.
result Foundational principle in linguistics linking distinguishing power and expressive power.

Study finds object detection systems have higher error rates for darker-skinned pedestrians.

problem Predictive bias in object detection systems for pedestrians with different skin tones.
method Annotated a large dataset and compared performance between skin tone groups, investigating factors like time of day and occlusion.
result Predictive bias in object detection systems is not solely due to more difficult scenes for darker-skinned pedestrians.

Learning robot objective functions from human input has become increasingly important, but state-of-the-art techniques assume that the human's desired objective lies within the robot's hypothesis space. When this is not true, even methods that keep track of uncertainty over the objective fail because they reason about …

2018-10-11abs ↗pdf ↗

Non-convex optimization is ubiquitous in machine learning. Majorization-Minimization (MM) is a powerful iterative procedure for optimizing non-convex functions that works by optimizing a sequence of bounds on the function. In MM, the bound at each iteration is required to \emph{touch} the objective function at the opti…

2015-06-25abs ↗pdf ↗

A robust visual tracking system requires an object appearance model that is able to handle occlusion, pose, and illumination variations in the video stream. This can be difficult to accomplish when the model is trained using only a single image. In this paper, we first propose a tracking approach based on affine subspa…

2014-03-03abs ↗pdf ↗

In this paper we introduce distinct approaches to loop braid groups, a generalisation of braid groups, and unify all the definitions that have appeared so far in literature, with a complete proof of the equivalence of these definitions. These groups have in fact been an object of interest in different domains of mathem…

2016-05-08abs ↗pdf ↗

Latent feature models are attractive for image modeling, since images generally contain multiple objects. However, many latent feature models ignore that objects can appear at different locations or require pre-segmentation of images. While the transformed Indian buffet process (tIBP) provides a method for modeling tra…

2012-06-27abs ↗pdf ↗

Paper proposes redundancy-free features for zero-shot object recognition.

problem Redundant visual features degrade zero-shot object recognition.
method Project original features into a new, statistically independent space.
result RFF-GZSL achieves competitive results on benchmark datasets.

Motivated by problems in search and detection we present a solution to a Combinatorial Multi-Armed Bandit (CMAB) problem with both heavy-tailed reward distributions and a new class of feedback, filtered semibandit feedback. In a CMAB problem an agent pulls a combination of arms from a set {1,...,k}\{1,...,k\} in each round, g…

2017-05-26abs ↗pdf ↗

A novel tracking method for dense honeybee colonies using pixel personality.

problem Tracking large numbers of densely-arranged, interacting objects in a 2D environment.
method Segmentation-based object detection followed by adaptive object recognition through visual appearance.
result Reconstructed ~46% of trajectories in 5 minutes and 71% of tracks for at least 2 minutes.

The coherent states are viewed as a powerful tool in differential geometry. It is shown that some objects in differential geometry can be expressed using quantities which appear in the construction of the coherent states. The following subjects are discussed via the coherent states: the geodesics, the conjugate locus a…

1998-07-14abs ↗pdf ↗

New examples of Z/2 harmonic 1-forms and their branching sets are explored.

problem Exploring the properties and examples of Z/2 harmonic 1-forms and their branching sets.
method Elementary constructions and families of Z2\Z_2 harmonic 1-forms.
result The branching set ΣΣ of a Z2\Z_2 harmonic 1-form can exhibit various features including non-trivial links, multiple covers, and immersed structures.

We use reinforcement learning (RL) to learn dexterous in-hand manipulation policies which can perform vision-based object reorientation on a physical Shadow Dexterous Hand. The training is performed in a simulated environment in which we randomize many of the physical properties of the system like friction coefficients…

2018-08-01abs ↗pdf ↗

Reducing barriers to entry in large-scale ML markets, study shows multi-objective learning can lower data requirements.

problem Barriers to entry in emerging markets for large-scale machine learning models.
method Defined a multi-objective high-dimensional regression framework to study reputational damage and data requirements.
result The number of data points needed for a new company to enter the market can be significantly smaller than the incumbent company's dataset size.

Frolicher and Nijenhuis recognized well in the middle of the previous century that the Lie bracket and its Jacobi identity could and should exist beyond Lie algebras. Nevertheless the conceptual meaning of their discovery has been obscured by the messy techniques they exploited. The principal objective in this paper is…

2009-04-07abs ↗pdf ↗

InfoNCE objective is equivalent to ELBO in RPM, linking to self-supervised learning.

problem Improving self-supervised learning methods by connecting them to variational inference.
method Recognizing RPM and showing InfoNCE as a simplified lower bound on MI, equal to ELBO in infinite sample limit.
result The actual InfoNCE objective is equal to the ELBO (up to a constant) in the infinite sample limit.