Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

2356 · Apr 202019922001200920172026
48 results for open-domain dialog

Study finds neural dialog models struggle with conversational tasks.

problem Insufficient understanding of dialog by neural models.
method Analysis of internal representations and evaluation of model performance.
result Neural dialog models lack key conversational skills like answering questions and inferring contradiction.

This study compares hierarchical and non-hierarchical models for open-domain multi-turn dialog generation.

problem Which kind of models (hierarchical or non-hierarchical) is better for open-domain multi-turn dialog generation?
method Systematically compared nearly all representative hierarchical and non-hierarchical models over the same experimental settings.
result Nearly all hierarchical models are worse than non-hierarchical models in open-domain multi-turn dialog generation, except for HRAN.

Open-domain dialog generation is a challenging problem; maximum likelihood training can lead to repetitive outputs, models have difficulty tracking long-term conversational goals, and training on standard movie or online datasets may lead to the generation of inappropriate, biased, or offensive text. Reinforcement Lear…

2019-09-17abs ↗pdf ↗

ClovaCall introduces a new Korean call speech corpus for contact centers.

problem Lack of large-scale call-based speech corpora for Korean dialog scenarios.
method Development of a new large-scale Korean call-based speech corpus (ClovaCall) in a restaurant reservation domain.
result Validation of the dataset with ASR models shows its effectiveness.

In an end-to-end dialog system, the aim of dialog state tracking is to accurately estimate a compact representation of the current dialog status from a sequence of noisy observations produced by the speech recognition and the natural language understanding modules. This paper introduces a novel method of dialog state t…

2016-06-13abs ↗pdf ↗

Representing a dialog policy as a recurrent neural network (RNN) is attractive because it handles partial observability, infers a latent representation of state, and can be optimized with supervised learning (SL) or reinforcement learning (RL). For RL, a policy gradient approach is natural, but is sample inefficient. I…

2016-12-18abs ↗pdf ↗

Prior work on training generative Visual Dialog models with reinforcement learning(Das et al.) has explored a Qbot-Abot image-guessing game and shown that this 'self-talk' approach can lead to improved performance at the downstream dialog-conditioned image-guessing task. However, this improvement saturates and starts d…

2019-09-23abs ↗pdf ↗

Task oriented dialog agents provide a natural language interface for users to complete their goal. Dialog State Tracking (DST), which is often a core component of these systems, tracks the system's understanding of the user's goal throughout the conversation. To enable accurate multi-domain DST, the model needs to enco…

2020-02-07abs ↗pdf ↗

New domains of discontinuity found for Anosov representations.

problem Understanding Anosov representations acting on homogeneous spaces.
method Constructing open domains of discontinuity for Anosov representations acting on specific homogeneous spaces.
result Describes the largest possible open domains of discontinuity for Zariski dense Anosov representations.

Machine reading using differentiable reasoning models has recently shown remarkable progress. In this context, End-to-End trainable Memory Networks, MemN2N, have demonstrated promising performance on simple natural language based reasoning tasks such as factual reasoning and basic deduction. However, other tasks, namel…

2016-10-13abs ↗pdf ↗

The Knowledge Base (KB) used for real-world applications, such as booking a movie or restaurant reservation, keeps changing over time. End-to-end neural networks trained for these task-oriented dialogs are expected to be immune to any changes in the KB. However, existing approaches breakdown when asked to handle such c…

2018-05-03abs ↗pdf ↗

We present Meena, a multi-turn open-domain chatbot trained end-to-end on data mined and filtered from public domain social media conversations. This 2.6B parameter neural network is simply trained to minimize perplexity of the next token. We also propose a human evaluation metric called Sensibleness and Specificity Ave…

2020-01-27abs ↗pdf ↗

Paper develops a model for verifying facts in tables without pre-retrieved evidence.

problem Verification of factual claims in structured data, especially in open-domain settings.
method Joint reranking-and-verification model that fuses evidence documents.
result Model achieves comparable performance to closed-domain state-of-the-art on TabFact dataset.

An increasing number of datasets contain multiple views, such as video, sound and automatic captions. A basic challenge in representation learning is how to leverage multiple views to learn better representations. This is further complicated by the existence of a latent alignment between views, such as between speech a…

2018-11-21abs ↗pdf ↗

The reduction of biharmonic maps equation in terms of the Maurer-Cartan form for all smooth map of any compact Riemannian manifolds into a compact Lie group with bi-invariant Riemannian metric is obtained. By this formula, all the biharmonic curves into a compact Lie group and all biharmonic maps from a 2-dimensional o…

2009-10-05abs ↗pdf ↗

In this paper, the description of biharmonic map equation in terms of the Maurer-Cartan form for all smooth map of a compact Riemannian manifold into a Riemannian symmetric space (G/K,h)(G/K,h) induced from the bi-invariant Riemannian metric hh on GG is obtained. By this formula, all biharmonic curves into symmetric space…

2011-01-17abs ↗pdf ↗

In this paper we analyze the problem of the geodesic connectedness of subsets of Riemannian manifolds. By using variational methods, the geodesic connectedness of open domains (whose boundaries can be not differentiable and not convex) of a smooth Riemannian manifold is proved. In some cases also the convexity of the d…

2000-04-12abs ↗pdf ↗

VALAN is a lightweight and scalable software framework for deep reinforcement learning based on the SEED RL architecture. The framework facilitates the development and evaluation of embodied agents for solving grounded language understanding tasks, such as Vision-and-Language Navigation and Vision-and-Dialog Navigation…

2019-12-06abs ↗pdf ↗

A family of probability distributions parametrized by an open domain ΛΛ in RnR^n defines the Fisher information matrix on this domain which is positive semi-definite. In information geometry the standard assumption has been that the Fisher information matrix tensor is positive definite defining in this way a Riemannia…

2015-03-29abs ↗pdf ↗

We consider an open domain with a compact boundary in an Euclidean space and a Schroedinger operator with magnetic field on this domain. We give sufficient conditions on the rate of growth of the magnetic field near the boundary which guarantees essential self-adjointness of this operator. From the physical point of vi…

2009-03-04abs ↗pdf ↗

In this paper we find, for any arbitrary finite topological type, a compact Riemann surface M,\mathcal{M}, an open domain MMM\subset\mathcal{M} with the fixed topological type, and a conformal complete minimal immersion X:MR3X:M\to\R^3 which can be extended to a continuous map X:MˉR3,X:\bar{M}\to\R^3, such that $X_{|\partial M…

2007-11-15abs ↗pdf ↗

Study shows distance to boundary is always attained on varifolds with bounded curvature.

problem Understanding varifolds with bounded mean curvature in Riemannian manifolds.
method Proves a barrier principle at infinity using sharp maximum principles.
result Distance to boundary is always attained on varifolds with bounded curvature.

Learning goal-oriented dialogues by means of deep reinforcement learning has recently become a popular research topic. However, commonly used policy-based dialogue agents often end up focusing on simple utterances and suboptimal policies. To mitigate this problem, we propose a class of novel temperature-based extension…

2018-07-02abs ↗pdf ↗