Study uses AI to refine loan assessments, improving credit default predictions.
problem Improving credit default prediction accuracy using AI-refined text.
method Comparative analysis of human-written and AI-refined loan assessments using deep learning techniques.
result AI-refined texts significantly enhance credit default predictions, especially when combined with structured data.
Computes Steenrod squares on Khovanov homology for knots up to 11 crossings.
problem Computing Steenrod squares on Khovanov homology.
method Spatial refinements of even and odd Khovanov homology, computation of Sq^2.
result Steenrod squares Sq^2 on Khovanov homology spaces are determined for knots up to 11 crossings.
Improved text-to-image alignment using iterative VQA feedback.
problem Misalignment between text prompts and generated images, especially for complex inputs.
method Decompose complex prompts into assertions, evaluate each using VQA, combine scores iteratively.
result Significantly higher correlation with human ratings compared to CLIP, BLIP scores.
Synthesizing high-quality images from text descriptions is a challenging problem in computer vision and has many practical applications. Samples generated by existing text-to-image approaches can roughly reflect the meaning of the given descriptions, but they fail to contain necessary details and vivid object parts. In…
In the theory of finite order knot invariants, the universal sl2 weight system maps the chord diagrams to polynomials in a single variable with integer coefficients. In this paper, we define a family of polynomials that generalize the Kreweras triangle (known to refine the normalized median Genocchi numbers),…
Deep neural networks are gaining increasing popularity for the classic text classification task, due to their strong expressive power and less requirement for feature engineering. Despite such attractiveness, neural text classification models suffer from the lack of training data in many real-world applications. Althou…
Improves deep generative model samples via gradient flow.
problem Inconsistent generation quality across samples.
method Discriminator Gradient Flow (DGflow) using entropy-regularized f-divergences.
result Significant improvement in generated sample quality.
Improved algorithm for misspecified MLMDPs with bounded regret and space/time complexities.
problem Misspecified linear Markov decision processes.
method Proposes an algorithm with three desirable properties: bounded regret, bounded space/time complexities, and no need for misspecification input.
result Regret scales as Kmax{εextmis,εexttol}, improving existing bounds. We study Kuperberg invariants for sutured manifolds in the case of a semidirect product of an involutory Hopf superalgebra H with its automorphism group Aut(H). These are topological invariants of balanced sutured 3-manifolds endowed with a homomorphism of the fundamental group into Aut(H) and possi…
IterefinE combines KG refinement with embeddings to improve KG quality.
problem Noisy Knowledge Graphs lead to poor performance in downstream tasks.
method Iterative KG refinement using embeddings and inference rules.
result Improved KG refinement leading to higher F1 scores.
This paper improves retrieval for LLMs in financial document Q&A.
problem Suboptimal text chunk retrieval by RAG causes inaccuracies in LLM responses.
method Sophisticated chunking techniques, query expansion, metadata annotations, re-ranking algorithms, and embedding fine-tuning.
result Enhanced retrieval quality improves LLM performance and reliability.
Amortized variational inference (AVI) replaces instance-specific local inference with a global inference network. While AVI has enabled efficient training of deep generative models such as variational autoencoders (VAE), recent empirical work suggests that inference networks can produce suboptimal variational parameter…
Simple bounds show most cross-sectional predictability findings are likely true.
problem Determining the validity of cross-sectional return predictability findings.
method Developed simple and intuitive bounds on the false discovery rate (FDR).
result Bounds show the FDR is small, indicating most findings are likely true.
New operations match Steenrod squares on Khovanov homology.
problem Matching Steenrod squares with operations on Khovanov homology.
method Applied cup-i products to Khovanov functor and proved agreement with Steenrod squares.
result Lipshitz-Sarkar's Sq^2 agrees with Morán's sq^2.
While neural sequence generation models achieve initial success for many NLP applications, the canonical decoding procedure with left-to-right generation order (i.e., autoregressive) in one-pass can not reflect the true nature of human revising a sentence to obtain a refined result. In this work, we propose XL-Editor, …
Improved financial network predictability using LLM for edge filtering.
problem Spurious edges in financial networks from textual similarity.
method Two-stage framework: sparse candidate graph + LLM edge classification.
result LLM-based edge filtering improves Sharpe ratio and reduces drawdown.
In this text, we establish the risk model based on AR(1) series and propose the basic model which has a dependent structure under intensity of claim number. Considering some properties of the risk model, we take advantage of newton iteration method to figure out the adjustment coefficient and estimate the exponential u…
This paper uses Heegaard Floer theory to study pseudo-Anosov flows and their periodic points.
problem Understanding the differential of Heegaard Floer chain complexes associated with pseudo-Anosov flows.
method Introduces a refined grading to analyze the homology of subcomplexes representing irreducible multi-orbits.
result The homology of these subcomplexes is 1-dimensional, providing insights into periodic points of pseudo-Anosov flows.
This study mainly investigates two common decoding problems in neural keyphrase generation: sequence length bias and beam diversity. To tackle the problems, we introduce a beam search decoding strategy based on word-level and ngram-level reward function to constrain and refine Seq2Seq inference at test time. Results sh…
VEC-SBM detects communities using side information like texts and images.
problem Community detection in social networks with side information.
method Proposes a novel algorithm based on iterative refinement techniques.
result Optimally recovers latent communities with side information.
Scarcity of labeled data is one of the most frequent problems faced in machine learning. This is particularly true in relation extraction in text mining, where large corpora of texts exists in many application domains, while labeling of text data requires an expert to invest much time to read the documents. Overall, st…
Efficiently distills pretrained text-to-image models without real data, improving FID and CLIP scores.
problem Slow iterative refinement process of diffusion-based text-to-image models.
method Guided Score identity Distillation with Long and Short Classifier-Free Guidance.
result Achieves state-of-the-art FID performance with competitive CLIP score.
This paper explores how NLP enhances insurance data analysis.
problem Traditional insurance data limitations and need for alternative data.
method Application of NLP techniques to transform and analyze unstructured text data.
result NLP techniques improve insurance data analysis and risk assessment.
DLM-One speeds up language generation by 500x with continuous models.
problem Efficiently generating text sequences in natural language processing.
method Score-distillation of continuous diffusion language models.
result Achieves up to 500x speedup in inference time with competitive performance.
One of the major challenges in machine learning nowadays is to provide predictions with not only high accuracy but also user-friendly explanations. Although in recent years we have witnessed increasingly popular use of deep neural networks for sequence modeling, it is still challenging to explain the rationales behind …
WSD uses a deterministic model to accelerate diffusion-based sampling.
problem Slow refinement process in diffusion models.
method Warm-start model that predicts an informed prior conditioned on input context.
result Significantly reduces the number of diffusion steps required for realistic samples.
AI advances impact asset management, offering new decision-making capabilities.
problem AI's impact on asset management and potential disruption.
method Analyzing how AI capabilities (reading, reasoning, decision-making) can be applied to asset management.
result AI can revolutionize asset management, but risks vary by fund type.
A bandit algorithm reduces regret in noisy, communication-constrained feedback.
problem Distributed stochastic multi-armed bandit with noisy, communication-constrained feedback.
method Proposes a multi-phase bandit algorithm, UE-UCB++, that matches an information-theoretic lower bound.
result Matches an information-theoretic lower bound of Ω(√(KT/σ²)) on the minimax regret.
ED-NeRF efficiently edits 3D scenes using latent space NeRF and improved loss functions.
problem Slow training speeds and inadequate editing loss functions in existing NeRF editing techniques.
method Embedding real-world scenes into latent space of LDM, using a unique refinement layer and an improved loss function.
result ED-NeRF achieves faster editing speed and improved output quality compared to state-of-the-art models.
We give improved algorithms for the ℓp-regression problem, minx∥x∥p such that Ax=b, for all p∈(1,2)∪(2,∞). Our algorithms obtain a high accuracy solution in O~p(m2p+∣p−2∣∣p−2∣)≤O~p(m31) iterations, where each iteration requires s…
Survey examines challenges of ML in avionic systems certification.
problem Challenges in current certification standards for ML in avionic systems.
method Literature review focusing on robustness and explainability of ML results.
result Current certification standards do not support ML in avionic systems.
Enhances thematic investing with stock embeddings from textual data.
problem Challenges in constructing thematic portfolios due to overlapping sector boundaries and evolving market dynamics.
method Introduces THEME, a framework that fine-tunes embeddings using hierarchical contrastive learning, aligning themes and stocks using their hierarchical relationship and incorporating stock returns.
result Theme-aligned portfolios demonstrate compelling performance, significantly outperforming large language models in thematic asset retrieval.
Capacity-Constrained Online Convex Optimization with Delayed Feedback
problem Online learning with delayed feedback under a hard capacity constraint
method Reduction to a delayed and weighted OCO problem using a scheduler
result First regret guarantees for capacity-constrained OCO under convex and strongly convex losses
Self-distillation improves constrained language generation by aligning models with target distributions.
problem Sparse and uninformative reward signals in constrained generation settings.
method Iteratively refining the base model through self-distillation, incorporating learned twist functions and proposals.
result Substantial gains in generation quality through improved model alignment with target distributions.
New inequality for refined knot invariants in a specific space.
problem General adjunction inequality for refined s-invariants does not hold. method Introduced an adjunction inequality for a specific spatial refinement in kCP2. result An adjunction inequality holds for the s-version of the Sq1-refinement in kCP2. We study refined topological string theory in the presence of orientifolds by counting second-quantized BPS states in M-theory. This leads us to propose a new integrality condition for both refined and unrefined topological strings when orientifolds are present. We define the SO(2N) refined Chern-Simons theory which co…
Hybrid approach combines topic and graph embeddings for legal document clustering.
problem Challenges in classifying legal texts due to domain-specific language and limited labeled data.
method Combines unsupervised topic and graph embeddings with a supervised model.
result Improves clustering quality over text-only or graph-only embeddings.
Refines neural network predictions using background knowledge for improved accuracy.
problem Compensate for lack of labeled data in neural networks.
method Introduces differentiable refinement functions and Iterative Local Refinement (ILR) algorithm to refine predictions efficiently and accurately.
result ILR finds competitive results in MNIST addition task and refines predictions on complex SAT formulas.
New research shows label refinement and weak training have limitations for aligning LLMs.
problem Limitations of refinement methods for aligning large language models.
method Analyzed probabilistic assumptions and alternative approaches to label refinement and weak training.
result Label refinement and weak training suffer from irreducible error, leaving a performance gap.
Avaya Conversational Intelligence(ACI) is an end-to-end, cloud-based solution for real-time Spoken Language Understanding for call centers. It combines large vocabulary, real-time speech recognition, transcript refinement, and entity and intent recognition in order to convert live audio into a rich, actionable stream o…
We refine Khovanov homology in the presence of an involution on the link. This refinement takes the form of a triply-graded theory, arising from a pair of filtrations. We focus primarily on strongly invertible knots and show, for instance, that this refinement is able to detect mutation.
In a previous paper we constructed a spectrum-level refinement of Khovanov homology. This refinement induces stable cohomology operations on Khovanov homology. In this paper we show that these cohomology operations commute with cobordism maps on Khovanov homology. As a consequence we obtain a refinement of Rasmussen's …
This paper explores the preference-based top-K rank aggregation problem. Suppose that a collection of items is repeatedly compared in pairs, and one wishes to recover a consistent ordering that emphasizes the top-K ranked items, based on partially revealed preferences. We focus on the Bradley-Terry-Luce (BTL) model…
New algorithm tackles multi-agent reinforcement learning with optimal convergence rate.
problem Multi-agent reinforcement learning with large state spaces and linear function approximations.
method Refined AVLPR framework with data-dependent pessimistic estimation and action-dependent bonuses.
result First algorithm with optimal O(T−1/2) convergence rate and no poly(Amax) dependency. Refined 3D index uses surgery and gradings to distinguish 3-manifolds.
problem Distinguishing 3-manifolds and gauge theories phases.
method Dehn surgery presentation, ideal triangulation, and enhanced flavor symmetries.
result Invariance of refined index under various transformations.
Quantification is a supervised learning task that consists in predicting, given a set of classes C and a set D of unlabelled items, the prevalence (or relative frequency) p(c|D) of each class c in C. Quantification can in principle be solved by classifying all the unlabelled items and counting how many of them have bee…
Deep adaptive sampling improves surrogate modeling for complex systems.
problem Statistical errors in random sampling for high-dimensional problems.
method DAS^2 method, using deep generative models to refine training sets.
result Reduces statistical errors in approximating solutions for low-regularity problems.
This study improves sentence embeddings from BERT models.
problem Capturing the underlying meaning of sentences using BERT models.
method Comprehensive review and testing of various sentence embedding extraction and refinement methods.
result Representation-shaping techniques significantly improve sentence embeddings from BERT-based and simple baseline models.