Hand-held system translates foreign menus for diet management.
problem Translation ambiguities and context-specific information for diet management.
method Portable multimedia device, machine translation, context-specific corpora, pre-processing steps, multimedia information.
result Higher accuracy and instant translations compared to Google Translate.
Emotion classification improved using brain signals from tactile enhanced multimedia.
problem Classifying viewer emotions in tactile enhanced multimedia.
method Frequency domain features from EEG data analyzed using SVM.
result Increased accuracy (76.19%) compared to time domain features (63.41%).
This dissertation improves classical compression techniques using deep learning.
problem Efficient storage and transmission of data in consumer and embedded applications.
method Leveraging deep learning to improve compression fidelity of classical algorithms.
result Improved compression ratios and visual quality using deep learning.
Bayesian neural networks improve reliability in multimedia forensics.
problem Challenges with out-of-distribution data in multimedia authentication.
method Proposes Bayesian neural networks (BNN) for forensic tasks.
result BNNs provide distributions for better reliability and out-of-distribution detection.
Mobile sensing is an emerging technology that utilizes agent-participatory data for decision making or state estimation, including multimedia applications. This article investigates the structure of mobile sensing schemes and introduces crowdsourcing methods for mobile sensing. Inspired by social network, one can estab…
Paper proposes detecting video manipulation using stream descriptors.
problem Misuse of manipulated video content.
method Binary classifiers on multimedia stream descriptors.
result Scalable approach can detect high-quality manipulations.
In the past few years, a lot of attention has been devoted to multimedia indexing by fusing multimodal informations. Two kinds of fusion schemes are generally considered: The early fusion and the late fusion. We focus on late classifier fusion, where one combines the scores of each modality at the decision level. To ta…
Calculating similarities between objects defined by many heterogeneous data modalities is an important challenge in many multimedia applications. We use a multi-modal topic model as a basis for defining such a similarity between objects. We propose to compare the resulting similarities from different model realizations…
Paper proposes a method to locate power grid recordings using ENF sequences.
problem Locating power grid recordings without concurrent power signals.
method Extract ENF sequences from power and audio recordings, develop multi-class SVM model.
result Validation of location authenticity of recordings using ENF sequences.
Paper classifies brain signals using eigenvalues for 2D and 3D educational content questions.
problem Classifying brain signals for 2D and 3D educational content questions.
method Eigenvalues of covariance matrix used as features; KNN and SVM classifiers applied.
result No significant difference in learning, memory retention, and recall between 2D and 3D educational content.
As a highlighting research topic in the multimedia area, cross-media retrieval aims to capture the complex correlations among multiple media types. Learning better shared representation and distance metric for multimedia data is important to boost the cross-media retrieval. Motivated by the strong ability of deep neura…
System identifies power grid location from media recordings.
problem Identifying the origin of power distribution grid from media recordings.
method Cascaded SVM and pole-matching classifiers for grid identification.
result Cascaded system improves accuracy by 15.57%.
Person re-identification (Re-ID) aims at matching images of the same person across disjoint camera views, which is a challenging problem in multimedia analysis, multimedia editing and content-based media retrieval communities. The major challenge lies in how to preserve similarity of the same person across video footag…
HyperLearn learns unified representations from multiple modalities efficiently.
problem Complex relational information in multimodal datasets.
method Hypergraph-based model with Graph Convolutional Networks.
result Unified representation learning for multimodal datasets without losing information.
Deep learning has enabled major advances in the fields of computer vision, natural language processing, and multimedia among many others. Developing a deep learning system is arduous and complex, as it involves constructing neural network architectures, managing training/trained models, tuning optimization process, pre…
In this paper, we introduce the Fairness GAN, an approach for generating a dataset that is plausibly similar to a given multimedia dataset, but is more fair with respect to protected attributes in allocative decision making. We propose a novel auxiliary classifier GAN that strives for demographic parity or equality of …
Face recognition system trained with noisy labels.
problem Label noise in training deep learning classifiers.
method Review and apply recent methods to manage noisy annotations.
result Improved performance of face recognition system with noisy labels.
Empirical law predicts accuracy of Google Translate's translation chains.
problem Predicting accuracy in machine translation with multiple hops.
method Empirical testing of Google Translate's sequential translation.
result Accuracy decreases with the number of translating hops, following a power law.
The authors of (Cho et al., 2014a) have shown that the recently introduced neural network translation systems suffer from a significant drop in translation quality when translating long sentences, unlike existing phrase-based translation systems. In this paper, we propose a way to address this issue by automatically se…
Paper proposes a method to learn word translations bidirectionally.
problem Word translation between languages.
method Jointly learns translations in both directions with minimal supervision.
result Improves accuracy of translations over previous methods.
Research has proven that stress reduces quality of life and causes many diseases. For this reason, several researchers devised stress detection systems based on physiological parameters. However, these systems require that obtrusive sensors are continuously carried by the user. In our paper, we propose an alternative a…
The Gaussian process latent variable model (GP-LVM) is a popular approach to non-linear probabilistic dimensionality reduction. One design choice for the model is the number of latent variables. We present a spike and slab prior for the GP-LVM and propose an efficient variational inference procedure that gives a lower …
Study on rigidity of translating hypersurfaces not in graphical direction.
problem Rigidity of translating hypersurfaces not in graphical direction.
method Proved rigidity results for complete graphical translating hypersurfaces under specific conditions.
result Entire graphical translating surfaces are flat under certain conditions.
Forward translation improves neural machine translation for sentences originally in source language.
problem Improving neural machine translation quality using synthetic data.
method Case study with French-English news translation, separating test sets by original language, analyzing domains, translationese, and noise.
result Forward translation delivers superior gains on sentences originally in source language, complementing back-translation on target language sentences.
The natural automorphism group of a translation surface is its group of translations. For finite translation surfaces of genus g > 1 the order of this group is naturally bounded in terms of g due to a Riemann-Hurwitz formula argument. In analogy with classical Hurwitz surfaces, we call surfaces which achieve the maxima…
A lot of attention has been devoted to multimedia indexing over the past few years. In the literature, we often consider two kinds of fusion schemes: The early fusion and the late fusion. In this paper we focus on late classifier fusion, where one combines the scores of each modality at the decision level. To tackle th…
Study on stable translation lengths of surface homeomorphisms and their approximations.
problem Understanding stable translation lengths of homeomorphisms and their finite approximations.
method Comparing stable translation lengths of homeomorphisms and their finite approximations on curve graphs.
result Stable translation length of homeomorphisms with dense periodic points equals the supremum of their approximations.
Researchers classify and describe Kα-translators in Euclidean space.
problem Classifying and describing Kα-translators in Euclidean space. method Rotationally symmetric and helicoidal motions.
result For each α, there is a Kα-translator intersecting orthogonally the rotation axis. The quality of machine translation is rapidly evolving. Today one can find several machine translation systems on the web that provide reasonable translations, although the systems are not perfect. In some specific domains, the quality may decrease. A recently proposed approach to this domain is neural machine translat…
Neural machine translation models trained for 5 South African languages.
problem Lack of resources and research for machine translation in African languages.
method Training neural machine translation models for 5 South African languages using modern techniques.
result Promises of neural machine translation for African languages.
Study measures gender bias in machine translation using multiple reference points.
problem Measuring and identifying gender bias in machine translation.
method Used an optimal non-biased translator, reference points from occupational statistics and survey.
result Found bias against both genders, but more against women, and found occupations have a greater effect than adjectives.
Classifies and constructs translators for curvature flows.
problem Understanding translating solitons in curvature flows.
method Developed rotational theory, introduced signed-neck framework.
result Classified and constructed catenoidal-type translators.
This paper uses LLMs and cycle consistency for better machine translation evaluation.
problem Evaluating translation quality and LLM capabilities without ground truth.
method Generate translation candidates, back-translate, and evaluate cycle consistency.
result Larger LLMs or more inference passes improve cycle consistency.
We propose a multi-wing harmonium model for mining multimedia data that extends and improves on earlier models based on two-layer random fields, which capture bidirectional dependencies between hidden topic aspects and observed inputs. This model can be viewed as an undirected counterpart of the two-layer directed mode…
Neural machine translation used to convert CUDA to OpenCL.
problem Translating CUDA to OpenCL programs.
method Training input set generation, pre/post processing, case study.
result Improved accuracy in translating CUDA to OpenCL.
Paper finds a non-existence theorem for certain translators in high dimensions.
problem Non-existence of certain translators in high-dimensional spaces.
method Developed a non-existence theorem and found an example of a translator.
result Non-existence of entire Qn−1-translators in Rn+1. The study explores how machine learning can enhance scientific research.
problem Improving scientific models with machine learning.
method Analysis of data-driven models versus manually added variables in regression.
result Complex models may not always improve over simpler ones in scientific contexts.
Constructing translating solitons from Lagrangian Grim Reapers.
problem Creating Lagrangian translating solitons from intersections of Grim Reapers.
method Desingularizing intersections with special Lagrangian Lawlor necks.
result Constructing Lagrangian translating solitons with multiple ends and loops.
Study classifies translators for mean curvature flow in 3D.
problem Classifying semigraphical translators for mean curvature flow in R3. method Morse-Radó theory and angular maximum principle.
result No solution to the translator equation on the upper half-plane with alternating boundary values.
Paper proposes AMSRE for multi-view data reduction.
problem Enhance performances of multi-view data tasks.
method Auto-weighted Multi-view Sparse Reconstructive Embedding (AMSRE).
result AMSRE effectively reduces multi-view data dimensions.
Study on singular points of translation surfaces under linearly dependent conditions.
problem Investigate singular points of translation surfaces under linearly dependent conditions.
method Use theories of generalised framed surfaces and framed surfaces.
result Introduce translation generalised framed surfaces and investigate their singular points.
Proves uniqueness of translators in 3D space.
problem Uniqueness of pitchfork and helicoid translators in mean curvature flow.
method Arc-counting argument and rotational maximum principle.
result Proves conjecture on uniqueness of translators.
An NMT system for Indic languages outperforms Google Translate.
problem Challenges in translating Indic languages efficiently.
method Encoder-decoder with attention mechanism for neural machine translation.
result Outperforms Google Translate with a 6 BLEU score margin on English-Gujarati translation.
New families of translation surfaces with multiple oblivious points discovered.
problem Identifying points on translation surfaces without nearby closed geodesics.
method Constructing new families of translation surfaces and proving existence in higher genera.
result Translation surfaces in every genus ≥3 have at least one oblivious point.
In this article we prove two non-existence results for translating solitons of the mean curvature flow (translators for short) in Rm+1. We also obtain an upper bound to the maximum height that a compact embedded translator in R3 can achieve. On the other hand, we study graphical perturbation…
CrossNet uses cross-consistency to improve unpaired image translation.
problem Image-to-image translation without paired data.
method Introduced a novel architecture with cross-translators and latent cross-consistency constraints.
result CrossNet outperforms state-of-the-art on various image translation tasks.
27 problems identified in automating movie/TV subtitle translation.
problem Challenges in translating movie/TV subtitles.
method Categorized problems into three categories and evaluated translation quality.
result Frontier NLP systems struggle with subtitles and require post-processing.
This paper classifies grim reapers in a specific product space.
problem Classifying grim reapers in a product space.
method Analyzing mean curvature flow and translations in $\h^2 imes
$.
result A full classification of grim reapers in $\h^2 imes
$.