Survey of deep learning for Hindi text classification.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
An NMT system for Indic languages outperforms Google Translate.
We develop a general theory of convex duality for certain singular control problems, taking the abstract results by Kramkov and Schachermayer (1999) for optimal expected utility from nonnegative random variables to the level of optimal expected utility from increasing, adapted controls. The main contributions are the f…
Project classifies Hinglish social content on platforms like Twitter, Reddit.
Dataset for measuring reading levels in India's children.
Visual Speech Recognition (VSR) is the process of recognizing or interpreting speech by watching the lip movements of the speaker. Recent machine learning based approaches model VSR as a classification problem; however, the scarcity of training data leads to error-prone systems with very low accuracies in predicting un…
New ASR system handles multiple languages without needing language-specific encoding.
TeLeS improves ASR confidence estimation by considering temporal alignment and lexical errors.
Paper identifies and solves a 'scrambled translation' issue in UNMT models.