This paper presents the first two editions of Visual Doom AI Competition, held in 2016 and 2017. The challenge was to create bots that compete in a multi-player deathmatch in a first-person shooter (FPS) game, Doom. The bots had to make their decisions based solely on visual information, i.e., a raw screen buffer. To p…
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Recent developments in deep reinforcement learning have enabled the creation of agents for solving a large variety of games given a visual input. These methods have been proven successful for 2D games, like the Atari games, or for simple tasks, like navigating in mazes. It is still an open question, how to address more…
We applied Generative Adversarial Networks (GANs) to learn a model of DOOM levels from human-designed content. Initially, we analysed the levels and extracted several topological features. Then, for each level, we extracted a set of images identifying the occupied area, the height map, the walls, and the position of ga…
A number of recent approaches to policy learning in 2D game domains have been successful going directly from raw input images to actions. However when employed in complex 3D environments, they typically suffer from challenges related to partial observability, combinatorial exploration spaces, path planning, and a scarc…
SMiRL learns to minimize surprise in unstable environments, improving agent performance.
In this paper, we prove interior Poincar{é} and Sobolev inequalities in Euclidean spaces and in Heisenberg groups, in the limiting case where the exterior (resp. Rumin) differential of a differential form is measured in L 1 norm. Unlike for L p , p > 1, the estimates are doomed to fail in top degree. The singular integ…
Unbiased wealth exchanges always lead to inequality.
Motivated by recent advance of machine learning using Deep Reinforcement Learning this paper proposes a modified architecture that produces more robust agents and speeds up the training process. Our architecture is based on Asynchronous Advantage Actor-Critic (A3C) algorithm where the total input dimensionality is halv…
We employ min-max methods to construct uncountably many, geometrically distinct, properly embedded geodesic lines in any asymptotically conical surface of non-negative scalar curvature, a setting where minimization schemes are doomed to fail. Our construction provides control of the Morse index of the geodesic lines we…
We present a model of predatory traders interacting with each other in the presence of a central reserve (which dissipates their wealth through say, taxation), as well as inflation. This model is examined on a network for the purposes of correlating complexity of interactions with systemic risk. We suggest the use of s…
MIME uses mutual information minimization for better exploration in environments with abrupt transitions.
GNNs may be limited by graph topology, affecting their learning outcomes.
Economy is demanding new models, able to understand and predict the evolution of markets. To this respect, Econophysics is offering models of markets as complex systems, such as the gas-like model, able to predict money distributions observed in real economies. However, this model reveals some technical hitches to expl…
Deep learning has been widely accepted as a promising solution for medical image segmentation, given a sufficiently large representative dataset of images with corresponding annotations. With ever increasing amounts of annotated medical datasets, it is infeasible to train a learning method always with all data from scr…
Learning robust value functions given raw observations and rewards is now possible with model-free and model-based deep reinforcement learning algorithms. There is a third alternative, called Successor Representations (SR), which decomposes the value function into two components -- a reward predictor and a successor ma…
Model shows how banks' fears of future defaults can cause immediate financial stress.
In the era of social media and networking platforms, Twitter has been doomed for abuse and harassment toward users specifically women. Monitoring the contents including sexism and sexual harassment in traditional media is easier than monitoring on the online social media platforms like Twitter, because of the large amo…
We consider the problem of learning to play first-person shooter (FPS) video games using raw screen images as observations and keyboard inputs as actions. The high-dimensionality of the observations in this type of applications leads to prohibitive needs of training data for model-free methods, such as the deep Q-netwo…
Economy is demanding new models, able to understand and predict the evolution of markets. To this respect, Econophysics offers models of markets as complex systems, that try to comprehend macro-, system-wide states of the economy from the interaction of many agents at micro-level. One of these models is the gas-like mo…
Computing optimal transport (OT) between measures in high dimensions is doomed by the curse of dimensionality. A popular approach to avoid this curse is to project input measures on lower-dimensional subspaces (1D lines in the case of sliced Wasserstein distances), solve the OT problem between these reduced measures, a…