Random play trains a DQN to win at Sungka.
problem Optimizing game strategies through self-play.
method Training a DQN agent with random play.
result DQN trained with random play converges fast and consistently wins.
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Random play trains a DQN to win at Sungka.