Clash Royale

Teaching an AI to Game the System: How a Cannon Found the Perfect Loophole

The agent found the loophole because the rules allowed it.

4 min readMachine Learning
Teaching an AI to Game the System: How a Cannon Found the Perfect Loophole
ClashRoyaleAi: an open-source, deterministic Clash Royale simulator for RL, with recurrent PPO, lookahead search and expert iteration [P]
From Machine Learning

The opponent plans by simulation: every second it scores each candidate play by running the match 10 seconds ahead in the engine.

Our PPO agent learned to park its Cannon behind its own King. Losing a building in a fight cost reward, and letting it decay cost nothing, so it found the loophole.

Read the original at Machine Learning