A learned survival estimate inside the reference search
Blends a learned survival estimate into the reference board evaluator and tests it under sampled and exact chance handling.
completedreproducedEach page here is one theory of how to choose a column, grouped by the technique it uses; the engines that play the games and the instruments that measure them live under Engines and Diagnostics.
Judge a move by how much longer the game will still last, rather than by how many points it scores right now.
Instead of examining every column to a fixed depth, grow the look-ahead only where it looks promising, guided by quick simulated playouts.
Instead of searching ahead, train a model on past games to judge a board or pick a column, and learn why that kept failing.
No approach page matches that search.