Drop7 Research

Global Game · August 2026 · 2026-08-v1

Human + AI leaderboard

Every entry is scored on gauntlet-01. Human scores come from server-replayed move sequences; AI scores come from the same scripted-round harness. This is a reproducible playground, not research-tier evidence.

Play this game →
#PlayerTypeVerified scoreMoves
1Gray throughput
extended-state policy · local artifact
Research →
ai245,69775
2Open-loop beam
heuristic-search · public policy · 8/22/2026
Research →
ai212,21365
3Greedy 1-ply
heuristic-search · public policy · 8/22/2026
Research →
ai207,95665
4Expectimax D2
fair-expectimax · public policy · 8/22/2026
Research →
ai191,98960
5TypeScript Expectimax D3
fair-expectimax · public policy · 8/22/2026
Research →
ai191,87360
6TypeScript Expectimax D4
fair-expectimax · public policy · 8/22/2026
Research →
ai191,87360
7Risk-sensitive D2
heuristic-search · public policy · 8/22/2026
Research →
ai160,23750
8MCTS
tree-search · public policy · 8/22/2026
Research →
ai158,18750
9Rollout H8
heuristic-search · public policy · 8/22/2026
Research →
ai122,52740
10Sparse expectimax D2
heuristic-search · public policy · 8/22/2026
Research →
ai120,96840

Computer policy leaderboard

Every autonomous policy plays the exact same predetermined rounds. The visible disc sequence is fixed by move number, and every gray disc hides a fixed value that takes its place when revealed. Two policies therefore face identical randomness on every move. How the benchmark works →

Playground evidence only. These competitions do not support qualification claims as per the benchmark guidelines.

#PolicyMeanMedianMinMaxMovesG01
1
Gray throughputextended state
heuristic-search
Research →
245,6970245,697245,69775.0246k
2
Open-loop beam
heuristic-search
Research →
212,2130212,213212,21365.0212k
3
Greedy 1-ply
heuristic-search
Research →
207,9560207,956207,95665.0208k
4
Expectimax D2
fair-expectimax
Research →
191,9890191,989191,98960.0192k
5
Risk-sensitive D2
heuristic-search
Research →
160,2370160,237160,23750.0160k
6
MCTS
tree-search
Research →
158,1870158,187158,18750.0158k
7
Rollout H8
heuristic-search
Research →
122,5270122,527122,52740.0123k
8
Sparse expectimax D2
heuristic-search
Research →
120,9680120,968120,96840.0121k

Flow diagnostics

PolicyClears / moveReveals / moveMax chainCensoredIllegalCompute
Gray throughput1.95/ 2.40 target1.03/ 1.40 target7000.2s
Open-loop beam1.71/ 2.40 target0.83/ 1.40 target6000.9s
Greedy 1-ply1.72/ 2.40 target0.85/ 1.40 target4001.0s
Expectimax D21.68/ 2.40 target0.88/ 1.40 target6005.7s
Risk-sensitive D21.74/ 2.40 target0.96/ 1.40 target6005.8s
MCTS1.70/ 2.40 target0.86/ 1.40 target5003.3s
Rollout H81.32/ 2.40 target0.53/ 1.40 target5001.0s
Sparse expectimax D21.55/ 2.40 target0.82/ 1.40 target5000.4s

The 2.4 clears / 1.4 reveals per move figures are diagnostic targets from limited task-record runs, not proven thresholds. Click any score above to replay that game move by move.