FuryBench Testing Framework

home toIntegralPlentyChessObsidianCaissaQuanticadeSerendipitySpaghetSchoenemannClarityIlluminaPawnocchioVineAnura
Finished
Caissa Witekdiff10.0+0.10
ELO    0.24 +- 1.43 (95%)
LLR    2.96 (-2.94, 2.94) [-3.00, 0.00]
GAMES  N: 60648 W: 14199 L: 14157 D: 32292
PENTA  [222, 7191, 15421, 7303, 187]
Integral Furydiff40.0+0.40
ELO    3.70 +- 2.40 (95%)
LLR    2.91 (-2.25, 2.89) [0.00, 3.00]
GAMES  N: 19354 W: 4858 L: 4652 D: 9844
PENTA  [20, 2150, 5125, 2368, 14]
Quanticade DarkNeutrino
Implement Pawn Pawn input set
diff60.0+0.60
ELO    -1.25 +- 8.85 (95%)
LLR    -0.14 (-2.25, 2.89) [0.00, 3.00]
GAMES  N: 1394 W: 326 L: 331 D: 737
PENTA  [2, 159, 379, 156, 1]
Quanticade Anematode
ojpowifjpawoifjepowijfpoi
diff4.0+0.04
ELO    0.56 +- 0.68 (95%)
LLR    -2.35 (-2.25, 2.89) [0.00, 3.00]
GAMES  N: 251810 W: 62154 L: 61745 D: 127911
PENTA  [1103, 27112, 69080, 27493, 1117]
Quanticade DarkNeutrino
Implement Pawn Pawn input set
diff10.0+0.10
ELO    -9.03 +- 5.07 (95%)
LLR    -2.32 (-2.25, 2.89) [0.00, 3.00]
GAMES  N: 5002 W: 1140 L: 1270 D: 2592
PENTA  [30, 654, 1248, 554, 15]
Quanticade DarkNeutrino
Pawn pawn
diffN=5000
ELO    13.17 +- 5.83 (95%)
LLR    3.07 (-2.25, 2.89) [0.00, 3.00]
GAMES  N: 7680 W: 2708 L: 2417 D: 2555
PENTA  [283, 795, 1479, 914, 369]
Integral Furydiff10.0+0.10
ELO    2.07 +- 1.63 (95%)
LLR    2.90 (-2.25, 2.89) [0.00, 3.00]
GAMES  N: 46272 W: 11652 L: 11376 D: 23244
PENTA  [162, 5354, 11823, 5640, 157]
Caissa Witeksmp diff2.0+0.02
ELO    1.82 +- 2.87 (95%)
LLR    2.97 (-2.94, 2.94) [-4.00, 0.00]
GAMES  N: 16772 W: 4307 L: 4219 D: 8246
PENTA  [114, 1996, 4082, 2076, 118]
Integral FurydiffN=20000
ELO    1.54 +- 1.17 (95%)
LLR    3.08 (-2.25, 2.89) [0.00, 3.00]
GAMES  N: 142280 W: 42474 L: 41845 D: 57961
PENTA  [3210, 16701, 30707, 17294, 3228]
Quanticade Anematode
move lover
diff4.0+0.04
ELO    2.30 +- 1.76 (95%)
LLR    2.94 (-2.25, 2.89) [0.00, 3.00]
GAMES  N: 39656 W: 9923 L: 9660 D: 20073
PENTA  [193, 4321, 10543, 4572, 199]
Caissa Witek
TimeManager: Cross-thread best move instability Each search thread counts how often its own root move changes between iterations. The main thread sums those counters, normalizes by thread count and folds the result into the soft time limit as one more multiplicative factor, so time allocation sees the helper threads disagreeing instead of only the main thread's own stability streak.
smp diff5.0+0.05
ELO    -5.48 +- 3.48 (95%)
LLR    -2.95 (-2.94, 2.94) [0.00, 3.00]
GAMES  N: 9516 W: 2180 L: 2330 D: 5006
PENTA  [12, 1212, 2460, 1062, 12]
Caissa Witek
QSearch: Node-level delta pruning Bail out of a quiescence node before move generation when even winning the most valuable piece on the board (plus a queen promotion when a pawn stands on the relative 7th) cannot reach alpha.
diff10.0+0.10
ELO    -4.59 +- 2.71 (95%)
LLR    -2.68 (-2.94, 2.94) [0.00, 2.00]
GAMES  N: 15914 W: 3629 L: 3839 D: 8446
PENTA  [53, 1915, 4210, 1747, 32]
Quanticade DarkNeutrino
2 speedups both doing the same thing tho
diff4.0+0.04
ELO    2.56 +- 2.67 (95%)
LLR    1.50 (-2.25, 2.89) [0.00, 3.00]
GAMES  N: 17914 W: 4417 L: 4285 D: 9212
PENTA  [101, 1990, 4655, 2098, 113]
Quanticade DarkNeutrino
search.c: Reduce more for tt_pv ahead of pruning
diff60.0+0.60
ELO    -0.70 +- 3.99 (95%)
LLR    -0.48 (-2.25, 2.89) [0.00, 3.00]
GAMES  N: 6426 W: 1490 L: 1503 D: 3433
PENTA  [1, 710, 1803, 699, 0]
Quanticade DarkNeutrino
search.c: Reduce more if this node is cutnode and we dont have TT move
diff10.0+0.10
ELO    -2.64 +- 2.91 (95%)
LLR    -2.26 (-2.25, 2.89) [0.00, 3.00]
GAMES  N: 13306 W: 3064 L: 3165 D: 7077
PENTA  [20, 1592, 3529, 1493, 19]