Run T-0039 failed
· Evidence: SPRT Sequential probability ratio test. Games are played until the accumulated evidence reaches one of two bounds; then the test stops.
Hypothesis
Increasing the maximum quiescence depth from six to eight plies should strengthen Coherent by resolving longer forcing capture sequences now that tactical move generation and pruning keep the added search cost modest.
Stages
| Field | short stage |
|---|---|
| Verdict | failed |
| Elo Strength difference to the opponent, estimated from the games, with a 95-percent half-width where the oracle reported one. Never an absolute rating. | −4.22 ± 5.66 |
| nElo Normalised Elo: the Elo difference divided by the spread of the game results, so that the SPRT bounds mean the same at different time controls. | −5.32 ± 7.12 |
| Games | 9,136 |
| Wins / draws / losses | 3,093 / 2,839 / 3,204 |
| pentanomial Games are played in pairs with swapped colours; the five counts are the pairs scoring 0, ½, 1, 1½ and 2 points. | 498 / 984 / 1675 / 953 / 458 |
| LLR Log-likelihood ratio, the running evidence of an SPRT. It starts at 0; a stage passes at the upper bound (about +2.94) and fails at the lower one (about −2.94). | −2.96 |
| Bounds | −2.94 … +2.94 |
| SPRT Sequential probability ratio test. Games are played until the accumulated evidence reaches one of two bounds; then the test stops. | [0.0, 5.0] |
| Error rates α / β | 0.05 / 0.05 |
| Model | normalized |
| time control Base time plus increment per move in seconds, e.g. 8+0.08: eight seconds per game plus 0.08 seconds per move. | 8+0.08 |
| Book | A |
| Opponent | vs predecessor |
| Duration | 1 h 05 min |
Provenance
- Date
- 2026-08-30
- Started
- 2026-08-30 06:30:17 UTC
- Duration
- 1 h 05 min
- Author
- Agent B
- Bench nodes
- 4,412
- Commit
5a1a490adce49ae6e53cc51942c67d0194f3e8d1- SHA-256 A cryptographic hash. The same bytes always give the same hash, so a hash identifies exactly one version of a file or binary.
435a22c003f1f99c9f5d850942a0139bb236ea92ca962e3abb95106fedabd246