| |
BGonline.org Forums
Should announcers see bot results/ HUGE hedge hog weakness
Posted By: equitin In Response To: Should announcers see bot results/ HUGE hedge hog weakness (Chuck Bower)
Date: Sunday, 30 August 2026, at 9:35 a.m.
There seems to be a lot of braggadocious chest beating among the new bots with, from what I can tell from a distance, zero reliable evidence.
I am curious about HedgeHog. The quotes below are from their /benchmarks page at https://hedgehog-bg.com/benchmarks.
It is not clear to me yet how they automated play against e.g. XG ("All competing engines are imported into the OGXF format...")
Each matchup is decided by 10 million head-to-head games of self-play.
All competing engines are imported into the OGXF format and run inside Hedgehog's own engine. The move generation, search, and cube logic are therefore identical across every model; the only thing that changes is the neural network doing the evaluation. These numbers reflect differences between the networks, not between separate software implementations, and may not match results published by the original programs.
0-ply means each engine plays straight from its neural network's evaluation, with no lookahead search. That isolates the quality of the network itself, which is what these benchmarks are meant to compare. Add search (1-ply, 2-ply, 3-ply) and every strong engine's lookahead corrects most of the same mistakes, so their results converge and the gap between networks shrinks toward zero. Differences that are clearly visible at 0-ply become hard to measure once deep search papers over them, so 0-ply is the most sensitive and honest way to tell two networks apart.
| |
BGonline.org Forums is maintained by Stick with WebBBS 5.12.