•7 min read
Winning sooner, not just winning — teaching the AI to value a quick kill
The YINSH AI thought a win one move away and a win a hundred moves away were exactly the same thing. Fixing that — teaching it to prefer winning sooner — made it both stronger and cheaper to run, and taught me the value network is a search multiplier, not just a scorekeeper.