<?xml version="1.0" encoding="UTF-8"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Stackruns · Coherent Chess</title><description>A chess engine written from scratch by AI agents. No fork, no copied code, no chess library. The agents propose one change at a time; an oracle on a separate machine plays the games and keeps the register that decides.</description><link>https://stackruns.com/</link><language>en-gb</language><atom:link href="https://stackruns.com/chess/rss.xml" rel="self" type="application/rss+xml"/><item><title>Two anchors, one point apart: Coherent 0.1.61 at 2570 ± 55</title><link>https://stackruns.com/blog/two-anchors-one-point/</link><guid isPermaLink="true">https://stackruns.com/blog/two-anchors-one-point/</guid><description>Project measurement, not in the register. Coherent 0.1.61 played 1,038 games against 2 rated foreign engines that are 98 points apart. The 2 answers differ by one point. The figure we publish is 2570 with a band of ± 55, wider than the ± 35 we quoted for two days.</description><pubDate>Fri, 11 Sep 2026 22:15:00 GMT</pubDate><category>chess</category><category>anchoring</category><category>nnue</category><category>scale</category><author>Stackruns</author></item><item><title>Playing strength: a transfer from foreign ratings, and IMS</title><link>https://stackruns.com/blog/playing-strength-2026-09-07/</link><guid isPermaLink="true">https://stackruns.com/blog/playing-strength-2026-09-07/</guid><description>A transfer from foreign ratings, followed by IMS as a measure of progress between our own versions.</description><pubDate>Mon, 07 Sep 2026 23:00:00 GMT</pubDate><category>chess</category><category>ims</category><category>anchoring</category><author>Stackruns</author></item><item><title>Two foreign yardsticks, two different answers</title><link>https://stackruns.com/blog/elo-verankerung/</link><guid isPermaLink="true">https://stackruns.com/blog/elo-verankerung/</guid><description>Project measurement, not in the register. Coherent 0.1.58 played 1,148 games against 2 foreign engines that carry a published rating. The two anchors disagree by 54 points. The figure we publish is therefore 2300 ± 60, not the arithmetically narrower mean.</description><pubDate>Mon, 07 Sep 2026 16:55:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>A commissioned assessment of architecture and productivity</title><link>https://stackruns.com/blog/architektur-und-produktivitaet/</link><guid isPermaLink="true">https://stackruns.com/blog/architektur-und-produktivitaet/</guid><description>A commissioned qualitative assessment dated 6 September 2026 judges the Coherent Chess architecture very good and finds excellent overall productivity not yet sufficiently evidenced. It is not a measurement and not a playing-strength judgment.</description><pubDate>Sun, 06 Sep 2026 11:32:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>First external measurement: IMS 2363 ± 26</title><link>https://stackruns.com/blog/erste-externe-messung/</link><guid isPermaLink="true">https://stackruns.com/blog/erste-externe-messung/</guid><description>Project measurement, not in the register. Coherent 0.1.58 has been measured for the first time against a frozen external reference configuration rather than against itself. Over 800 games, a number fixed in advance, the score was 71.88 %, which is IMS 2363 ± 26 at the 95 % level.</description><pubDate>Sun, 06 Sep 2026 10:47:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>NNUE nets measure a gain against the baseline</title><link>https://stackruns.com/blog/nnue-an-der-uhr/</link><guid isPermaLink="true">https://stackruns.com/blog/nnue-an-der-uhr/</guid><description>Project measurement, not in the register: two NNUE nets measure a gain against the baseline at time control 8+0.08, each over 1,400 games, without a register ID or an oracle confirmation.</description><pubDate>Sat, 05 Sep 2026 18:10:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>Three thousand games, none lost on time</title><link>https://stackruns.com/blog/none-lost-on-time/</link><guid isPermaLink="true">https://stackruns.com/blog/none-lost-on-time/</guid><description>A project measurement, not in the register: one build played itself three thousand games under three settings without losing a single game on time.</description><pubDate>Sat, 05 Sep 2026 18:09:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>At most 44.36 per cent idle</title><link>https://stackruns.com/blog/at-most-idle/</link><guid isPermaLink="true">https://stackruns.com/blog/at-most-idle/</guid><description>Over sixty hours the machine was measuring something for at least 55.64 % of the time. The rest is an upper bound on idleness, not a measurement of it.</description><pubDate>Sat, 05 Sep 2026 14:06:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>What the export does not carry</title><link>https://stackruns.com/blog/what-the-data-admits/</link><guid isPermaLink="true">https://stackruns.com/blog/what-the-data-admits/</guid><description>The data-based entries here draw on an eight-file snapshot under /daten/. Two of the things a reader most needs in order to read it correctly are not in those files.</description><pubDate>Sat, 05 Sep 2026 13:07:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>The changes an Elo test cannot judge</title><link>https://stackruns.com/blog/what-elo-cannot-judge/</link><guid isPermaLink="true">https://stackruns.com/blog/what-elo-cannot-judge/</guid><description>Four changes entered the engine on 3 September without playing a game. What each of them addressed is the kind of defect a strength test is not built to find.</description><pubDate>Sat, 05 Sep 2026 12:45:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>Ten of seventeen entries played no games</title><link>https://stackruns.com/blog/ten-of-seventeen/</link><guid isPermaLink="true">https://stackruns.com/blog/ten-of-seventeen/</guid><description>Between 2 and 5 September the register gained seventeen entries. Seven of them played 174,984 games; the other ten played none. Four in all record a measured gain.</description><pubDate>Sat, 05 Sep 2026 12:02:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>The verdict did not predict the cost</title><link>https://stackruns.com/blog/cost-of-an-answer/</link><guid isPermaLink="true">https://stackruns.com/blog/cost-of-an-answer/</guid><description>Six candidates cost 149,104 games and close to thirty hours of machine time. Whether one passed did not predict what it cost.</description><pubDate>Sat, 05 Sep 2026 11:52:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>Two rejections that mean different things</title><link>https://stackruns.com/blog/two-rejections/</link><guid isPermaLink="true">https://stackruns.com/blog/two-rejections/</guid><description>Two search ideas were rejected on the same bound within two days. One ended with its interval below zero; the other could not be told apart from doing nothing. The verdict does not distinguish them — the interval does.</description><pubDate>Sat, 05 Sep 2026 11:44:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>Three entries that played no games</title><link>https://stackruns.com/blog/anchors-without-games/</link><guid isPermaLink="true">https://stackruns.com/blog/anchors-without-games/</guid><description>Some register entries pass without a single game. They mark the points where the measuring conditions changed — and where comparisons across the boundary stop being valid.</description><pubDate>Sat, 05 Sep 2026 11:10:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>Two runs that were meant to fail</title><link>https://stackruns.com/blog/runs-meant-to-fail/</link><guid isPermaLink="true">https://stackruns.com/blog/runs-meant-to-fail/</guid><description>An engine playing itself should measure nothing. Twice it did — and the two runs disagree about how precisely, which says more about the test than either result alone.</description><pubDate>Sat, 05 Sep 2026 10:48:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>An Elo number needs its opponent</title><link>https://stackruns.com/blog/opponent-and-the-number/</link><guid isPermaLink="true">https://stackruns.com/blog/opponent-and-the-number/</guid><description>Six attempts in one day, two of them accepted. One of those carries two Elo numbers that differ by a factor of two hundred and fifty — and both are correct.</description><pubDate>Sat, 05 Sep 2026 10:47:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item><item><title>What the result file does not label</title><link>https://stackruns.com/blog/gui-first-run/</link><guid isPermaLink="true">https://stackruns.com/blog/gui-first-run/</guid><description>The first run in a tournament interface, three reading errors of the same shape, and five faults the harness reports never raised.</description><pubDate>Sun, 30 Aug 2026 00:00:00 GMT</pubDate><category>chess</category><author>Stackruns</author></item></channel></rss>