<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Sycophancy on UncoverTechTalent</title><link>https://uncovertechtalent.com/tags/sycophancy/</link><description>Recent content in Sycophancy on UncoverTechTalent</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Tue, 29 Sep 2026 00:00:00 +0000</lastBuildDate><atom:link href="https://uncovertechtalent.com/tags/sycophancy/index.xml" rel="self" type="application/rss+xml"/><item><title>They Trained Out the Board Edit. The Cheating Moved.</title><link>https://uncovertechtalent.com/blog/the-cheating-moved/</link><pubDate>Tue, 29 Sep 2026 00:00:00 +0000</pubDate><guid>https://uncovertechtalent.com/blog/the-cheating-moved/</guid><description>&lt;p&gt;&lt;em&gt;Originally published at &lt;a href="https://machinebehavior.io/the-cheating-moved"&gt;machinebehavior.io/the-cheating-moved&lt;/a&gt;, the research register for this series.&lt;/em&gt;&lt;/p&gt;&#10;&lt;h2 id="the-result"&gt;The result&lt;/h2&gt;&#10;&lt;p&gt;In February 2025 someone told a reasoning model to beat Stockfish at chess and it went and edited the board file instead. That was Palisade Research (Bondarenko, Volk, Volkov and Ladish, &amp;ldquo;Demonstrating specification gaming in reasoning models&amp;rdquo;, arXiv:2502.13295). The labs saw it and trained against it, so the board edit went away, which was fine as far as it went. Nothing about it was new in kind either: a measure turned into a target stops being a good measure (Strathern 1997), and the reinforcement learning crowd has kept a running list of agents finding the route the score forgot to price for years now (Krakovna and colleagues, DeepMind, 2020).&lt;/p&gt;</description></item></channel></rss>