<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>securesein — Reinforcement learning</title><description>RLHF, RLVR, reward models, policy optimisation.</description><link>https://securesein.com/</link><item><title>Curiosity at the Edge of Chaos: What a New RL Paper Actually Shows</title><link>https://securesein.com/blog/curiosity-edge-of-chaos/</link><guid isPermaLink="true">https://securesein.com/blog/curiosity-edge-of-chaos/</guid><description>A new reinforcement learning paper claims curiosity should emerge from an agent&apos;s own internal dynamics rather than a hand-tuned bonus term — a skeptical read of what the experiments actually support.</description><pubDate>Sat, 12 Sep 2026 00:00:00 GMT</pubDate><author>Sebastiaan with AI</author><category>Research — Paper</category><category>reinforcement-learning</category><category>deep-learning</category></item></channel></rss>