<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel>
<title>Harleen Kaur — Engineering notes</title><link>https://hk-775.github.io/hk-775/blog/</link>
<description>Engineering decisions, experiments, and evidence for AI workloads.</description>
<language>en</language><atom:link href="https://hk-775.github.io/hk-775/blog/feed.xml" rel="self" type="application/rss+xml"/>
<item><title>Designing an Evidence Trail for Agent Actions</title><link>https://hk-775.github.io/hk-775/blog/designing-an-evidence-trail-for-agent-actions.html</link><guid isPermaLink="true">https://hk-775.github.io/hk-775/blog/designing-an-evidence-trail-for-agent-actions.html</guid><pubDate>Mon, 05 Oct 2026 12:00:00 GMT</pubDate><description>I trace one synthetic agent action from its request to its observed effect, then examine what event hashes, snapshots, and replay can establish—and what needs additional verification.</description><category>Agent engineering</category><category>Evidence</category><category>Action governance</category></item>
<item><title>When rules beat decision models</title><link>https://hk-775.github.io/hk-775/blog/when-rules-beat-decision-models.html</link><guid isPermaLink="true">https://hk-775.github.io/hk-775/blog/when-rules-beat-decision-models.html</guid><pubDate>Fri, 02 Oct 2026 12:00:00 GMT</pubDate><description>I replaced a label-matching diagnostic with an executable support workflow. On 144 synthetic test episodes, the rules baseline outperformed the tested zero-shot model configurations.</description><category>Evaluation</category><category>Tool selection</category><category>Agent engineering</category></item>
</channel></rss>
