<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>Prasun Srivastava — Writing</title><description>Essays on AI-assisted software development: harnesses, mechanical gates, review at agent throughput, and keeping control of codebases AI writes most of.</description><link>https://prasunsrivastava.com/</link><language>en-us</language><item><title>Why I run every eval more than once</title><link>https://prasunsrivastava.com/writing/run-every-eval-more-than-once/</link><guid isPermaLink="true">https://prasunsrivastava.com/writing/run-every-eval-more-than-once/</guid><description>I ran the same AI code review five times at temperature zero, with everything held identical, and got a different answer two times out of five. A single run can push you to deploy the wrong model without ever knowing. This is how I size the number of eval repeats by what the decision is worth.</description><pubDate>Wed, 22 Jul 2026 00:00:00 GMT</pubDate><author>prasun@prasunsrivastava.com (Prasun Srivastava)</author></item><item><title>Testing a multi-agent AI code-review gate against a single prompt</title><link>https://prasunsrivastava.com/writing/review-gate-vs-single-prompt/</link><guid isPermaLink="true">https://prasunsrivastava.com/writing/review-gate-vs-single-prompt/</guid><description>I rebuilt Cloudflare&apos;s multi-agent review gate and ran it against a plain single-prompt review, on 12 commits from a real codebase&apos;s own history. The single prompt caught more of the known bugs at about a quarter of the cost, and the runs surfaced 9 shipped bugs nobody had caught.</description><pubDate>Wed, 15 Jul 2026 00:00:00 GMT</pubDate><author>prasun@prasunsrivastava.com (Prasun Srivastava)</author></item><item><title>Auditing my own AI coding harness for enforcement</title><link>https://prasunsrivastava.com/writing/auditing-my-own-ai-coding-harness-for-enforcement/</link><guid isPermaLink="true">https://prasunsrivastava.com/writing/auditing-my-own-ai-coding-harness-for-enforcement/</guid><description>I audited the AI coding harness I wrote for a production codebase in 2025: 12,000 lines of rules, zero mechanical checks. What I moved to gates, what stayed prose, and why agents made the enforcement layer necessary.</description><pubDate>Wed, 08 Jul 2026 00:00:00 GMT</pubDate><author>prasun@prasunsrivastava.com (Prasun Srivastava)</author></item></channel></rss>