<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
<channel>
  <title>pgmnemo</title>
  <link>https://pgmnemo.com/blog/</link>
  <description>Measurements, negative results, and engineering notes from building agent memory as a PostgreSQL extension.</description>
  <language>en</language>
  <atom:link href="https://pgmnemo.com/feed.xml" rel="self" type="application/rss+xml"/>
  <lastBuildDate>Thu, 14 Aug 2026 09:00:00 +0000</lastBuildDate>
  <item>
    <title>We A/B tested agent memory against no memory at all — on a live fleet. Here is what broke.</title>
    <link>https://pgmnemo.com/blog/agent-memory-ab-test</link>
    <guid isPermaLink="true">https://pgmnemo.com/blog/agent-memory-ab-test</guid>
    <pubDate>Thu, 14 Aug 2026 09:00:00 +0000</pubDate>
    <description>Three arms, 435 randomized runs on a production agent fleet, 30 days of real work. Always-on hybrid recall beat the no-memory control on task success (66.9% vs 55.3%, not significant after Bonferroni — every caveat stated). Selective recall, the arm this category sells as intelligence, performed exactly like having no memory at all (p=0.98) despite injecting more context. The de-identified dataset and a standard-library analysis script that reproduces every number are published in the repository.</description>
  </item>
</channel>
</rss>
