<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>StackUnseen</title>
    <link>https://stackunseen.com</link>
    <description>What I built, tested and would use again in an enterprise: architecture maps, decisions, code and failures from real AI systems.</description>
    <language>en</language>
    <atom:link href="https://stackunseen.com/feed.xml" rel="self" type="application/rss+xml"/>
    <item>
      <title>Build log: shipping Unseen UI to npm</title>
      <link>https://stackunseen.com/journal/build-log-unseen-ui-alpha</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/build-log-unseen-ui-alpha</guid>
      <pubDate>Thu, 08 Oct 2026 00:00:00 GMT</pubDate>
      <description>How the component library behind this site went from a private workspace to two published packages, including the three things the first consumer broke.</description>
      <category>Build log</category>
      <category>unseen-ui</category><category>react</category><category>developer-tools</category>
    </item>
    <item>
      <title>Week 41: five areas, one journey</title>
      <link>https://stackunseen.com/journal/weekly-note-2026-w41</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/weekly-note-2026-w41</guid>
      <pubDate>Thu, 08 Oct 2026 00:00:00 GMT</pubDate>
      <description>The site moved to a five-area structure, every entry now hangs off a map node, and the first problem-to-production journey is live.</description>
      <category>Weekly note</category>
      <category>weekly</category><category>meta</category>
    </item>
    <item>
      <title>LangGraph vs the OpenAI Agents SDK</title>
      <link>https://stackunseen.com/journal/langgraph-vs-agents-sdk</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/langgraph-vs-agents-sdk</guid>
      <pubDate>Wed, 07 Oct 2026 00:00:00 GMT</pubDate>
      <description>Two ways to write the same supervisor. Compared on control flow, tracing, provider coupling, testing and what each makes hard.</description>
      <category>Comparison</category>
      <category>agents</category><category>orchestration</category><category>comparison</category>
    </item>
    <item>
      <title>Week 40: alpha out, radar in</title>
      <link>https://stackunseen.com/journal/weekly-note-2026-w40</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/weekly-note-2026-w40</guid>
      <pubDate>Mon, 05 Oct 2026 00:00:00 GMT</pubDate>
      <description>Unseen UI alpha.1 reached npm, the tools page grew a radar ring, and one retrieval experiment got abandoned on purpose.</description>
      <category>Weekly note</category>
      <category>weekly</category><category>unseen-ui</category><category>rag</category>
    </item>
    <item>
      <title>Every question your AI readiness review asks was answered months ago</title>
      <link>https://stackunseen.com/journal/guide-production-ai-readiness</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/guide-production-ai-readiness</guid>
      <pubDate>Sun, 04 Oct 2026 00:00:00 GMT</pubDate>
      <description>The last six days of a sixty-day series, and the pattern is that operability gets bought early or it does not get bought at all.</description>
      <category>Deep dive</category>
      <category>architecture</category><category>observability</category><category>tracing</category>
    </item>
    <item>
      <title>RAG production-readiness checklist</title>
      <link>https://stackunseen.com/journal/rag-production-readiness-checklist</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/rag-production-readiness-checklist</guid>
      <pubDate>Sat, 03 Oct 2026 00:00:00 GMT</pubDate>
      <description>Twelve checks to pass before a retrieval system answers a real user. Tick them locally; progress stays in your browser.</description>
      <category>Checklist</category>
      <category>rag</category><category>evaluation</category><category>operations</category>
    </item>
    <item>
      <title>What an AI gateway actually costs to run</title>
      <link>https://stackunseen.com/journal/what-an-ai-gateway-costs</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/what-an-ai-gateway-costs</guid>
      <pubDate>Fri, 02 Oct 2026 00:00:00 GMT</pubDate>
      <description>The operational bill for one gateway in front of a dozen tool servers, including the costs nobody budgets for.</description>
      <category>Article</category>
      <category>mcp</category><category>gateway</category><category>operations</category><category>cost</category>
    </item>
    <item>
      <title>Working with Claude, practices that hold up</title>
      <link>https://stackunseen.com/journal/working-with-claude-practices</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/working-with-claude-practices</guid>
      <pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate>
      <description>Prompting, agents and tools, Claude Code, evaluation and safety. The habits that make Claude-based systems reliable, as a checklist you can run against your own setup.</description>
      <category>Checklist</category>
      <category>prompting</category><category>agents</category><category>evaluation</category><category>guardrails</category>
    </item>
    <item>
      <title>Pattern: the outbox for agent actions</title>
      <link>https://stackunseen.com/journal/outbox-pattern-for-agent-actions</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/outbox-pattern-for-agent-actions</guid>
      <pubDate>Mon, 28 Sep 2026 00:00:00 GMT</pubDate>
      <description>Agents that write to systems of record need the same transactional outbox that event-driven services use. This entry covers the shape and the trade-offs.</description>
      <category>Architecture pattern</category>
      <category>agents</category><category>integrations</category><category>architecture</category><category>reliability</category>
    </item>
    <item>
      <title>Adding a second agent does not add intelligence, it adds a contract</title>
      <link>https://stackunseen.com/journal/guide-agent-coordination-contracts</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/guide-agent-coordination-contracts</guid>
      <pubDate>Sun, 27 Sep 2026 00:00:00 GMT</pubDate>
      <description>Six posts on multi-agent systems, and the failures were never inside an agent. They were between two of them.</description>
      <category>Deep dive</category>
      <category>multi-agent</category><category>protocols</category><category>memory</category>
    </item>
    <item>
      <title>Designing a Multi-Agent System for Real-World Use</title>
      <link>https://stackunseen.com/journal/designing-a-multi-agent-system</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/designing-a-multi-agent-system</guid>
      <pubDate>Sat, 26 Sep 2026 00:00:00 GMT</pubDate>
      <description>How to design and run a multi-agent system in production, using the supervisor pattern, with its trade-offs and where human approval belongs.</description>
      <category>Architecture pattern</category>
      <category>agents</category><category>orchestration</category><category>architecture</category>
    </item>
    <item>
      <title>When to bring in a compliance review</title>
      <link>https://stackunseen.com/journal/when-to-bring-in-a-compliance-review</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/when-to-bring-in-a-compliance-review</guid>
      <pubDate>Thu, 24 Sep 2026 00:00:00 GMT</pubDate>
      <description>The changes that should pull legal, privacy or compliance into an AI project early, and what to have ready when you do. Not legal advice; a way to ask at the right time.</description>
      <category>Checklist</category>
      <category>governance</category><category>security</category><category>approvals</category>
    </item>
    <item>
      <title>Security review checklist for an AI feature</title>
      <link>https://stackunseen.com/journal/security-review-checklist-for-ai-features</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/security-review-checklist-for-ai-features</guid>
      <pubDate>Tue, 22 Sep 2026 00:00:00 GMT</pubDate>
      <description>What to check before an assistant, RAG app or agent goes in front of real users. Grouped by area, ticked off locally; progress stays in your browser.</description>
      <category>Checklist</category>
      <category>security</category><category>identity-and-access</category><category>prompt-injection</category><category>guardrails</category>
    </item>
    <item>
      <title>The retrieval cache that served stale policies</title>
      <link>https://stackunseen.com/journal/the-retrieval-cache-that-served-stale-policies</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/the-retrieval-cache-that-served-stale-policies</guid>
      <pubDate>Sun, 20 Sep 2026 00:00:00 GMT</pubDate>
      <description>A well-meaning cache in front of retrieval kept answering from last quarter's HR policy for eleven days.</description>
      <category>Failure story</category>
      <category>rag</category><category>operations</category><category>caching</category>
    </item>
    <item>
      <title>Six ways to wire agents together, and the same three things break every time</title>
      <link>https://stackunseen.com/journal/guide-multi-agent-design-patterns</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/guide-multi-agent-design-patterns</guid>
      <pubDate>Sat, 19 Sep 2026 00:00:00 GMT</pubDate>
      <description>The topology gets all the design attention. Ownership, termination and traceability are what decide whether it survives contact with production.</description>
      <category>Deep dive</category>
      <category>multi-agent</category><category>pipelines</category><category>protocols</category>
    </item>
    <item>
      <title>Nothing broke, and the system is still getting worse</title>
      <link>https://stackunseen.com/journal/ops-drift-and-operating-cadence</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/ops-drift-and-operating-cadence</guid>
      <pubDate>Fri, 18 Sep 2026 00:00:00 GMT</pubDate>
      <description>Drift is quiet by construction. SLOs and a standing review are what make it audible.</description>
      <category>Explainer</category>
      <category>drift</category><category>operations</category>
    </item>
    <item>
      <title>You cannot roll back a prompt you never versioned</title>
      <link>https://stackunseen.com/journal/ops-incidents-runbooks-and-rollbacks</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/ops-incidents-runbooks-and-rollbacks</guid>
      <pubDate>Thu, 17 Sep 2026 00:00:00 GMT</pubDate>
      <description>AI failures cross model, prompt, retrieval, tool and data boundaries. The runbook and the rollback controls have to as well.</description>
      <category>Explainer</category>
      <category>incidents</category><category>operations</category>
    </item>
    <item>
      <title>Real users ask questions your test set never imagined</title>
      <link>https://stackunseen.com/journal/ops-production-eval-signals</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/ops-production-eval-signals</guid>
      <pubDate>Wed, 16 Sep 2026 00:00:00 GMT</pubDate>
      <description>Offline evals cover the questions you thought of. Production tells you the ones you did not.</description>
      <category>Explainer</category>
      <category>evaluation</category><category>observability</category>
    </item>
    <item>
      <title>Your AI feature has unit economics whether you measured them or not</title>
      <link>https://stackunseen.com/journal/ops-cost-latency-and-capacity</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/ops-cost-latency-and-capacity</guid>
      <pubDate>Tue, 15 Sep 2026 00:00:00 GMT</pubDate>
      <description>Cost and latency belong to the workflow, not the model call. Budget them before rollout, not after the invoice.</description>
      <category>Explainer</category>
      <category>cost</category><category>capacity</category>
    </item>
    <item>
      <title>Logging the answer tells you almost nothing</title>
      <link>https://stackunseen.com/journal/ops-traces-logs-and-spans</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/ops-traces-logs-and-spans</guid>
      <pubDate>Mon, 14 Sep 2026 00:00:00 GMT</pubDate>
      <description>A trace that links intent, prompt, retrieval, tools and output is the only thing that makes an AI failure debuggable.</description>
      <category>Explainer</category>
      <category>tracing</category><category>observability</category>
    </item>
    <item>
      <title>Most agent controls do not actually control anything</title>
      <link>https://stackunseen.com/journal/guide-agent-control-and-supervision</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/guide-agent-control-and-supervision</guid>
      <pubDate>Sat, 12 Sep 2026 00:00:00 GMT</pubDate>
      <description>Six days of notes on supervising autonomous systems, and the same failure shape kept turning up: the control exists, it is documented, and nothing in the running system is bound by it.</description>
      <category>Deep dive</category>
      <category>multi-agent</category><category>human-review</category><category>agents</category>
    </item>
    <item>
      <title>The architecture review that happens before launch, not after the incident</title>
      <link>https://stackunseen.com/journal/arch-reference-architecture-and-readiness</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/arch-reference-architecture-and-readiness</guid>
      <pubDate>Fri, 11 Sep 2026 00:00:00 GMT</pubDate>
      <description>Quality, safety, cost, latency, resilience and operations have to be visible on one diagram, at the same time.</description>
      <category>Article</category>
      <category>architecture</category><category>readiness</category>
    </item>
    <item>
      <title>An approve button is not human oversight</title>
      <link>https://stackunseen.com/journal/arch-approvals-events-and-tenant-boundaries</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/arch-approvals-events-and-tenant-boundaries</guid>
      <pubDate>Thu, 10 Sep 2026 00:00:00 GMT</pubDate>
      <description>Approvals, events and tenant boundaries are architecture. Bolt them on as UI and they become theatre.</description>
      <category>Explainer</category>
      <category>approvals</category><category>multi-tenancy</category>
    </item>
    <item>
      <title>You are already building a control plane, badly</title>
      <link>https://stackunseen.com/journal/arch-control-plane-and-data-plane</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/arch-control-plane-and-data-plane</guid>
      <pubDate>Wed, 09 Sep 2026 00:00:00 GMT</pubDate>
      <description>Policy, memory, evals and rollout get rebuilt inside every feature until someone names the layer they belong to.</description>
      <category>Explainer</category>
      <category>architecture</category>
    </item>
    <item>
      <title>Most agents are workflows wearing a costume</title>
      <link>https://stackunseen.com/journal/arch-choosing-the-right-ai-pattern</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/arch-choosing-the-right-ai-pattern</guid>
      <pubDate>Tue, 08 Sep 2026 00:00:00 GMT</pubDate>
      <description>Chatbot, workflow, agent, RAG. Pick the wrong one and you spend a quarter debugging autonomy nobody asked for.</description>
      <category>Comparison</category>
      <category>architecture</category><category>patterns</category>
    </item>
    <item>
      <title>Every model call should go through something you own</title>
      <link>https://stackunseen.com/journal/arch-model-gateway-and-provider-strategy</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/arch-model-gateway-and-provider-strategy</guid>
      <pubDate>Mon, 07 Sep 2026 00:00:00 GMT</pubDate>
      <description>Routing, budgets, fallbacks and policy need one place to live. Scattered SDK calls give you none of them.</description>
      <category>Explainer</category>
      <category>gateway</category><category>architecture</category>
    </item>
    <item>
      <title>The hard part of building an agent is not making it act. It is making it stop.</title>
      <link>https://stackunseen.com/journal/guide-agentic-ai-foundations</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/guide-agentic-ai-foundations</guid>
      <pubDate>Sat, 05 Sep 2026 00:00:00 GMT</pubDate>
      <description>Six days of notes on agentic systems, and not one of the fixes that actually worked lived inside the model.</description>
      <category>Deep dive</category>
      <category>agents</category><category>guardrails</category><category>planning</category>
    </item>
    <item>
      <title>A wrong answer tells you nothing about which agent was wrong</title>
      <link>https://stackunseen.com/journal/multiagent-traces-evals-and-recovery</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/multiagent-traces-evals-and-recovery</guid>
      <pubDate>Fri, 04 Sep 2026 00:00:00 GMT</pubDate>
      <description>Multi-agent systems need trace timelines, trajectory evals and designed recovery, because the final output hides everything that produced it.</description>
      <category>Explainer</category>
      <category>multi-agent</category><category>tracing</category>
    </item>
    <item>
      <title>An agent that stops when it feels finished has no stop condition</title>
      <link>https://stackunseen.com/journal/multiagent-permissions-budgets-stop-conditions</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/multiagent-permissions-budgets-stop-conditions</guid>
      <pubDate>Thu, 03 Sep 2026 00:00:00 GMT</pubDate>
      <description>Permissions belong to roles, budgets belong to runs, and termination has to be something a machine can check.</description>
      <category>Explainer</category>
      <category>multi-agent</category><category>budgets</category>
    </item>
    <item>
      <title>Every multi-agent pattern is an answer to one question: who controls the run</title>
      <link>https://stackunseen.com/journal/multiagent-who-controls-the-run</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/multiagent-who-controls-the-run</guid>
      <pubDate>Wed, 02 Sep 2026 00:00:00 GMT</pubDate>
      <description>Supervisor, orchestrator-worker, planner-executor, pipeline and judge differ mainly in where control lives and when it returns.</description>
      <category>Explainer</category>
      <category>multi-agent</category><category>orchestration</category>
    </item>
    <item>
      <title>Write role contracts, not agent personalities</title>
      <link>https://stackunseen.com/journal/multiagent-role-contracts-and-handoffs</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/multiagent-role-contracts-and-handoffs</guid>
      <pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate>
      <description>&quot;You are a meticulous senior researcher&quot; is a costume. Inputs, outputs, permissions and done criteria are a contract.</description>
      <category>Explainer</category>
      <category>multi-agent</category><category>handoffs</category>
    </item>
    <item>
      <title>Most teams reaching for multiple agents do not need them</title>
      <link>https://stackunseen.com/journal/multiagent-when-multiple-agents-actually-help</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/multiagent-when-multiple-agents-actually-help</guid>
      <pubDate>Mon, 31 Aug 2026 00:00:00 GMT</pubDate>
      <description>Separation of work has to buy quality, control or throughput. Usually it buys a harder debugging problem.</description>
      <category>Comparison</category>
      <category>multi-agent</category>
    </item>
    <item>
      <title>Production RAG does not fail loudly, and that is the whole problem</title>
      <link>https://stackunseen.com/journal/guide-production-rag-operations</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/guide-production-rag-operations</guid>
      <pubDate>Sat, 29 Aug 2026 00:00:00 GMT</pubDate>
      <description>Six days of notes on operating retrieval systems after launch, where nearly every real failure arrives dressed as a good answer.</description>
      <category>Deep dive</category>
      <category>rag</category><category>structured-data</category><category>multimodal</category>
    </item>
    <item>
      <title>You cannot instruct a model into data protection</title>
      <link>https://stackunseen.com/journal/safety-data-exfiltration-and-evidence</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/safety-data-exfiltration-and-evidence</guid>
      <pubDate>Fri, 28 Aug 2026 00:00:00 GMT</pubDate>
      <description>Minimise what enters context, watch the paths data can leave by, and keep the evidence an incident will demand.</description>
      <category>Explainer</category>
      <category>security</category><category>audit</category>
    </item>
    <item>
      <title>Blast radius is a design parameter</title>
      <link>https://stackunseen.com/journal/safety-blast-radius-and-containment</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/safety-blast-radius-and-containment</guid>
      <pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate>
      <description>Narrow tools, staged writes, sandboxes and a tested kill switch. Containment is built before it is needed.</description>
      <category>Explainer</category>
      <category>security</category><category>containment</category>
    </item>
    <item>
      <title>Untrusted text does not get to give orders</title>
      <link>https://stackunseen.com/journal/safety-untrusted-input-validation</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/safety-untrusted-input-validation</guid>
      <pubDate>Wed, 26 Aug 2026 00:00:00 GMT</pubDate>
      <description>Injection defence is layered validation and clear authority, not a better-worded system prompt.</description>
      <category>Explainer</category>
      <category>prompt-injection</category><category>security</category>
    </item>
    <item>
      <title>Your agent should not be a superuser</title>
      <link>https://stackunseen.com/journal/safety-agent-identity-and-scope</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/safety-agent-identity-and-scope</guid>
      <pubDate>Tue, 25 Aug 2026 00:00:00 GMT</pubDate>
      <description>Bind every AI action to a user, a service and a run, then scope the verb rather than only the data.</description>
      <category>Explainer</category>
      <category>identity-and-access</category>
    </item>
    <item>
      <title>Safety is an architecture decision, not a moderation setting</title>
      <link>https://stackunseen.com/journal/safety-architecture-not-moderation</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/safety-architecture-not-moderation</guid>
      <pubDate>Mon, 24 Aug 2026 00:00:00 GMT</pubDate>
      <description>Assets, actors, abuse cases, controls, owners, evidence. Six things a team can actually write down before launch.</description>
      <category>Explainer</category>
      <category>security</category><category>guardrails</category>
    </item>
    <item>
      <title>Most RAG failures are faithful answers to bad evidence</title>
      <link>https://stackunseen.com/journal/guide-evidence-driven-rag-patterns</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/guide-evidence-driven-rag-patterns</guid>
      <pubDate>Sat, 22 Aug 2026 00:00:00 GMT</pubDate>
      <description>Six posts on reranking, citations and decomposition, and the uncomfortable thing they turn out to have in common.</description>
      <category>Deep dive</category>
      <category>rag</category><category>retrieval</category><category>reranking</category>
    </item>
    <item>
      <title>Launch day is when evaluation starts</title>
      <link>https://stackunseen.com/journal/evals-online-evaluation-after-launch</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/evals-online-evaluation-after-launch</guid>
      <pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate>
      <description>Offline scores expire on contact with real users. Sampling, groundedness, drift and outcomes are the parts that keep paying.</description>
      <category>Explainer</category>
      <category>evaluation</category>
    </item>
    <item>
      <title>If it runs after the deploy, it is a postmortem</title>
      <link>https://stackunseen.com/journal/evals-regression-gates-in-the-release-path</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/evals-regression-gates-in-the-release-path</guid>
      <pubDate>Thu, 20 Aug 2026 00:00:00 GMT</pubDate>
      <description>Regression gates only work when they sit in the release path, with a threshold, a slice and an owner.</description>
      <category>Explainer</category>
      <category>evaluation</category><category>release</category>
    </item>
    <item>
      <title>A judge that cannot name the failure is not a judge</title>
      <link>https://stackunseen.com/journal/evals-judges-rubrics-and-calibration</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/evals-judges-rubrics-and-calibration</guid>
      <pubDate>Wed, 19 Aug 2026 00:00:00 GMT</pubDate>
      <description>Deterministic checks first, one rubric per criterion, and a calibration loop against human review.</description>
      <category>Explainer</category>
      <category>evaluation</category>
    </item>
    <item>
      <title>Your golden set is a product, not a spreadsheet</title>
      <link>https://stackunseen.com/journal/evals-golden-set-is-a-product</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/evals-golden-set-is-a-product</guid>
      <pubDate>Tue, 18 Aug 2026 00:00:00 GMT</pubDate>
      <description>Labels, edge cases, versions, an owner. Skip those and the number moves without anyone knowing why.</description>
      <category>Explainer</category>
      <category>evaluation</category><category>golden-set</category>
    </item>
    <item>
      <title>An eval that cannot block a release is just a report</title>
      <link>https://stackunseen.com/journal/evals-what-an-eval-is-allowed-to-block</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/evals-what-an-eval-is-allowed-to-block</guid>
      <pubDate>Mon, 17 Aug 2026 00:00:00 GMT</pubDate>
      <description>Decide which decision the number is allowed to stop, then build the harness around that.</description>
      <category>Explainer</category>
      <category>evaluation</category>
    </item>
    <item>
      <title>No model can reason about evidence your retriever never returned</title>
      <link>https://stackunseen.com/journal/guide-retrieval-foundations</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/guide-retrieval-foundations</guid>
      <pubDate>Sat, 15 Aug 2026 00:00:00 GMT</pubDate>
      <description>Six days of retrieval notes, and not one of the failures announced itself.</description>
      <category>Deep dive</category>
      <category>retrieval</category><category>rag</category><category>search</category>
    </item>
    <item>
      <title>You cannot run RAG on user complaints</title>
      <link>https://stackunseen.com/journal/rag-citations-evals-and-operations</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/rag-citations-evals-and-operations</guid>
      <pubDate>Fri, 14 Aug 2026 00:00:00 GMT</pubDate>
      <description>Groundedness scores, per-claim citations and an operations dashboard are what tell you quality is drifting before your users do.</description>
      <category>Explainer</category>
      <category>citations</category><category>evaluation</category>
    </item>
    <item>
      <title>Chunk for the question, not for the token limit</title>
      <link>https://stackunseen.com/journal/rag-chunking-and-context-packaging</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/rag-chunking-and-context-packaging</guid>
      <pubDate>Thu, 13 Aug 2026 00:00:00 GMT</pubDate>
      <description>Chunk size is a decision about what a complete answer looks like. Reranking, context budgets and refusal thresholds finish the job.</description>
      <category>Explainer</category>
      <category>chunking</category><category>rag</category>
    </item>
    <item>
      <title>Test retrieval on its own or you will tune the wrong thing</title>
      <link>https://stackunseen.com/journal/rag-retrieval-is-its-own-subsystem</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/rag-retrieval-is-its-own-subsystem</guid>
      <pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate>
      <description>Most &quot;the model is wrong&quot; problems are recall problems. Hybrid search, query rewriting and versioned indexes are where they get fixed.</description>
      <category>Explainer</category>
      <category>retrieval</category><category>rag</category>
    </item>
    <item>
      <title>Ingestion is a data product, not a loader script</title>
      <link>https://stackunseen.com/journal/rag-ingestion-is-a-data-product</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/rag-ingestion-is-a-data-product</guid>
      <pubDate>Tue, 11 Aug 2026 00:00:00 GMT</pubDate>
      <description>Parsing sets your quality ceiling, failures need a destination, and the update and delete paths are the half nobody tests.</description>
      <category>Explainer</category>
      <category>ingestion</category><category>rag</category>
    </item>
    <item>
      <title>Your vector index does not know who is allowed to read it</title>
      <link>https://stackunseen.com/journal/rag-trusted-sources-and-freshness</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/rag-trusted-sources-and-freshness</guid>
      <pubDate>Mon, 10 Aug 2026 00:00:00 GMT</pubDate>
      <description>Embeddings carry no access control lists. Ownership, permissions, freshness and lineage have to be designed before anything is indexed.</description>
      <category>Explainer</category>
      <category>freshness</category><category>rag</category>
    </item>
    <item>
      <title>Everything that makes an AI demo impressive is a constraint you have not added yet</title>
      <link>https://stackunseen.com/journal/guide-from-llm-demo-to-ai-product</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/guide-from-llm-demo-to-ai-product</guid>
      <pubDate>Sat, 08 Aug 2026 00:00:00 GMT</pubDate>
      <description>Six posts on the distance between a system that answers well once and a system people are willing to depend on.</description>
      <category>Deep dive</category>
      <category>rag</category><category>grounding</category><category>hallucination</category>
    </item>
    <item>
      <title>One MCP server is an integration. Thirty is why you need a gateway</title>
      <link>https://stackunseen.com/journal/mcp-gateway-field-notes</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/mcp-gateway-field-notes</guid>
      <pubDate>Thu, 06 Aug 2026 00:00:00 GMT</pubDate>
      <description>Notes from building an MCP gateway. The protocol standardised how agents call tools, then stayed silent about credentials, context budgets and who may call what. That silence gets expensive as the servers multiply.</description>
      <category>Deep dive</category>
      <category>mcp</category><category>gateway</category><category>security</category><category>identity-and-access</category>
    </item>
    <item>
      <title>The 3am agent run that nobody is watching</title>
      <link>https://stackunseen.com/journal/infra-scheduled-runs-and-operations</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/infra-scheduled-runs-and-operations</guid>
      <pubDate>Wed, 05 Aug 2026 00:00:00 GMT</pubDate>
      <description>Scheduled agent work fails quietly by default. Traces, release gates and a runbook are what make it fail loudly.</description>
      <category>Explainer</category>
      <category>scheduling</category><category>operations</category>
    </item>
    <item>
      <title>A retried agent job is a second chance to send the same email</title>
      <link>https://stackunseen.com/journal/infra-retries-and-idempotency</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/infra-retries-and-idempotency</guid>
      <pubDate>Tue, 04 Aug 2026 00:00:00 GMT</pubDate>
      <description>Retries are the first thing teams switch on and the last thing they design for.</description>
      <category>Explainer</category>
      <category>retries</category><category>idempotency</category>
    </item>
    <item>
      <title>Four kinds of agent state, and only one of them is memory</title>
      <link>https://stackunseen.com/journal/infra-four-kinds-of-agent-state</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/infra-four-kinds-of-agent-state</guid>
      <pubDate>Mon, 03 Aug 2026 00:00:00 GMT</pubDate>
      <description>Conversation, task, artifacts and product data end up in one store. Then a deletion request arrives and applies to all of it.</description>
      <category>Explainer</category>
      <category>state</category><category>memory</category>
    </item>
    <item>
      <title>Your AI system's biggest problem is probably not the model</title>
      <link>https://stackunseen.com/journal/guide-ai-systems-basics</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/guide-ai-systems-basics</guid>
      <pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate>
      <description>Six days of production AI notes, and the model choice never once turned out to be the thing that broke.</description>
      <category>Deep dive</category>
      <category>models</category><category>architecture</category><category>generation</category>
    </item>
    <item>
      <title>Your tool registry is an access control list wearing a different name</title>
      <link>https://stackunseen.com/journal/infra-tool-registry-is-access-control</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/infra-tool-registry-is-access-control</guid>
      <pubDate>Fri, 31 Jul 2026 00:00:00 GMT</pubDate>
      <description>A list of tools is documentation. A registry decides who may act, as whom, and leaves proof behind.</description>
      <category>Explainer</category>
      <category>tools</category><category>identity-and-access</category>
    </item>
    <item>
      <title>Autonomy is a runtime decision, not a model decision</title>
      <link>https://stackunseen.com/journal/infra-autonomy-is-a-runtime-decision</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/infra-autonomy-is-a-runtime-decision</guid>
      <pubDate>Thu, 30 Jul 2026 00:00:00 GMT</pubDate>
      <description>Agents get more capable by default. They only get bounded on purpose.</description>
      <category>Explainer</category>
      <category>autonomy</category><category>runtime</category>
    </item>
    <item>
      <title>When MCP earns its overhead, and when a direct API is better</title>
      <link>https://stackunseen.com/journal/mcp-versus-direct-apis</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/mcp-versus-direct-apis</guid>
      <pubDate>Tue, 28 Jul 2026 00:00:00 GMT</pubDate>
      <description>One question settles most of these arguments: will more than one AI application need this capability?</description>
      <category>Comparison</category>
      <category>mcp</category><category>apis</category>
    </item>
    <item>
      <title>A clean connector does not make a reliable agent</title>
      <link>https://stackunseen.com/journal/mcp-connector-versus-agent</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/mcp-connector-versus-agent</guid>
      <pubDate>Mon, 27 Jul 2026 00:00:00 GMT</pubDate>
      <description>Connector correctness and agent behaviour are separate problems. Most teams only test the first one.</description>
      <category>Explainer</category>
      <category>mcp</category>
    </item>
    <item>
      <title>Tools, resources and prompts are not interchangeable</title>
      <link>https://stackunseen.com/journal/mcp-tools-resources-and-prompts</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/mcp-tools-resources-and-prompts</guid>
      <pubDate>Sat, 25 Jul 2026 00:00:00 GMT</pubDate>
      <description>Choosing the wrong primitive hands the model authority a person was supposed to hold.</description>
      <category>Explainer</category>
      <category>mcp</category>
    </item>
    <item>
      <title>Where your MCP server runs is a production decision</title>
      <link>https://stackunseen.com/journal/mcp-where-your-server-runs</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/mcp-where-your-server-runs</guid>
      <pubDate>Fri, 24 Jul 2026 00:00:00 GMT</pubDate>
      <description>Transport looks like a technical detail during the demo. It is really a choice about ownership, reachability and trust.</description>
      <category>Explainer</category>
      <category>mcp</category><category>deployment</category>
    </item>
    <item>
      <title>What the MCP protocol actually standardises</title>
      <link>https://stackunseen.com/journal/mcp-what-the-protocol-standardises</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/mcp-what-the-protocol-standardises</guid>
      <pubDate>Thu, 23 Jul 2026 00:00:00 GMT</pubDate>
      <description>It settles how capabilities are described and discovered. Every decision about whether to trust them is still yours.</description>
      <category>Explainer</category>
      <category>mcp</category>
    </item>
    <item>
      <title>From model demos to mission-ready AI systems</title>
      <link>https://stackunseen.com/journal/from-model-demos-to-mission-ready-ai-systems</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/from-model-demos-to-mission-ready-ai-systems</guid>
      <pubDate>Mon, 20 Jul 2026 00:00:00 GMT</pubDate>
      <description>A practical map for moving from impressive AI demos to systems people can trust.</description>
      <category>Article</category>
      <category>architecture</category>
    </item>
    <item>
      <title>Why governance belongs in the architecture</title>
      <link>https://stackunseen.com/journal/why-governance-belongs-in-the-architecture</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-governance-belongs-in-the-architecture</guid>
      <pubDate>Sun, 19 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why governance belongs in the architecture, not in a document after launch.</description>
      <category>Explainer</category>
      <category>governance</category>
    </item>
    <item>
      <title>Why production AI is coordinated infrastructure</title>
      <link>https://stackunseen.com/journal/why-production-ai-is-coordinated-infrastructure</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-production-ai-is-coordinated-infrastructure</guid>
      <pubDate>Sun, 19 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why production AI is the coordination of data, models, tools, controls, evaluation, and people.</description>
      <category>Article</category>
      <category>architecture</category>
    </item>
    <item>
      <title>Why observability is not optional for AI systems</title>
      <link>https://stackunseen.com/journal/why-observability-is-not-optional-for-ai-systems</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-observability-is-not-optional-for-ai-systems</guid>
      <pubDate>Sat, 18 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why production AI needs traces across prompts, context, retrieval, tools, costs, and decisions.</description>
      <category>Explainer</category>
      <category>observability</category><category>tracing</category>
    </item>
    <item>
      <title>Why evaluating the model is not enough</title>
      <link>https://stackunseen.com/journal/why-evaluating-the-model-is-not-enough</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-evaluating-the-model-is-not-enough</guid>
      <pubDate>Sat, 18 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why evaluating agents means measuring the whole workflow users experience.</description>
      <category>Explainer</category>
      <category>evaluation</category>
    </item>
    <item>
      <title>Why retrieved content must stay untrusted</title>
      <link>https://stackunseen.com/journal/why-retrieved-content-must-stay-untrusted</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-retrieved-content-must-stay-untrusted</guid>
      <pubDate>Sat, 18 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why retrieved content must stay data, not become authority over the system.</description>
      <category>Explainer</category>
      <category>prompt-injection</category><category>security</category>
    </item>
    <item>
      <title>Why multi-agent systems need decision rules</title>
      <link>https://stackunseen.com/journal/why-multi-agent-systems-need-decision-rules</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-multi-agent-systems-need-decision-rules</guid>
      <pubDate>Fri, 17 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why agents need arbitration rules before disagreement reaches the final answer.</description>
      <category>Explainer</category>
      <category>governance</category><category>multi-agent</category>
    </item>
    <item>
      <title>Why handoffs need contracts</title>
      <link>https://stackunseen.com/journal/why-handoffs-need-contracts</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-handoffs-need-contracts</guid>
      <pubDate>Wed, 15 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why reliable handoffs need structured state, evidence, open questions, and ownership.</description>
      <category>Explainer</category>
      <category>handoffs</category><category>contracts</category>
    </item>
    <item>
      <title>Why shared memory becomes a permission problem</title>
      <link>https://stackunseen.com/journal/why-shared-memory-becomes-a-permission-problem</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-shared-memory-becomes-a-permission-problem</guid>
      <pubDate>Tue, 14 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why shared memory becomes an access-control problem in multi-agent systems.</description>
      <category>Explainer</category>
      <category>memory</category><category>identity-and-access</category>
    </item>
    <item>
      <title>Why agents need a shared language</title>
      <link>https://stackunseen.com/journal/why-agents-need-a-shared-language</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-agents-need-a-shared-language</guid>
      <pubDate>Mon, 13 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why agents need shared message contracts before they can collaborate reliably.</description>
      <category>Explainer</category>
      <category>multi-agent</category><category>protocols</category>
    </item>
    <item>
      <title>Why role design matters more than agent count</title>
      <link>https://stackunseen.com/journal/why-role-design-matters-more-than-agent-count</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-role-design-matters-more-than-agent-count</guid>
      <pubDate>Sun, 12 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why agent roles need clear purpose, authority, tools, memory, and evaluation.</description>
      <category>Explainer</category>
      <category>multi-agent</category>
    </item>
    <item>
      <title>When decentralized agent behavior makes sense</title>
      <link>https://stackunseen.com/journal/when-decentralized-agent-behavior-makes-sense</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/when-decentralized-agent-behavior-makes-sense</guid>
      <pubDate>Sat, 11 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why decentralized agent behavior needs convergence rules before it earns production trust.</description>
      <category>Explainer</category>
      <category>multi-agent</category>
    </item>
    <item>
      <title>Why AI debate needs a reliable judge</title>
      <link>https://stackunseen.com/journal/why-ai-debate-needs-a-reliable-judge</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-ai-debate-needs-a-reliable-judge</guid>
      <pubDate>Sat, 11 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why AI debate is only useful when the judge has reliable criteria.</description>
      <category>Explainer</category>
      <category>multi-agent</category><category>evaluation</category>
    </item>
    <item>
      <title>Why shared state needs structure and attribution</title>
      <link>https://stackunseen.com/journal/why-shared-state-needs-structure-and-attribution</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-shared-state-needs-structure-and-attribution</guid>
      <pubDate>Wed, 08 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why shared state needs structure, attribution, and versioning to stay debuggable.</description>
      <category>Explainer</category>
      <category>multi-agent</category><category>shared-state</category>
    </item>
    <item>
      <title>Why peer agents need protocols, not vibes</title>
      <link>https://stackunseen.com/journal/why-peer-agents-need-protocols-not-vibes</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-peer-agents-need-protocols-not-vibes</guid>
      <pubDate>Mon, 06 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why peer agents need communication rules, round limits, and a final decision owner.</description>
      <category>Explainer</category>
      <category>multi-agent</category><category>protocols</category>
    </item>
    <item>
      <title>Why dynamic dispatch needs ownership rules</title>
      <link>https://stackunseen.com/journal/why-dynamic-dispatch-needs-ownership-rules</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-dynamic-dispatch-needs-ownership-rules</guid>
      <pubDate>Sun, 05 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why dynamic dispatch needs visible routing decisions and clear final ownership.</description>
      <category>Explainer</category>
      <category>multi-agent</category>
    </item>
    <item>
      <title>Why pipelines make AI handoffs inspectable</title>
      <link>https://stackunseen.com/journal/why-pipelines-make-ai-handoffs-inspectable</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-pipelines-make-ai-handoffs-inspectable</guid>
      <pubDate>Sun, 05 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why predictable AI workflows often benefit from pipeline structure before agent autonomy.</description>
      <category>Explainer</category>
      <category>multi-agent</category><category>pipelines</category>
    </item>
    <item>
      <title>Why supervisor agents help and where they bottleneck</title>
      <link>https://stackunseen.com/journal/why-supervisor-agents-help-and-where-they-bottleneck</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-supervisor-agents-help-and-where-they-bottleneck</guid>
      <pubDate>Sat, 04 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why supervisor agents help only when routing, review, and authority are clear.</description>
      <category>Explainer</category>
      <category>multi-agent</category>
    </item>
    <item>
      <title>When hierarchy helps multi-agent work scale</title>
      <link>https://stackunseen.com/journal/when-hierarchy-helps-multi-agent-work-scale</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/when-hierarchy-helps-multi-agent-work-scale</guid>
      <pubDate>Sat, 04 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why hierarchy helps when work has real layers of authority and responsibility.</description>
      <category>Explainer</category>
      <category>multi-agent</category>
    </item>
    <item>
      <title>Why one giant agent is rarely the cleanest design</title>
      <link>https://stackunseen.com/journal/why-one-giant-agent-is-rarely-the-cleanest-design</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-one-giant-agent-is-rarely-the-cleanest-design</guid>
      <pubDate>Fri, 03 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why multi-agent systems should reduce complexity through specialization, not add agents for novelty.</description>
      <category>Explainer</category>
      <category>multi-agent</category>
    </item>
    <item>
      <title>Why autonomy needs budgets</title>
      <link>https://stackunseen.com/journal/why-autonomy-needs-budgets</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-autonomy-needs-budgets</guid>
      <pubDate>Wed, 01 Jul 2026 00:00:00 GMT</pubDate>
      <description>Why autonomous systems need limits on cost, time, actions, retries, and risk.</description>
      <category>Explainer</category>
      <category>budgets</category><category>cost</category>
    </item>
    <item>
      <title>Why production agents need recovery design</title>
      <link>https://stackunseen.com/journal/why-production-agents-need-recovery-design</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-production-agents-need-recovery-design</guid>
      <pubDate>Mon, 29 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why retries, fallbacks, checkpoints, rollback, and escalation belong in the first design.</description>
      <category>Explainer</category>
      <category>recovery</category><category>reliability</category>
    </item>
    <item>
      <title>Why guardrails should enable safe action</title>
      <link>https://stackunseen.com/journal/why-guardrails-should-enable-safe-action</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-guardrails-should-enable-safe-action</guid>
      <pubDate>Sun, 28 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why guardrails should shape safe progress, not just block uncertain behavior.</description>
      <category>Explainer</category>
      <category>guardrails</category>
    </item>
    <item>
      <title>Where human review actually belongs in AI workflows</title>
      <link>https://stackunseen.com/journal/where-human-review-actually-belongs-in-ai-workflows</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/where-human-review-actually-belongs-in-ai-workflows</guid>
      <pubDate>Sat, 27 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why human review needs timing, authority, and context to be useful.</description>
      <category>Explainer</category>
      <category>human-review</category><category>agents</category>
    </item>
    <item>
      <title>Why agent memory is not one database</title>
      <link>https://stackunseen.com/journal/why-agent-memory-is-not-one-database</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-agent-memory-is-not-one-database</guid>
      <pubDate>Sat, 27 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why useful agent memory has to be separated, scoped, and governed.</description>
      <category>Explainer</category>
      <category>memory</category><category>agents</category>
    </item>
    <item>
      <title>Why reflection should change the next action</title>
      <link>https://stackunseen.com/journal/why-reflection-should-change-the-next-action</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-reflection-should-change-the-next-action</guid>
      <pubDate>Sat, 27 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why reflection is valuable only when it decides whether to revise, retry, escalate, or stop.</description>
      <category>Explainer</category>
      <category>reflection</category><category>agents</category>
    </item>
    <item>
      <title>Why tool access is where AI becomes operational risk</title>
      <link>https://stackunseen.com/journal/why-tool-access-is-where-ai-becomes-operational-risk</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-tool-access-is-where-ai-becomes-operational-risk</guid>
      <pubDate>Wed, 24 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why giving AI tools means designing permissions, observability, rollback, and approval paths.</description>
      <category>Explainer</category>
      <category>tools</category><category>identity-and-access</category>
    </item>
    <item>
      <title>Why planning reduces wasted AI actions</title>
      <link>https://stackunseen.com/journal/why-planning-reduces-wasted-ai-actions</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-planning-reduces-wasted-ai-actions</guid>
      <pubDate>Tue, 23 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why planning matters when actions have cost, risk, or dependencies.</description>
      <category>Explainer</category>
      <category>planning</category><category>agents</category>
    </item>
    <item>
      <title>Why agent behavior is a loop, not a single prompt</title>
      <link>https://stackunseen.com/journal/why-agent-behavior-is-a-loop-not-a-single-prompt</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-agent-behavior-is-a-loop-not-a-single-prompt</guid>
      <pubDate>Mon, 22 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why reliable agent behavior comes from loop design, not a single clever prompt.</description>
      <category>Explainer</category>
      <category>agents</category>
    </item>
    <item>
      <title>Why agents need outcomes, boundaries, and stop conditions</title>
      <link>https://stackunseen.com/journal/why-agents-need-outcomes-boundaries-and-stop-conditions</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-agents-need-outcomes-boundaries-and-stop-conditions</guid>
      <pubDate>Sun, 21 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why agents need clear outcomes, boundaries, budgets, and stop conditions before autonomy.</description>
      <category>Explainer</category>
      <category>agents</category><category>guardrails</category>
    </item>
    <item>
      <title>Why RAG must be evaluated in parts</title>
      <link>https://stackunseen.com/journal/why-rag-must-be-evaluated-in-parts</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-rag-must-be-evaluated-in-parts</guid>
      <pubDate>Sun, 21 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why RAG evaluation has to measure retrieval, grounding, generation, and citations separately.</description>
      <category>Explainer</category>
      <category>evaluation</category><category>rag</category>
    </item>
    <item>
      <title>Why weak evidence should trigger recovery, not confidence</title>
      <link>https://stackunseen.com/journal/why-weak-evidence-should-trigger-recovery-not-confidence</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-weak-evidence-should-trigger-recovery-not-confidence</guid>
      <pubDate>Sat, 20 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why weak evidence should trigger recovery, not a fluent answer with false confidence.</description>
      <category>Explainer</category>
      <category>recovery</category><category>rag</category>
    </item>
    <item>
      <title>Why reflection only matters when tied to evidence</title>
      <link>https://stackunseen.com/journal/why-reflection-only-matters-when-tied-to-evidence</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-reflection-only-matters-when-tied-to-evidence</guid>
      <pubDate>Sat, 20 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why reflection only helps when it checks the answer against evidence and changes behavior.</description>
      <category>Explainer</category>
      <category>reflection</category><category>rag</category>
    </item>
    <item>
      <title>Why freshness is part of correctness</title>
      <link>https://stackunseen.com/journal/why-freshness-is-part-of-correctness</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-freshness-is-part-of-correctness</guid>
      <pubDate>Fri, 19 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why freshness becomes part of correctness when the world changes faster than your index.</description>
      <category>Explainer</category>
      <category>freshness</category><category>rag</category>
    </item>
    <item>
      <title>Why evidence is becoming multimodal</title>
      <link>https://stackunseen.com/journal/why-evidence-is-becoming-multimodal</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-evidence-is-becoming-multimodal</guid>
      <pubDate>Wed, 17 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why evidence quality changes when the source is visual, tabular, audio, or multimodal.</description>
      <category>Explainer</category>
      <category>multimodal</category><category>rag</category>
    </item>
    <item>
      <title>Why production knowledge often lives in systems, not PDFs</title>
      <link>https://stackunseen.com/journal/why-production-knowledge-often-lives-in-systems-not-pdfs</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-production-knowledge-often-lives-in-systems-not-pdfs</guid>
      <pubDate>Mon, 15 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why production knowledge often lives in systems of record, not only in documents.</description>
      <category>Explainer</category>
      <category>structured-data</category><category>rag</category>
    </item>
    <item>
      <title>Why some answers live in relationships, not documents</title>
      <link>https://stackunseen.com/journal/why-some-answers-live-in-relationships-not-documents</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-some-answers-live-in-relationships-not-documents</guid>
      <pubDate>Sun, 14 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why some answers depend on relationships that plain document search can miss.</description>
      <category>Explainer</category>
      <category>graph-rag</category><category>rag</category>
    </item>
    <item>
      <title>Why one user question may need many searches</title>
      <link>https://stackunseen.com/journal/why-one-user-question-may-need-many-searches</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-one-user-question-may-need-many-searches</guid>
      <pubDate>Sat, 13 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why one user question may need multiple searches before the evidence is good enough.</description>
      <category>Explainer</category>
      <category>retrieval</category><category>rag</category>
    </item>
    <item>
      <title>Why complex AI questions need decomposition</title>
      <link>https://stackunseen.com/journal/why-complex-ai-questions-need-decomposition</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-complex-ai-questions-need-decomposition</guid>
      <pubDate>Sat, 13 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why complex questions become more reliable when the system answers them in parts.</description>
      <category>Explainer</category>
      <category>planning</category><category>rag</category>
    </item>
    <item>
      <title>Why RAG needs both broad context and precise evidence</title>
      <link>https://stackunseen.com/journal/why-rag-needs-both-broad-context-and-precise-evidence</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-rag-needs-both-broad-context-and-precise-evidence</guid>
      <pubDate>Wed, 10 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why good RAG often needs both precise snippets and enough surrounding context.</description>
      <category>Explainer</category>
      <category>retrieval</category><category>rag</category>
    </item>
    <item>
      <title>Why every important AI claim needs provenance</title>
      <link>https://stackunseen.com/journal/why-every-important-ai-claim-needs-provenance</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-every-important-ai-claim-needs-provenance</guid>
      <pubDate>Mon, 08 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why important AI claims need receipts that users and teams can inspect.</description>
      <category>Explainer</category>
      <category>citations</category><category>provenance</category>
    </item>
    <item>
      <title>Why reranking is where retrieval becomes useful</title>
      <link>https://stackunseen.com/journal/why-reranking-is-where-retrieval-becomes-useful</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-reranking-is-where-retrieval-becomes-useful</guid>
      <pubDate>Sun, 07 Jun 2026 00:00:00 GMT</pubDate>
      <description>How reranking turns broad retrieval candidates into evidence the answer can depend on.</description>
      <category>Explainer</category>
      <category>reranking</category><category>retrieval</category>
    </item>
    <item>
      <title>Why relevant data can still be unauthorized data</title>
      <link>https://stackunseen.com/journal/why-relevant-data-can-still-be-unauthorized-data</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-relevant-data-can-still-be-unauthorized-data</guid>
      <pubDate>Sun, 07 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why relevant evidence is still wrong evidence if the user was not allowed to see it.</description>
      <category>Explainer</category>
      <category>identity-and-access</category><category>retrieval</category>
    </item>
    <item>
      <title>Why vector stores are infrastructure, not magic memory</title>
      <link>https://stackunseen.com/journal/why-vector-stores-are-infrastructure-not-magic-memory</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-vector-stores-are-infrastructure-not-magic-memory</guid>
      <pubDate>Sat, 06 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why vector databases need to be operated like infrastructure, not treated like magic memory.</description>
      <category>Explainer</category>
      <category>vector-database</category>
    </item>
    <item>
      <title>Why keyword search and semantic search both matter</title>
      <link>https://stackunseen.com/journal/why-keyword-search-and-semantic-search-both-matter</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-keyword-search-and-semantic-search-both-matter</guid>
      <pubDate>Sat, 06 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why exact words and semantic meaning both matter in production retrieval.</description>
      <category>Explainer</category>
      <category>search</category><category>retrieval</category>
    </item>
    <item>
      <title>Why hybrid search is often the practical default</title>
      <link>https://stackunseen.com/journal/why-hybrid-search-is-often-the-practical-default</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-hybrid-search-is-often-the-practical-default</guid>
      <pubDate>Sat, 06 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why combining retrieval signals is often more practical than betting on one search method.</description>
      <category>Explainer</category>
      <category>search</category><category>retrieval</category>
    </item>
    <item>
      <title>Why embeddings are useful and easy to overtrust</title>
      <link>https://stackunseen.com/journal/why-embeddings-are-useful-and-easy-to-overtrust</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-embeddings-are-useful-and-easy-to-overtrust</guid>
      <pubDate>Fri, 05 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why semantic similarity is useful, but not the same as correctness or trust.</description>
      <category>Explainer</category>
      <category>embeddings</category><category>rag</category>
    </item>
    <item>
      <title>Why chunk boundaries shape answer quality</title>
      <link>https://stackunseen.com/journal/why-chunk-boundaries-shape-answer-quality</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-chunk-boundaries-shape-answer-quality</guid>
      <pubDate>Wed, 03 Jun 2026 00:00:00 GMT</pubDate>
      <description>How chunk boundaries decide what evidence the system can actually retrieve.</description>
      <category>Explainer</category>
      <category>chunking</category><category>rag</category>
    </item>
    <item>
      <title>Why retrieval quality starts before search</title>
      <link>https://stackunseen.com/journal/why-retrieval-quality-starts-before-search</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-retrieval-quality-starts-before-search</guid>
      <pubDate>Tue, 02 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why answer quality starts when knowledge enters the system, not when the user asks.</description>
      <category>Explainer</category>
      <category>ingestion</category><category>rag</category>
    </item>
    <item>
      <title>Why RAG is an evidence design problem, not a buzzword</title>
      <link>https://stackunseen.com/journal/why-rag-is-an-evidence-design-problem-not-a-buzzword</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-rag-is-an-evidence-design-problem-not-a-buzzword</guid>
      <pubDate>Mon, 01 Jun 2026 00:00:00 GMT</pubDate>
      <description>Why RAG is really about trusted evidence, not just letting the model search.</description>
      <category>Explainer</category>
      <category>rag</category>
    </item>
    <item>
      <title>When to use a chatbot, workflow, or agent</title>
      <link>https://stackunseen.com/journal/when-to-use-a-chatbot-workflow-or-agent</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/when-to-use-a-chatbot-workflow-or-agent</guid>
      <pubDate>Sun, 31 May 2026 00:00:00 GMT</pubDate>
      <description>A practical way to choose between conversation, repeatable workflows, and bounded autonomy.</description>
      <category>Comparison</category>
      <category>agents</category><category>architecture</category>
    </item>
    <item>
      <title>Why production AI needs structured outputs</title>
      <link>https://stackunseen.com/journal/why-production-ai-needs-structured-outputs</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-production-ai-needs-structured-outputs</guid>
      <pubDate>Sat, 30 May 2026 00:00:00 GMT</pubDate>
      <description>How structured output turns model text into something software can safely use.</description>
      <category>Explainer</category>
      <category>structured-output</category>
    </item>
    <item>
      <title>Why a capable model is still not a product</title>
      <link>https://stackunseen.com/journal/why-a-capable-model-is-still-not-a-product</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-a-capable-model-is-still-not-a-product</guid>
      <pubDate>Sat, 30 May 2026 00:00:00 GMT</pubDate>
      <description>Why strong model capability still needs workflow, product, and ownership design.</description>
      <category>Article</category>
      <category>product</category>
    </item>
    <item>
      <title>Why confident AI answers still need evidence</title>
      <link>https://stackunseen.com/journal/why-confident-ai-answers-still-need-evidence</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-confident-ai-answers-still-need-evidence</guid>
      <pubDate>Wed, 27 May 2026 00:00:00 GMT</pubDate>
      <description>Why confident language still needs evidence, boundaries, and uncertainty paths.</description>
      <category>Explainer</category>
      <category>grounding</category><category>hallucination</category>
    </item>
    <item>
      <title>Why creativity settings are product decisions</title>
      <link>https://stackunseen.com/journal/why-creativity-settings-are-product-decisions</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-creativity-settings-are-product-decisions</guid>
      <pubDate>Mon, 25 May 2026 00:00:00 GMT</pubDate>
      <description>How sampling settings shape the user experience, not just the writing style.</description>
      <category>Explainer</category>
      <category>sampling</category><category>product</category>
    </item>
    <item>
      <title>Why more context can make an AI system worse</title>
      <link>https://stackunseen.com/journal/why-more-context-can-make-an-ai-system-worse</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-more-context-can-make-an-ai-system-worse</guid>
      <pubDate>Sun, 24 May 2026 00:00:00 GMT</pubDate>
      <description>Why a larger context window does not remove the need for context discipline.</description>
      <category>Explainer</category>
      <category>context</category>
    </item>
    <item>
      <title>Why vague prompts become vague systems</title>
      <link>https://stackunseen.com/journal/why-vague-prompts-become-vague-systems</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-vague-prompts-become-vague-systems</guid>
      <pubDate>Sun, 24 May 2026 00:00:00 GMT</pubDate>
      <description>Why prompts become operational instructions once real users depend on the system.</description>
      <category>Explainer</category>
      <category>prompting</category>
    </item>
    <item>
      <title>Why LLMs feel intelligent but still generate one step at a time</title>
      <link>https://stackunseen.com/journal/why-llms-feel-intelligent-but-still-generate-one-step-at-a-time</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-llms-feel-intelligent-but-still-generate-one-step-at-a-time</guid>
      <pubDate>Sat, 23 May 2026 00:00:00 GMT</pubDate>
      <description>The practical reason LLM behavior needs constraints, validation, and repeatable output design.</description>
      <category>Explainer</category>
      <category>models</category><category>generation</category>
    </item>
    <item>
      <title>Why tokenization quietly affects cost, limits, and reliability</title>
      <link>https://stackunseen.com/journal/why-tokenization-quietly-affects-cost-limits-and-reliability</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-tokenization-quietly-affects-cost-limits-and-reliability</guid>
      <pubDate>Sat, 23 May 2026 00:00:00 GMT</pubDate>
      <description>How a low-level text detail quietly becomes a product constraint for cost, latency, and reliability.</description>
      <category>Explainer</category>
      <category>tokens</category><category>cost</category>
    </item>
    <item>
      <title>Why model choice is rarely the first production AI problem</title>
      <link>https://stackunseen.com/journal/why-model-choice-is-rarely-the-first-production-ai-problem</link>
      <guid isPermaLink="true">https://stackunseen.com/journal/why-model-choice-is-rarely-the-first-production-ai-problem</guid>
      <pubDate>Fri, 22 May 2026 00:00:00 GMT</pubDate>
      <description>Why production AI succeeds or fails in the system around the model, not in the model choice alone.</description>
      <category>Explainer</category>
      <category>models</category><category>architecture</category>
    </item>
  </channel>
</rss>