Comparison · · 1 min read
LangGraph vs the OpenAI Agents SDK
Two ways to write the same supervisor. Compared on control flow, tracing, provider coupling, testing and what each makes hard.
Topic
Part of AI & Models: all AI & Models entries · the AI & Models lens on the map
Comparison · · 1 min read
Two ways to write the same supervisor. Compared on control flow, tracing, provider coupling, testing and what each makes hard.
Checklist · · 24 checks
Prompting, agents and tools, Claude Code, evaluation and safety. The habits that make Claude-based systems reliable, as a checklist you can run against your own setup.
Architecture pattern · · 1 min read
Agents that write to systems of record need the same transactional outbox that event-driven services use. This entry covers the shape and the trade-offs.
Architecture pattern · · 1 min read
How to design and run a multi-agent system in production, using the supervisor pattern, with its trade-offs and where human approval belongs.
Deep dive · · 6 min read
Six days of notes on supervising autonomous systems, and the same failure shape kept turning up: the control exists, it is documented, and nothing in the running system is bound by it.
Deep dive · · 6 min read
Six days of notes on agentic systems, and not one of the fixes that actually worked lived inside the model.
Explainer · · 1 min read
Why human review needs timing, authority, and context to be useful.
Explainer · · 2 min read
Why useful agent memory has to be separated, scoped, and governed.
Explainer · · 1 min read
Why reflection is valuable only when it decides whether to revise, retry, escalate, or stop.
Explainer · · 1 min read
Why planning matters when actions have cost, risk, or dependencies.
Explainer · · 2 min read
Why reliable agent behavior comes from loop design, not a single clever prompt.
Explainer · · 1 min read
Why agents need clear outcomes, boundaries, budgets, and stop conditions before autonomy.
Comparison · · 2 min read
A practical way to choose between conversation, repeatable workflows, and bounded autonomy.
LangChain · Useful reference · A practical talk on graph-based agent control flow with the failure modes called out.
OWASP · Recommended · The threat list to check any tool-calling agent against before it touches production data.
Anthropic · Recommended · The clearest written case for simple, composable agent patterns over frameworks.