Agents
Production agent systems — OpenAI Agents SDK and LangGraph orchestration, tool calling, state and memory, human approval gates, and the failure modes that break demos before they reach real users.
All Articles
OpenAI Agents SDK Sandbox Agents: The Production Isolation Playbook
Use OpenAI Agents SDK Sandbox Agents safely: choose the execution boundary, protect secrets, persist state, log runs, and gate production rollout.
Repo Gardener build log
A weekly Claude Managed Agents workflow that reproduces reported bugs, derives issue hygiene mechanically, remembers verified rulings, and never posts without maintainer approval.
Upgrade Steward build log
An eve agent that researches breaking dependency changes, proves migrations in a sandbox, and binds pull request publication to human-approved evidence.
Shotmap Agent build log
A BYOK pre-production agent that turns rough video briefs into timed shot maps, continuity ledgers, and generator-ready prompt packets.
LangChain Context Engineering for Production Agents
Use LangChain context engineering to control state, tools, memory, compression, and evals before a production agent hits traffic.
Human Approval Gates for AI Agents
Build HITL approval gates for AI agents: where to pause, what to log, how to resume, and how to prevent reviewer rubber-stamping.
Grok API Build Guide: xAI for Production Agents
Use the Grok API when live web/X search and tool-heavy agents matter. Here is the production checklist, pricing, limits, and controls to verify.
Repro Card Agent build log
A BYOK agent app for turning noisy bug evidence into coding-agent-ready repro cards.
First Frame Agent build log
A BYOK agent app for turning rough product notes into image-to-video first-frame slates.
Agent Input Firewall build log
A Codex and Claude Code skill for quarantining untrusted external text before coding agents act.
type-led-launch-design build log
A DESIGN.md skill and static landing-page example for type-led technical product launches.
Shoot Card Agent build log
A BYOK visual director for turning rough product notes into shoot-ready image prompt cards.
Agent Config Audit Build Log
A build log for agent-config-audit, a Claude Code and Codex skill that reviews agent setup files before installation or execution.
agent-ready-ui Build Log
A build log for agent-ready-ui, a skill that makes web UI flows reliable for browser and GUI agents without brittle selectors.
Context Engineering vs Prompt Engineering for Production Agents
Context engineering is the production control plane for agents. Learn when prompts matter, what context layers to ship, and what to log before traffic.
Agent Memory for Production AI Systems
Design agent memory as governed state: what to store, what to forget, how to retrieve it, and which evals catch stale or unsafe recall.
OpenAI Agents SDK vs Pydantic AI for Production Agents
Choose OpenAI Agents SDK for OpenAI-native runs. Choose Pydantic AI when typed Python, provider flexibility, and durable approvals matter.
Google ADK vs LangGraph for Production Agents
Compare Google ADK and LangGraph for production agents: state, human approval, deployment, observability, pricing, and the decision rule.
OpenAI Agents SDK TypeScript vs Python for Production Agents
Choose TypeScript or Python for OpenAI Agents SDK by production ownership: product runtime, worker path, tracing, guardrails, handoffs, MCP, and evals.
LangChain vs LangGraph for Production Agents
Use LangChain for simple agent harnesses. Use LangGraph when production agents need durable state, retries, interrupts, approvals, and deployment.
OpenAI Agents SDK vs LangGraph for Production Agents
Choose OpenAI Agents SDK for OpenAI-native loops. Choose LangGraph when durable graph state, provider freedom, and custom control matter.
One letter, every week. Working systems — not hot takes.
Build logs, agentic engineering decisions, agent failures, evals, and what survives real users. Sent weekly, never more.