Agents

Production agent systems — OpenAI Agents SDK and LangGraph orchestration, tool calling, state and memory, human approval gates, and the failure modes that break demos before they reach real users.

Articles21
Topics6

All Articles

OpenAI Agents SDK Sandbox Agents: The Production Isolation Playbook

OpenAI Agents SDK Sandbox Agents: The Production Isolation Playbook

Use OpenAI Agents SDK Sandbox Agents safely: choose the execution boundary, protect secrets, persist state, log runs, and gate production rollout.

Repo Gardener build log

Repo Gardener build log

A weekly Claude Managed Agents workflow that reproduces reported bugs, derives issue hygiene mechanically, remembers verified rulings, and never posts without maintainer approval.

Upgrade Steward build log

Upgrade Steward build log

An eve agent that researches breaking dependency changes, proves migrations in a sandbox, and binds pull request publication to human-approved evidence.

Shotmap Agent build log

Shotmap Agent build log

A BYOK pre-production agent that turns rough video briefs into timed shot maps, continuity ledgers, and generator-ready prompt packets.

LangChain Context Engineering for Production Agents

LangChain Context Engineering for Production Agents

Use LangChain context engineering to control state, tools, memory, compression, and evals before a production agent hits traffic.

Human Approval Gates for AI Agents

Human Approval Gates for AI Agents

Build HITL approval gates for AI agents: where to pause, what to log, how to resume, and how to prevent reviewer rubber-stamping.

Grok API Build Guide: xAI for Production Agents

Grok API Build Guide: xAI for Production Agents

Use the Grok API when live web/X search and tool-heavy agents matter. Here is the production checklist, pricing, limits, and controls to verify.

Repro Card Agent build log

A BYOK agent app for turning noisy bug evidence into coding-agent-ready repro cards.

First Frame Agent build log

A BYOK agent app for turning rough product notes into image-to-video first-frame slates.

Agent Input Firewall build log

A Codex and Claude Code skill for quarantining untrusted external text before coding agents act.

type-led-launch-design build log

A DESIGN.md skill and static landing-page example for type-led technical product launches.

Shoot Card Agent build log

A BYOK visual director for turning rough product notes into shoot-ready image prompt cards.

Agent Config Audit Build Log

A build log for agent-config-audit, a Claude Code and Codex skill that reviews agent setup files before installation or execution.

agent-ready-ui Build Log

A build log for agent-ready-ui, a skill that makes web UI flows reliable for browser and GUI agents without brittle selectors.

Context Engineering vs Prompt Engineering for Production Agents

Context Engineering vs Prompt Engineering for Production Agents

Context engineering is the production control plane for agents. Learn when prompts matter, what context layers to ship, and what to log before traffic.

Agent Memory for Production AI Systems

Agent Memory for Production AI Systems

Design agent memory as governed state: what to store, what to forget, how to retrieve it, and which evals catch stale or unsafe recall.

OpenAI Agents SDK vs Pydantic AI for Production Agents

OpenAI Agents SDK vs Pydantic AI for Production Agents

Choose OpenAI Agents SDK for OpenAI-native runs. Choose Pydantic AI when typed Python, provider flexibility, and durable approvals matter.

Google ADK vs LangGraph for Production Agents

Google ADK vs LangGraph for Production Agents

Compare Google ADK and LangGraph for production agents: state, human approval, deployment, observability, pricing, and the decision rule.

OpenAI Agents SDK TypeScript vs Python for Production Agents

OpenAI Agents SDK TypeScript vs Python for Production Agents

Choose TypeScript or Python for OpenAI Agents SDK by production ownership: product runtime, worker path, tracing, guardrails, handoffs, MCP, and evals.

LangChain vs LangGraph for Production Agents

Use LangChain for simple agent harnesses. Use LangGraph when production agents need durable state, retries, interrupts, approvals, and deployment.

OpenAI Agents SDK vs LangGraph for Production Agents

OpenAI Agents SDK vs LangGraph for Production Agents

Choose OpenAI Agents SDK for OpenAI-native loops. Choose LangGraph when durable graph state, provider freedom, and custom control matter.

Newsletter

One letter, every week. Working systems — not hot takes.

Build logs, agentic engineering decisions, agent failures, evals, and what survives real users. Sent weekly, never more.

Weekly. No spam. Unsubscribe anytime.