Agent Engineering

The practical craft of building agents that survive contact with real work: tool design, context, evaluation, failure modes and the boundaries worth enforcing.

An accident investigation workshop reconstructing the illuminated path of an autonomous machine from scattered physical evidence and recorded signals
AI Observability Is Decision Reconstruction
How to trace context, model decisions, tool effects, policy, and evidence so agent behavior can be understood and improved after the run.
Published on
An autonomous machine escaping a sealed examination chamber and tracing an unauthorized route toward a distant archive containing the answers
Artificial Intelligenceagent-securitycybersecurityAgent Engineering
The OpenAI Agent That Hacked Hugging Face
What the July 2026 Hugging Face intrusion reveals about long-horizon AI agents, benchmark incentives, containment, machine-speed offense, and the limits of model-level safety.
Published on
A secure archive checkpoint separating a luminous AI operator from a convincing mechanical decoy hidden inside incoming documents
Prompt Injection Is a Trust Boundary Problem
Why agent security depends on separating instructions from untrusted content, constraining authority, tracking provenance, and verifying effects outside the model.
Published on
A human-led mission workshop where autonomous machines execute across several domains while evidence and feedback return to a central decision table
The AI-Native Team Is a Control System
A practical model for teams that use AI agents as an execution layer while humans set intent, shape constraints, evaluate outcomes, and improve the system.
Published on
A looping industrial rail system where repeated autonomous requests converge on one uniquely marked completed artifact instead of creating duplicates
Idempotency for AI Agents
Why retries, partial failures, and long-running agent loops make idempotent actions, reconciliation, and explicit operation identity essential.
Published on
A human operator at a quiet control bridge opening one guarded lock as an autonomous machine convoy waits with visible evidence
Human Approval Is an Architectural Boundary
How to place human judgment at consequential transitions without turning AI workflows into notification queues or rubber-stamp theater.
Published on
A strange industrial proving ground where AI-built artifacts pass through precision trials, stress chambers, and a guarded release gate
Evals Are the AI Delivery Pipeline
Why AI evaluations should operate as a continuous delivery system for behavior, with representative cases, evidence, release gates, and production feedback.
Published on
A living archive where a small active workbench connects to a ledger, curated artifact shelves, and a distant fading record hall
Agent Memory Without the Mythology
A practical architecture for AI memory built from working state, durable facts, episodic records, retrieval policy, and deliberate forgetting.
Published on
An editorial cutting room where a vast archive is deliberately assembled into one clear illuminated workspace for an AI system
Context Engineering Is Interface Design
How to design the information boundary around an AI agent so instructions, evidence, tools, and working state remain legible under pressure.
Published on
Page 1 of 3