Working Notes on Agent Systems/Brad Zhang

@teach_fireworks / X longform

Production-level agents require an infrastructure that can run for a long time, have controlled access, and leave evidence.

Production-level agents require an infrastructure that can run for a long time, have controlled access, and leave evidence. This 5-minute essence unfolds along seven questions:...

July 16, 2026 · 1 min read

Production-level agents require an infrastructure that can run for a long time, have controlled access, and leave evidence.
Figure 1 / source image

Production-level agents require an infrastructure that can run for a long time, have controlled access, and leave evidence.

This 5-minute essence unfolds along seven questions: how Agents take the initiative to complete tasks, how to obtain independent identities and minimal permissions, how Agents collaborate through APIs and MCP, why Harness is getting thinner, and how to use Outcome, Budget, and Evals constraints to continue iterating.

When Agent gains stronger independent execution capabilities, the team also needs to prevent prototype sprawl.

Common product direction, audit links and stopping conditions determine whether automation can ultimately be transformed into organizational output.

Key points •Agent identity, permissions and Audit Log should be designed uniformly in the production architecture.

  • MCP is suitable for exposing Tools with clear boundaries rather than copying the entire internal system.
  • Thin Harness gives the model more room for strategy while retaining necessary Guardrail.
  • Outcome, Budget and Evals jointly define the Agent's operating boundaries.

Visual summary

Article argument map

Generated from the post's content graph

FORMATTOPICCAPABILITYMARKETcoverscoverscoverscoverssignalssignalssignalssignalsFORMATlongform noteTOPICagentsTOPICharness engineeringTOPICevaluationTOPICtoolingCAPABILITYagent workflowCAPABILITYharness engineeringCAPABILITYAI-native workbenchCAPABILITYevaluationCAPABILITYproduct surface
Mermaid outline
flowchart LR
  format-long_post["longform note"]
  topic-agents["agents"]
  topic-harness-engineering["harness engineering"]
  topic-evaluation["evaluation"]
  topic-tooling["tooling"]
  capability-agent-workflow["agent workflow"]
  capability-harness-engineering["harness engineering"]
  capability-ai-native-workbench["AI-native workbench"]
  capability-evaluation["evaluation"]
  capability-product-surface["product surface"]
  format-long_post -->|covers| topic-agents
  format-long_post -->|covers| topic-harness-engineering
  format-long_post -->|covers| topic-evaluation
  format-long_post -->|covers| topic-tooling
  format-long_post -->|signals| capability-agent-workflow
  format-long_post -->|signals| capability-harness-engineering
  format-long_post -->|signals| capability-ai-native-workbench
  format-long_post -->|signals| capability-evaluation

Visual structure

Essay structure map

Built from summary and key paragraph positions

Production-level agents require an infrastructure that can run for a long time, have...THESISProduction-levelagents require aninfrastructure thatcan run for a longSIGNALProduction-levelagents require aninfrastructure thatcan run for a longOPERATORKey points •Agentidentity, permissionsand Audit Log shouldbe designed uniformlyIMPLICATION•Outcome, Budget andEvals jointly definethe Agent's operatingboundaries.
Mermaid outline
flowchart LR
  thesis["Production-level agents require an infrastructure that can run for a long time, have controlled access, and..."]
  signal["Production-level agents require an infrastructure that can run for a long time, have controlled access, and..."]
  operator["Key points •Agent identity, permissions and Audit Log should be designed uniformly in the production archit..."]
  implication["•Outcome, Budget and Evals jointly define the Agent's operating boundaries."]
  thesis -->|frames| signal
  signal -->|develops| operator
  operator -->|lands in| implication

Source: View the original post