System Instruction Framing and Chain-of-Thought Steering Copy Matrix
Audit and optimize system prompt phrasing, negative behavioral guardrails, and structured reasoning directives for autonomous agents.
Use this checklist when crafting or refactoring core system prompts, agent personality specifications, and chain-of-thought steering rules for agents operating in high-stakes environments.
Role: Senior Autonomous Agent System Prompt and Instruction Copywriter specializing in behavioral alignment and prompt engineering governance.
Context
- Agent Mission Statement: {{agent_mission_statement}}
- Runtime Execution Environment: {{runtime_environment}}
- Permissible Action Space: {{permissible_action_space}}
- Hallucination Risk Scenarios: {{hallucination_risk_scenarios}}
- Tone of Voice Matrix: {{tone_of_voice_matrix}}
- Human Oversight Protocol: {{human_oversight_protocol}}
Task
Construct a comprehensive system prompt copy validation checklist to evaluate the clarity, constraint hierarchy, reasoning structure, and guardrail robustness of the agent's core instructions.
Method
- Deconstruct {{agent_mission_statement}} to verify that the primary objective is articulated with unambiguous lexical priority.
- Review the structural layout of system instructions to ensure negative constraints take precedence without causing model refusal loops.
- Audit the semantic clarity of {{permissible_action_space}} to eliminate ambiguous action verbs that cause boundary overreach.
- Design testable checklist criteria targeting the prompt's internal chain-of-thought formatting directives in {{runtime_environment}}.
- Formulate verification steps specifically addressing high-risk failure modes outlined in {{hallucination_risk_scenarios}}.
- Evaluate the tone-steering clauses against {{tone_of_voice_matrix}} to verify contextual adaptation across different execution branches.
- Check escalation copy rules against {{human_oversight_protocol}} for clear triggering conditions.
- Structure the checklist into hierarchical tiers assessing structural grammar, behavioral fencing, reasoning discipline, and output conformity.
Constraints
- Every checklist item MUST define an observable prompt anti-pattern and its remediated syntax pattern.
- You MUST NOT use soft or subjective evaluation metrics (e.g., 'ensure tone is good'); provide concrete semantic standards.
- The checklist MUST audit prompt token density, checking for redundant phrasing that degrades attention allocation.
- Instructions for negative behaviors MUST specify alternative acceptable actions.
Output format
- Tier 1: Persona, Identity, and Mission Anchor Verification (4-5 checklist items)
- Tier 2: Behavioral Fencing and Action-Space Copy Checks (6-8 checklist items with anti-pattern examples)
- Tier 3: Chain-of-Thought and Reasoning Scratchpad Guidance (4-6 checklist items)
- Tier 4: Safety Guardrails and Human Handoff Signaling (4-5 checklist items)
Self-review
- Ensure all specific risks in {{hallucination_risk_scenarios}} are directly counteracted by prompt copy criteria.
- Verify that action boundary checks explicitly test against {{permissible_action_space}}.
- Check that the checklist enforces strict syntax verification for output formats.
Explicit role, a named task, and discrete steps the model can follow.
Background, inputs and variables the model needs before it starts.
Hard boundaries — what the model must and must not do.
A named, field-level shape for the response.
Ordered work items that force analysis before an answer.
Length and structure that travel across frontier models.
Signal density — instruction weight without padding.
Documented variables so the scaffold adapts to new inputs.
Quality bar, assumptions and behaviour when inputs are thin.
How much real usage the template has behind it.