Design effective AI agent systems with clear architecture, capability planning, and tool integration. Use when building agent systems, designing agent workflows, planning tool ecosystems, defining agent roles and boundaries, or architecting multi-agent systems. Covers agent patterns, agentic loops, capability design, and system composition.
npx skills add https://github.com/majiayu000/claude-skill-registry --skill Agent Design Architecture
Effective agent systems require thoughtful architecture that balances capability, safety, and maintainability. This skill guides designing agent systems from first principles.
An agent is a system that:
Key distinction from simple chatbots:
User Request → Model Reasoning → Tool Selection → Tool Execution → Result → Response
When to use: Single-step tasks, straightforward tool selection
Example: "Find the latest sales data and summarize"
Complexity: Low | Autonomy: Low
Design considerations:
Goal → Reasoning → Action Selection → Execute → Observe → Success?
↑_____No_________↓
Reflect & Replan
When to use: Complex, multi-step tasks; planning required
Example: "Analyze customer complaints, identify patterns, propose solutions"
Complexity: Medium | Autonomy: Medium
Design considerations:
Goal → Decompose → Create Plan → Execute Step → Verify → Adapt Plan
When to use: Complex goals requiring explicit planning; stakeholder visibility
Example: "Develop a product roadmap balancing technical debt and features"
Complexity: High | Autonomy: High
Design considerations:
Coordinator Agent
├─ Research Agent (gathers info)
├─ Analysis Agent (processes data)
└─ Synthesis Agent (creates output)
When to use: Complex domains requiring specialized sub-agents
Example: "Write a comprehensive research report"
Complexity: High | Autonomy: High
Design considerations:
Event → Trigger Rules → Immediate Action → Update State
When to use: Real-time monitoring, rapid response
Example: "Monitor system alerts and escalate critical issues"
Complexity: Medium | Autonomy: Medium
Design considerations:
User Interface
↓
Orchestration Layer
↙ ↓ ↓ ↓ ↘
Agent Agent Agent Agent Agent
(specialized roles)
When to use: Organization-scale automation, diverse domains
Example: "Coordinate content creation across marketing, sales, support"
Complexity: Very High | Autonomy: High
Design considerations:
Clarify what the agent owns vs. what it doesn't:
AGENT OWNS:
✓ Reasoning and planning for goals
✓ Tool orchestration within domain
✓ Data gathering and synthesis
✓ Communicating confidence/uncertainty
HUMAN OWNS:
✓ Final decisions on high-stakes items
✓ Setting priorities and constraints
✓ Defining success criteria
✓ Judging value of outcomes
SHARED RESPONSIBILITY:
~ Monitoring for issues
~ Learning from outcomes
~ Ethical judgment calls
Tools are how agents affect the world. Thoughtful tool design is critical.
Tool Selection Criteria:
Tool Description Template:
Name: [Clear, unambiguous]
Purpose: [What problem does this solve?]
When to use: [Specific situations]
When NOT to use: [Common mistakes]
Required inputs: [Parameters, formats, constraints]
Output: [Structure, semantics, error cases]
Limitations: [What it doesn't do]
Side effects: [What else happens when called?]
Cost: [Latency, computational, financial]
Example: Good Tool Description
Name: CheckInventoryLevel
Purpose: Query current stock levels for a specific product
When to use: Before committing to delivery dates or promotions
When NOT to use: Don't use to predict future demand (use ForecastDemand instead)
Required inputs: product_id (string), warehouse_id (string, optional)
Output: {
available_units: int,
reserved_units: int,
low_stock_threshold: int,
last_updated: timestamp,
confidence: "high" | "medium" | "low"
}
Limitations: Data updates every 15 minutes; historical data not available
Side effects: Increments query count (throttled at 100/min)
Cost: ~50ms latency
Explicit criteria prevent agent drift:
Success criteria (agent should optimize for):
Failure modes (agent should avoid):
Example:
GOAL: Write product descriptions for 100 SKUs
SUCCESS LOOKS LIKE:
✓ Descriptions are 50-100 words
✓ Highlight unique features accurately
✓ Reading level is accessible (8th grade)
✓ Include relevant keywords naturally
✓ Completed in <2 hours
AVOID:
✗ Making up product features
✗ Copying competitor descriptions
✗ Generic templated language
✗ Inaccurate specifications
✗ Exceeding 2-hour time budget
Define what agent can decide vs. what requires human approval:
Autonomy Levels:
Decision tree for setting boundaries:
Is decision REVERSIBLE?
→ Yes → Is impact LOW-STAKES?
→ Yes → AUTONOMOUS (agent decides freely)
→ No → DELEGATED (with monitoring)
→ No → Is decision TIME-SENSITIVE?
→ Yes → DELEGATED (with human review)
→ No → RECOMMENDED (with human approval)
Atomic Tools (Single responsibility):
GetCustomerHistory() - retrieve onlyUpdateOrderStatus() - modify onlyValidateEmailAddress() - verify onlyComposite Tools (Common combinations):
CreateAndAssignTicket() - create ticket + assign to agent + notifyAnalyzeSentimentAndRoute() - analyze + determine category + routeRule of thumb: Start atomic, compose at the agent logic level, create composites only for frequent, safe combinations.
Input Validation:
Tool receives: customer_id
Validate:
✓ Is non-null string
✓ Matches expected format (UUID)
✓ Customer exists in system
✓ Current user authorized to access
→ Only then: Execute
Output Safety:
Tool computes: credit_score
Before returning to agent:
✓ Is score within expected range?
✓ Is score based on fresh data?
✓ Should sensitive fields be redacted?
→ Return with confidence/caveat info
Rate Limiting & Quotas:
Tool: SendEmail
Limits:
- 10 emails per agent per minute
- 100 emails per day total
- No sending outside business hours
- Escalate if limits approached
What's the ideal flow if everything works?
User: "Analyze my Q3 sales and recommend changes"
1. Agent: Gather sales data for Q3
2. Agent: Segment by product/region/customer
3. Agent: Calculate trends (YoY, MoM)
4. Agent: Identify top performers and underperformers
5. Agent: Research market context
6. Agent: Synthesize recommendations
7. Agent: Present findings with confidence levels
8. User: Review and decide
Where could things go wrong?
FAILURE POINT: Sales data incomplete/corrupted
DETECTION: Verify data completeness before analysis
HANDLING: Report gaps, ask for manual data, skip analysis
FAILURE POINT: Market context misleading
DETECTION: Cross-reference multiple sources
HANDLING: Flag conflicting info, present both interpretations
FAILURE POINT: Recommendation outside feasibility
DETECTION: Run recommendation through feasibility checker
HANDLING: Adjust recommendation or escalate
FAILURE POINT: Analysis takes too long
DETECTION: Track elapsed time against budget
HANDLING: Report progress, offer partial results, ask to continue
How does agent recover from failures?
Retry Logic:
Escalation:
State Management:
Sequential Handoff (Agent A → Agent B):
ResearchAgent gathers data
→ Passes to AnalysisAgent
→ Passes to WritingAgent
→ Returns result
Best for: Linear workflows
Challenge: One agent bottleneck blocks others
Parallel Execution (A & B & C run concurrently):
ResearchAgent-A gathers market data
ResearchAgent-B gathers competitor data
ResearchAgent-C gathers internal data
→ All results → AnalysisAgent (merges)
Best for: Gathering info from multiple sources
Challenge: Coordinating and merging results
Broadcast (Agent communicates to many):
CoordinatorAgent creates plan
→ Sends to all sub-agents
→ Each executes their part
→ Reports back progress
Best for: Orchestrating specialization
Challenge: Managing dependencies and failures
Shared State (All agents access common state):
{
"goal": "..."
"status": "in_progress",
"findings": [...],
"blockers": [...],
"decisions_needed": [...]
}
Advantage: Agents always see current picture
Risk: Conflicts if multiple agents modify
Passed State (State passes between agents):
Agent-A → outputs with state → Agent-B reads and updates → Agent-C reads updated state
Advantage: Clear causality
Risk: Out-of-sync if agents modify independently
Event Logging (Append-only record):
Events:
- Agent-A started
- Agent-A found X
- Agent-B started
- Agent-B found Y (contradicts X)
- Issue flagged: contradiction
Advantage: Auditability, no conflicts
Risk: Complexity in reconstructing state
When designing an agent system, verify:
| Anti-Pattern | Problem | Solution |
|------------------|-----------|------------|
| Too much autonomy | Agent makes poor decisions without guidance | Define clear boundaries and constraints |
| Unclear tool descriptions | Agent misuses tools, makes mistakes | Invest in precise tool documentation |
| No iteration limit | Agent loops forever, wastes resources | Set max iterations + explicit exit criteria |
| Silent failures | Agent appears successful but isn't | Return confidence scores, flag assumptions |
| No state tracking | Can't debug or resume | Log decisions and reasoning |
| Unrealistic tool set | Agent "hallucinates" because tools insufficient | Audit tool sufficiency for goal class |
| Single point of failure | One tool breaks, whole workflow fails | Provide tool alternatives/fallbacks |
| Vague success criteria | Agent doesn't know what "good" looks like | Define explicit metrics and examples |
Complete development kit for Microsoft 365 Copilot declarative agents with three comprehensive workflows (basic, advanced, validation), TypeSpec support, and Microsoft 365 Agents Toolkit integration
Format and structurally validate local treatment-plan documentation after clinical decisions have already been supplied and verified by authorized licensed professionals. Use for source traceability, clinician-authored intervention records, goals and checkpoints, shared-decision records, reconciliation handoffs, and release gates—not for clinical decision-making.
> provider/change budget/修改卖家/修改预算/draft/草稿/我的任务/my tasks/what am I working on/关闭/取消任务/决策列表/decision list/指定服务商/browse (sender.role = COUNTERPARTY, not you); (3) literal "Read the okx-ai skill" (or legacy "Read the okx-agent-task skill") in the envelope.
Automate payer review of prior authorization (PA) requests. This skill should be used when users say "Review this PA request", "Process prior authorization for [procedure]", "Assess medical necessity", "Generate PA decision", or when processing clinical documentation for coverage policy validation and authorization decisions.
Expert in designing and building autonomous AI agents. Masters tool use, memory systems, planning strategies, and multi-agent orchestration.
Autonomous agents are AI systems that can independently decompose goals, plan actions, execute tools, and self-correct without constant human guidance. The challenge isn't making them capable - it's making them reliable. Every extra decision multiplies failure probability.
Orchestrates design workflows by routing work through brainstorming, multi-agent review, and execution readiness in the correct order.
Structured persuasion for tech leads, PMs, and founders—not activity logs. Five scenarios (kickoff, status update, wrap-up, investor pitch, solution selling) on one 5-part framework (Hook→Context→Proposal→Evidence→Ask). AI prompts for missing materials and audience context; pre-submit checklist. Claude Code plugin; Cursor, Codex, and chat via prompts.
Take majiayu000/agent design architecture from the repository into ~/.claude/skills for personal
use, or into .claude/skills inside a project.
The agent identifies a skill by the name field in its header. Two skills with the
same name cannot sit side by side — one of them will be ignored.