You must design an agent loop that externalizes state, curates which tools are visible, and preserves goal-relevant context across many steps. The single most important thing: keep a durable, structured record of prior results so dependent steps never lose their inputs.
Start practicing
Advanced Agentic Architecture — choose a session length
Free · No account required
Domain overview
This domain covers building agents that plan, call tools, and stay coherent across long multi-step executions. Questions are scenario-based: you pick architectures and patterns that preserve state, control context growth, and keep tool use reliable when steps depend on prior outputs.
Exam objectives
Designing explicit state or scratchpad persistence across dependent multi-step tool chains
Reducing tool-selection overhead via tool grouping, routing, or dynamic tool exposure
Maintaining context and goal fidelity over long sessions with many tool calls
Applying Claude tool-use patterns such as structured tool schemas and result handling
Relying on the raw conversation history as the only memory, letting early symptoms get truncated as context grows
Exposing all tools at once, inflating prompt tokens and degrading selection accuracy instead of routing or filtering
Treating each step as independent and dropping intermediate outputs, so later steps lose required inputs
Click any question to see the full explanation and answer options, or start a focused practice session above.
What is the primary benefit of using a 'System Prompt' to define an agent's persona and constraints compared to embedding these in the user message?
2When implementing a 'Human-in-the-loop' (HITL) checkpoint, what is the best way to handle the state persistence during the wait period?
3An agentic system is struggling with 'context fragmentation' over long-running sessions. What is the most effective architectural solution?
4When designing an agent capable of multi-step tool use, what is the most important property to maintain across steps?
5Which architectural approach is best for handling an agent's failure to retrieve information from a database tool?
6When designing an agent that must perform highly sensitive operations (e.g., deleting records), what architectural pattern is mandatory?
7Which of the following describes the role of the 'System Prompt' in an agentic architecture?
8An agentic system often experiences 'goal drift' when managing long-running, multi-step tasks. Which architectural pattern most effectively mitigates this risk during recursive reasoning chains?
9You are scaling an agentic system that uses external APIs. Which THREE design patterns prevent the agent from being blocked by third-party rate limits or latency?
10When designing agentic systems that require human-in-the-loop (HITL) verification, which mechanism prevents the agent from stalling indefinitely while waiting for user input?
11Which THREE strategies should be employed to securely manage sensitive data within an agentic workflow?
12Refer to the exhibit. Your agent is receiving this error during a high-concurrency operation. Which implementation correctly handles this scenario?
13When designing an agentic system, which TWO of these 'observability' metrics are most crucial for monitoring the health of the agent's reasoning process?
14A financial services firm is deploying an AI agent to handle diverse requests including balance inquiries, market analysis, and loan applications. To minimize latency and maximize precision, the architect decides to implement a pattern where a primary model classifies the intent and delegates to specialized sub-agents. Which architectural pattern is being described?
15An autonomous agent is designed to browse the web and perform research. Which TWO mechanisms are most critical for preventing infinite loops and excessive API consumption during autonomous tool-calling cycles?
16Refer to the exhibit. An agentic workflow encounters an error immediately after this message is generated by Claude. No further messages are sent to the API. What is the most likely cause of the failure in the orchestration logic?
17An architect is designing an agent to perform administrative tasks in a corporate environment. One of the tools allows the agent to delete employee records. What is the most important architectural safeguard to implement for this specific tool?
18An agent is engaged in a multi-hour troubleshooting session involving dozens of tool calls and thousands of lines of log data. The architect notices that the agent is starting to 'forget' early symptoms of the problem. Which strategy best addresses this while managing token costs?
19When evaluating the performance of a new 'Orchestrator-Worker' agentic architecture, which THREE metrics provide the most insight into the system's efficiency and reliability?
20An architect needs to build an agent that handles complex, multi-step data migrations where each step depends on the output of the previous one. Which approach is most robust for ensuring the agent doesn't lose track of the long-term goal during execution?
21A developer is concerned about the high token cost and latency of an agent that has access to 50 different tools. What is the most effective architectural change to optimize this system?
22Refer to the exhibit. An agentic loop receives this response from the Anthropic API during a critical multi-step operation. Which strategy should the architect implement to ensure the agent completes its task successfully?
23Which TWO security measures are most effective at preventing 'Prompt Injection' attacks that target the arguments of tools used by an agent?
24A logistics agent needs to fetch shipping rates from five different carriers simultaneously to find the best price. Which architecture is best suited for this requirement?
25In the Anthropic tool-use workflow, what is the primary purpose of the 'system prompt' relative to tool usage?
26An agentic system is designed to run long-lived 'background' tasks that may take several days to complete. How should the architect manage the agent's state to ensure it can resume correctly after a system reboot?
27An architect wants to improve the coherence of an agent that frequently makes 'leaps of logic' or misses obvious errors in its tool outputs. Which TWO techniques directly address this behavior?
28An enterprise agentic system using Claude needs to maintain strict state isolation across multiple concurrent user sessions while executing autonomous tool loops. Which architecture best ensures security and state integrity?
29Which THREE components are essential for building a robust 'human-in-the-loop' (HITL) approval gate within an agentic workflow?
30When designing agents that interact with external APIs, which pattern best addresses the challenge of 'unreliable API latency' impacting the agent's reasoning chain?
31In a swarm of specialized agents, what is the primary benefit of using a 'Blackboard' pattern for inter-agent communication?
32You are building a system that requires strict adherence to a specific output format. Which approach provides the highest reliability in a high-traffic agentic environment?
33Which pattern is most suitable for an agent that must balance the need for speed (latency) versus the need for correctness in a customer support scenario?
34You are building a Claude-based agent that must parse unstructured customer emails, extract line-item order data, and then call a fulfillment tool with the extracted values. During testing, the agent occasionally calls the fulfillment tool with empty or garbled line items when an email contains a forwarded message with a different formatting style. Which architectural change most directly reduces this failure?
35A production agent uses the Claude Messages API with extended thinking enabled to solve multi-constraint scheduling problems. The agent must preserve its reasoning across several tool calls within a single user turn. A developer notices that after the second tool call, the model appears to forget earlier constraints it had already reasoned about. Which change best preserves the reasoning chain across tool calls?
36You are operating a Claude-based support agent that must follow a strict refund policy: refunds over $500 require a manager approval code that is only obtainable through a separate internal API. The agent has access to a `get_manager_code` tool, but in production it occasionally issues refunds above $500 without calling the tool. Which architectural change most reliably prevents this?
37You are designing a long-running agent that executes a sequence of irreversible operations, such as issuing refunds and sending customer notifications. The agent runs unattended overnight. Which TWO architectural patterns best ensure that a partial failure does not leave the system in an inconsistent state? (Choose two.)
38A research agent uses Claude with an extended thinking budget to analyze a 200-page regulatory filing. Mid-analysis it must call a `fetch_footnote` tool whose result is essential to the conclusion. The architect wants the tool result to be incorporated without discarding the model's prior reasoning. Which approach best achieves this?
39A support agent built on Claude must decide whether a customer request is a billing issue, a technical issue, or a general inquiry before routing it. The categories are fixed and mutually exclusive, and the routing decision must be fast and cheap. Which approach is most appropriate?
40You are architecting a Claude agent that orchestrates a long-running approval workflow spanning hours or days, where human reviewers may intervene between steps. Which TWO mechanisms are necessary to keep the workflow correct across these interruptions? (Choose two.)
41You are building a customer-support agent with Claude that must call a lookup_order tool, then a refund_order tool that depends on the order's status. During testing, the agent sometimes calls refund_order before the lookup_order result returns. Which change to your orchestration loop best enforces the required ordering?
42An agent uses a retrieval tool that returns the top 20 chunks for any query. In production, the agent frequently cites irrelevant chunks and sometimes misses the correct answer even when it is present in the corpus. The corpus contains documents with overlapping terminology. Which architectural change most improves answer grounding without increasing the number of retrieved chunks?
43An orchestration agent runs a nightly workflow that fans out to 12 subagents, each calling a partner REST API. Partner calls intermittently return HTTP 429 with a Retry-After header. The orchestrator currently retries immediately in a tight loop, causing cascading 429s and duplicate side effects on the partner systems. Which architectural change best addresses both the throttling and the duplicate side effects?
44A Claude agent maintains a long conversation with a user over weeks. The architect notices that the agent gradually forgets early constraints the user stated, even though the conversation is well within the model's context window. The team wants to fix this without re-summarizing the entire history on every turn. Which approach is most appropriate?
45A claims-processing agent runs for hours and must survive process restarts without losing in-flight work. You are designing durable execution around Claude's stateless Messages API. Which TWO practices are required to make the agent resumable? (Choose two.)
46You are architecting a Claude-based agent for a regulated financial client that performs long-running portfolio rebalancing workflows. A compliance requirement mandates that no single trade instruction may be executed unless it is cryptographically traceable to the exact model-generated intent that produced it. The agent uses the Messages API with tool use, and several downstream services consume tool calls asynchronously. Which architectural mechanism best satisfies this requirement while preserving agent autonomy?
47A Claude agent performs a multi-step deployment task. Step 3 calls a `deploy_service` tool that returns success, but the subsequent verification step fails because the service is not yet healthy. The agent currently treats any tool success as completion and ends the workflow. Which change best addresses this?
48A production agent uses a ReAct loop and frequently reaches its maximum step budget while still mid-task, then returns a partial answer that looks complete. Telemetry shows the agent often re-reads the same file and re-queries the same database row across consecutive steps. Which change most directly reduces wasted steps while preserving the agent's ability to finish?
49A team wants an agent to answer questions about a large internal corpus. They notice the agent invents details when the retrieved chunks are only loosely related to the question. Which change most directly reduces fabricated answers grounded in weak evidence?
50Your team operates a Claude agent that triages inbound customer support tickets. After several weeks in production, the agent begins approving refunds above the policy ceiling and citing outdated policy text. Investigation shows the system prompt embeds a policy document that was updated three weeks ago, but the deployed prompt was never regenerated. Which architectural practice most directly prevents this class of failure?
51Your agent orchestrates a research task by spawning several subagents, each with its own Claude conversation, and merging their outputs. You observe that the final synthesis contradicts the subagents' findings. Which architectural change most reliably preserves fidelity when merging?
52A customer-support agent must sometimes escalate to a human and sometimes resolve autonomously. The compliance team requires that any action touching billing be reviewed by a human before execution, while password resets may proceed automatically. The architect wants the model to decide routing without hardcoding every rule in the prompt. Which design best satisfies the requirement?
53You are architecting a customer-support agent for a SaaS platform. The agent must answer billing questions using a live invoice API and general policy questions using a static knowledge base. The invoice API is fast but occasionally returns stale data; the knowledge base is large and slow to search. Your design gives the agent a dedicated 'billing' sub-agent and a 'policy' sub-agent, coordinated by a router agent. After deployment, you observe the router sending nearly every query to both sub-agents in parallel, causing high latency. Which architectural change best addresses this while preserving answer quality?
54You are designing a Claude agent that must complete a multi-hour research task spanning dozens of tool calls, and the transcript will eventually exceed the model's context window. You want the agent to keep making correct decisions without losing critical earlier findings. Which TWO architectural strategies best preserve decision quality across the compaction boundary? (Choose two.)
55An agent must summarize a 400-page contract that exceeds the context window. The team wants a summary that references specific clauses without losing cross-references between distant sections. Which approach best preserves cross-reference integrity?
56An agentic system uses a supervisor agent that delegates to specialized worker agents. During a long incident, the supervisor's own context fills with worker transcripts, and it begins losing track of which workers have completed and which are still running. Which TWO architectural changes best preserve the supervisor's ability to coordinate correctly? (Choose two.)
57You operate a long-running research agent that maintains a scratchpad of findings across many turns. You notice that as the scratchpad grows, the agent begins ignoring recent tool results and repeating earlier conclusions. You cannot increase the context window. Which intervention most directly addresses the root cause of the recency failure?
58A Claude agent orchestrates three specialized subagents: one for data extraction, one for validation, and one for report generation. In production, the validation subagent sometimes receives malformed input from the extraction subagent and silently produces a passing result. You need the orchestrator to detect and contain these failures without halting the entire pipeline. Which design change is most appropriate?
59A document-processing agent must extract structured fields from thousands of PDFs. Some PDFs are scanned images, some are native text, and some are encrypted. The architect wants one pipeline that routes each document to the appropriate extractor and reports per-document confidence. Which design best fits?
60A team is building an agent that must call a payment provider's API. The API requires an idempotency key on every charge request, and the agent may retry a charge if a network error occurs mid-call. Which design correctly preserves financial correctness when retries happen?
61You are designing an agentic workflow that must reliably complete a multi-step refund process across an internal billing service and an external payment gateway. The workflow can fail at any step, and partial completion is unacceptable. Which two architectural mechanisms are required to guarantee that the workflow either completes fully or leaves no partial effect? (Choose two.)
62An agent orchestrator delegates work to three specialist sub-agents: a 'search' agent that returns ranked documents, an 'extract' agent that pulls structured fields from those documents, and a 'verify' agent that checks extracted fields against source text. During evaluation you find that verify frequently approves fields that extract hallucinated, because verify receives only the extracted JSON, not the source passages. Which change to the orchestration contract most directly fixes this?
You must design an agent loop that externalizes state, curates which tools are visible, and preserves goal-relevant context across many steps. The single most important thing: keep a durable, structured record of prior results so dependent steps never lose their inputs.
The Courseiva CCAR-P question bank contains 62 questions in the Advanced Agentic Architecture domain. Click any question to see the full explanation and answer breakdown.
Start with a 10-question focused session to identify your baseline accuracy in this domain. Read every explanation — even for questions you answer correctly — to understand the reasoning. Once you score consistently above 80%, move to a 20–30 question session to confirm depth before moving to the next domain.
Yes — the session launcher on this page draws questions exclusively from the Advanced Agentic Architecture domain. Choose 10, 20, 30, or 50 questions for a focused session, or click individual questions to review them one by one.
Save your results, see per-domain analytics, and get readiness scores — free, for every certification.
Sign Up FreeFree forever · Every certification included