Be able to wire a ReAct loop: send tools, read stop_reason, execute requested tools, and return every tool_result with its tool_use_id before the next turn. The critical thing is correctly pairing parallel tool results and detecting repeated calls to stop runaway loops.
Start practicing
Agentic Architecture and Orchestration — choose a session length
Free · No account required
Domain overview
This domain covers how Claude-based agents plan, call tools, and recover from failure. Expect scenario questions on the ReAct reason-act loop, parallel tool_use blocks and their matching tool_result messages, loop detection, stop_reason handling, and choosing sampling settings for deterministic tool-driven work.
Exam objectives
Structuring the Messages API tool_use and tool_result content blocks so Claude can continue after a tool call
Handling stop_reason values such as tool_use and end_turn to decide when to loop or terminate
Detecting and breaking repeated identical tool calls, for example via call history or iteration caps
Setting low temperature for deterministic extraction and API-calling agents rather than creative sampling
Returning tool_result blocks without the matching tool_use_id, or omitting results for some parallel calls, so Claude cannot pair outputs to requests
Ignoring stop_reason and re-prompting blindly, causing infinite loops or premature termination instead of controlled iteration
Using high temperature for data extraction agents, producing inconsistent tool arguments and unreliable structured output
Click any question to see the full explanation and answer options, or start a focused practice session above.
Refer to the exhibit. The provided JSON represents a partial response from Claude. According to the Anthropic API specification for tool use, what is the mandatory next step for the orchestrator to continue the agentic loop?
2A company is deploying an agent that interacts with a sensitive financial database. To maintain security and accuracy, the architect wants to ensure that any 'delete' or 'transfer' actions are reviewed by a human before execution. Which agentic pattern should be used?
3When designing a multi-agent system, what is the primary benefit of the 'Orchestrator-Worker' pattern compared to a single-agent architecture with 20 tools?
4An agent tasked with data analysis repeatedly fails because the 'tool_result' it receives is too large for the context window, causing subsequent calls to truncate. What is the most effective architectural adjustment to resolve this?
5An orchestrator is designed to handle 'Parallel Tool Use' for a travel booking agent. Which TWO conditions must be met for the architect to successfully implement this feature?
6In the context of agentic orchestration, what is the primary purpose of a 'ReAct' (Reason-Act) loop?
7An architect is evaluating an agent's performance using 'Trajectory Evaluation.' Which TWO aspects of the agent's behavior are most effectively measured by this method?
8An agent is designed to summarize long documents by first splitting them into chunks and then calling a 'summarize_chunk' tool for each part. What is the most efficient orchestration strategy to minimize the total time taken for this task?
9A security audit of an agentic architecture reveals a risk of 'Indirect Prompt Injection' where the agent reads a malicious email and then uses its 'delete_account' tool on itself. Which architectural change best mitigates this risk?
10An agentic architecture uses Claude to resolve customer support tickets by invoking database lookup and refund tools. During execution, the model frequently makes redundant API calls for the same user ID within a single turn. Which orchestration pattern best resolves this issue while maintaining deterministic behavior?
11An architect is designing a system where Claude must handle complex multi-step customer inquiries. The current implementation uses a single prompt, but it often fails on high-reasoning tasks. Which agentic pattern should be implemented to allow a primary model to decompose the user's request into sub-tasks for specialized worker agents?
12When designing an autonomous agent loop using Claude 3.5 Sonnet that involves tool use, which TWO strategies are most effective for preventing the agent from entering an infinite loop when a tool returns an error?
13Refer to the exhibit. An architect is reviewing the tool-use integration for a custom Claude agent. What is the primary purpose of setting the 'is_error' field to true in this specific API payload?
14An architect is optimizing a multi-agent system for a legal firm. The system uses a 'Router' agent followed by several 'Specialist' agents. Which THREE techniques will most effectively minimize latency in this specific architecture?
15Refer to the exhibit showing a configuration for Claude 3.5 Sonnet. Which architectural requirement is strictly necessary when implementing this 'computer_20241022' tool in a production environment?
16An organization is deploying an agentic system that must maintain long-running conversations over several days, involving hundreds of tool calls. What is the most critical architectural consideration for managing the 'messages' array to prevent exceeding Claude's context window?
17Refer to the exhibit. In a parallel tool-use scenario where Claude generates three such blocks in a single response, how should the orchestration layer handle the 'tool_result' messages to ensure Claude can correctly process the outputs?
18An architect is building an agent that uses a 'Plan-and-Execute' pattern. The agent first generates a multi-step plan, then executes each step. During execution, the agent discovers that Step 2 is impossible due to a missing API key. How should the agent be designed to handle this?
19An architect is evaluating the use of 'Prompt Caching' within a complex Orchestrator-Workers agentic loop. Which TWO benefits are most significant for this specific use case?
20A developer wants to build a simple agent that can answer questions about a company's internal documentation. The developer implements a 'Router' that chooses between a 'Search' tool and a 'Direct Answer' prompt. What is the primary advantage of this router-based approach over a single prompt?
21When implementing the 'Evaluator-Optimizer' pattern for code generation with Claude 3.5 Sonnet, which TWO components are essential for the feedback loop to be effective?
22A large-scale agentic system uses Claude to manage a supply chain. The agent must interact with a legacy SQL database, a modern REST API, and a real-time streaming service. What is the most robust way to handle tool definitions in this heterogeneous environment?
23An architect notices that an agent using Claude 3.5 Sonnet frequently 'hallucinates' tool arguments when the tool schema is very large (over 50 parameters). What is the most effective architectural change to improve tool call accuracy?
24A financial services company is building a customer support assistant that needs to handle loan applications, account balance checks, and general FAQ queries. Which orchestration pattern should be used to minimize token usage while ensuring that the model uses specific system instructions for each distinct task?
25An architect is designing an agentic system to handle complex software engineering tasks. The system must break down a high-level requirement into specific coding milestones, assign those milestones to specialized sub-agents, and synthesize the results. Which orchestration strategy is most appropriate?
26When implementing a Tool Use (function calling) loop with Claude, which TWO steps are required for the application to successfully complete a multi-step task involving external data?
27Refer to the exhibit. What is the correct next step for the application developer to ensure the agent continues its task?
28In an Evaluator-Optimizer agentic workflow, which THREE components are essential for creating a successful iterative refinement loop for high-quality content generation?
29An agentic system is designed to use a search tool to find information. During testing, the agent occasionally gets stuck in a loop, repeatedly calling the same search tool with the same parameters. What is the most effective way to prevent this behavior programmatically?
30You are building an agent that needs to summarize 50 different documents. Which orchestration approach provides the best balance between speed and the model's ability to cross-reference information between all documents?
31Which TWO practices are recommended when defining tool schemas (JSON descriptions) for an agentic system to improve reliability and reduce errors?
32Where should the general 'rules of engagement' for an agent (such as tone, persona, and high-level safety constraints) be placed in an Anthropic API call?
33In a high-stakes agentic workflow, such as one that executes financial trades, why is 'Human-in-the-Loop' (HITL) considered a critical architectural component?
34Which 'stop_reason' does the Anthropic API return when the model has finished its current thought process and is waiting for the user to provide the results of a tool execution?
35What is the recommended temperature setting for an agent that primarily uses tools to perform data extraction and API calls?
36An architect is designing an agentic workflow on the Anthropic Messages API where a single Claude model must first break a user goal into ordered subtasks, then execute each subtask with tools, and finally synthesize a final answer. The architect wants to minimize the number of API round trips while still giving the model a chance to observe each tool result before deciding the next action. Which orchestration approach best satisfies these requirements?
37An architect is building a Claude-based agent that must reconcile invoice disputes. The orchestrator runs a single loop where Claude may emit several tool_use blocks in one assistant turn: one block calls lookup_invoice, another calls fetch_payment_history, and a third calls get_contract_terms. A junior developer proposes executing the first tool_use block, feeding its tool_result back to Claude, and only then deciding whether to run the other two. What is the architecturally correct way for the orchestrator to handle this turn?
38A travel-booking agent built on the Claude Messages API handles multi-turn conversations. During a session, the agent calls a search_flights tool and receives results. On the next user turn the agent suddenly invents a flight that was never returned by any tool. The architect reviews the request payload and sees that the prior assistant turn containing the tool_use block and the following user turn containing the tool_result block were both sent back to the API. What is the most likely cause of the hallucinated flight?
39An architect is building a Claude agent that must extract structured data from scanned invoices and then call an ERP API for each invoice. During early testing, the agent occasionally calls the ERP API with incomplete fields because it invents missing values. The architect wants to prevent these hallucinated field values from reaching the ERP. Which approach best addresses this requirement using the Anthropic Messages API?
40A team runs a Claude agent that orchestrates a multi-step data pipeline. The agent uses the Messages API and passes the full conversation history on every request. After the history grows beyond roughly 180,000 tokens, the agent begins to fail intermittently with context length errors even though the configured model supports 200,000 tokens. Logs show the agent appends each tool result verbatim, including large JSON payloads from a database tool. Which change best resolves the failures while preserving the agent's ability to reason over prior steps?
41A claims-processing agent must extract structured fields from scanned documents, then decide whether to auto-approve, request more evidence, or escalate to a human. The architect wants deterministic, auditable control over which branch executes rather than letting Claude narrate the decision. Which TWO design choices best achieve that? (Choose two.)
42You are building a customer-support agent with the Anthropic Messages API. The agent must call a `get_order_status` tool, and the result is required before it can decide whether to issue a refund. The tool sometimes takes several seconds to respond. Which approach best ensures the agent does not prematurely finalize its answer before the tool result is available?
43An architect is building a customer-support agent that calls a `lookup_order` tool. The tool occasionally returns HTTP 503 errors during peak hours. The architect wants the agent to recover from transient failures without human intervention. Which approach best fits Claude's agentic tool-use loop?
44A support agent running on the Claude Messages API must look up a customer's order status before answering. The developer writes a single request that includes the get_order_status tool definition and the user's question, then expects the model to return the order status directly in its first response. What actually happens, and what must the developer implement?
45An architect is designing a Claude agent that must autonomously handle customer disputes end to end. The agent may issue refunds up to $200, but any refund above that amount must be approved by a human reviewer. The system must not block the entire workflow while waiting for approval. Which orchestration pattern best satisfies these constraints?
46An agent that reconciles invoices is failing intermittently. Logs show it sometimes calls `approve_payment` before `verify_vendor`, even though the system prompt lists `verify_vendor` first. The agent has 12 tools available. Which change most directly reduces this ordering violation?
47An architect is designing a research agent whose context window is filling with raw web page HTML. The agent must keep working across dozens of sources without exceeding the context limit. Which architecture best preserves reasoning quality over a long session?
48A research agent must query three independent sources (a database, a web search API, and a file store) to assemble a report. Latency matters, and the sources do not depend on each other. Which orchestration pattern should you use with Claude's tool use to minimize wall-clock time?
49An architect is configuring a Claude agent that must decide when to search an internal knowledge base versus when to answer directly from its own knowledge. The team wants the agent to retrieve only when the question references internal policies, products, or procedures. Which configuration most directly supports this decision?
50An orchestration layer runs a research agent that can call a web_search tool and a summarize_document tool. The architect wants the agent to keep iterating until it has gathered enough sources, but must cap total model calls and total tool executions to control cost and prevent runaway loops. Which approach best enforces those caps while preserving the agent's ability to decide when it is done?
51A document-processing agent extracts structured fields from scanned contracts and must return a JSON object matching a strict schema. The agent also calls an OCR tool when text is unreadable. Which TWO design choices best ensure reliable, schema-valid output? (Choose two.)
52An architect is designing a Claude agent that must coordinate a research subagent and a summarization subagent to produce a market brief. The research subagent gathers sources, and the summarization subagent condenses them. The architect wants to avoid the orchestrator losing track of which subagent produced which artifact across many turns. Which TWO architectural practices best support reliable artifact tracking and handoff? (Choose two.)
53A team wants Claude to call an internal `create_invoice` tool only after the user has explicitly approved the line items. Which mechanism should the architect use to enforce this gate?
54A document-triage agent uses Claude with a read_attachment tool. During testing, the agent calls read_attachment for a PDF, receives a tool_result containing the extracted text, and then immediately calls read_attachment again for the same file id, repeating several times until the orchestrator stops it. The tool itself works correctly and returns valid content each time. What is the most effective architectural fix?
55A support agent has two tools: `search_knowledge_base` and `escalate_ticket`. For simple questions it should answer directly; for unresolved issues it should escalate. Which `tool_choice` setting best supports this behavior?
56An agent must classify incoming support tickets and then route each to one of four specialized sub-agents. The architect notices the router sometimes dispatches the same ticket to two sub-agents. Which design change most directly prevents duplicate dispatch?
57A production agent uses the Claude Messages API with tool use to manage cloud resources. During a run, Claude returns a response whose stop_reason is "tool_use" containing three tool_use content blocks that are fully independent of one another. The orchestrator loops and processes them one at a time, executing each tool and sending a separate user message containing a single tool_result block before each next call. Runs that used to complete in seconds now take minutes and occasionally exceed the request timeout. What is the most likely architectural cause of the slowdown?
58An architect is hardening an agent that uses tools to read internal wikis and send Slack messages. Which TWO practices most directly reduce the risk that untrusted wiki content causes the agent to take unauthorized actions? (Choose two.)
59A document-processing agent reads contracts and must extract a list of payment milestones, each with an amount, currency, and due date. Downstream billing systems cannot tolerate hallucinated fields, and the extraction must be validated programmatically before anything is written to the database. The architect wants Claude to emit data that can be parsed and checked with certainty rather than relying on post-hoc regex over prose. Which approach best satisfies this requirement?
60A research agent must answer questions that require looking up several unrelated facts, such as the population of three different cities. The architect observes that Claude frequently emits one search tool call, waits for the result, then emits the next, repeating until all facts are gathered. This serial pattern roughly triples latency versus what the workload should allow. Which change most directly enables the model to request all three independent lookups in a single assistant turn?
61An architect is designing a multi-step research agent that must gather facts, draft a report, and revise it based on a critique. The team wants to reduce the blast radius of mistakes and keep each stage's context focused, rather than letting one long conversation accumulate every intermediate result. Which TWO design choices best support this goal? (Choose two.)
Be able to wire a ReAct loop: send tools, read stop_reason, execute requested tools, and return every tool_result with its tool_use_id before the next turn. The critical thing is correctly pairing parallel tool results and detecting repeated calls to stop runaway loops.
The Courseiva CCAR-F question bank contains 61 questions in the Agentic Architecture and Orchestration domain. Click any question to see the full explanation and answer breakdown.
Start with a 10-question focused session to identify your baseline accuracy in this domain. Read every explanation — even for questions you answer correctly — to understand the reasoning. Once you score consistently above 80%, move to a 20–30 question session to confirm depth before moving to the next domain.
Yes — the session launcher on this page draws questions exclusively from the Agentic Architecture and Orchestration domain. Choose 10, 20, 30, or 50 questions for a focused session, or click individual questions to review them one by one.
Save your results, see per-domain analytics, and get readiness scores — free, for every certification.
Sign Up FreeFree forever · Every certification included