Courseiva

Claude Certified Architect - Professional (CCAR-P) — Questions 1–75

262 questions total · 4pages · All types, answers revealed

Page 1 of 4

Page 2
1
MCQmedium

You are the lead architect on a Claude-powered internal knowledge assistant. Six weeks into the pilot, the executive sponsor asks for a one-page status update that will be forwarded to the CFO, who has never attended a demo. Which artifact should you produce?

A.A live 30-minute demo for the CFO so she can see the assistant answer real questions.
B.The full evaluation notebook with prompt variants, scoring rubrics, and per-case results so the CFO can audit the methodology.
C.A one-page executive brief covering pilot objectives, measured adoption and quality metrics, run-rate cost, top risks, and the decision requested.
D.A forward of the weekly engineering standup notes with a short cover message summarizing the highlights.
AnswerC

An executive brief is forwardable, self-contained, and frames the pilot in the CFO's language: outcomes, spend, risk, and a specific ask. Because the CFO has no demo context, the brief must carry the narrative alone, which is exactly what this artifact does while respecting the sponsor's one-page constraint.

Why this answer

A CFO with no demo context needs a forwardable artifact that stands alone: objectives, evidence of value, cost, risk, and a clear ask. The executive brief is the only option that is self-contained, brief, and framed for a financial audience, satisfying both the sponsor's request and the CFO's decision needs.

Exam trap

The trap here is assuming a more technical or more interactive artifact signals rigor, when executive audiences actually need a short, self-contained, decision-oriented document they can read without you in the room.

2
MCQmedium

A project manager is overseeing a multi-phase implementation of Claude-based automated workflows. During the transition to production, the CISO expresses concerns regarding data privacy. Which strategy best addresses the stakeholder’s requirements while maintaining project momentum?

A.Present the technical documentation of the Claude model architecture to demonstrate overall robustness.
B.Delay the deployment until the CISO completes an independent audit of the third-party infrastructure.
C.Initiate a collaborative risk assessment focusing on data residency, encryption, and API access controls.
D.Ask the legal department to provide a formal exemption for the project to bypass standard security reviews.
AnswerC

Collaborative risk assessment directly aligns with the CISO's mandate to ensure security and compliance. By focusing on tangible controls like residency and encryption, you provide concrete evidence that the system complies with organizational standards, which facilitates an informed decision-making process and fosters institutional confidence in the project.

Why this answer

Engaging the CISO early through a structured Data Protection Impact Assessment (DPIA) bridges the gap between technical implementation and governance requirements. By proactively addressing security controls and data handling procedures within the Claude API architecture, you satisfy regulatory scrutiny and prevent late-stage blockers. This approach ensures that privacy is treated as a core design principle rather than an afterthought, which is crucial for building organizational trust during enterprise AI deployment.

Exam trap

Candidates often suggest 'bypassing the CISO' or 'delaying the launch'. Both are poor strategies; proactive collaboration is the only way to ensure security alignment without stalling the project.

3
MCQeasy

Which mechanism best facilitates secure, ephemeral access to Anthropic API keys for developers within a containerized CI environment?

A.Hardcoding the API keys in the source code repository with encryption.
B.Storing API keys in a local .env file on every developer workstation.
C.Using a secrets management service to inject keys as environment variables at runtime.
D.Distributing shared API keys via a secure internal messaging channel.
AnswerC

Injecting secrets as environment variables at runtime is a best practice that keeps sensitive information out of the repository. This method supports automated rotation and granular access control, ensuring that only authorized services can retrieve the credentials, thereby enhancing security and reducing the operational burden on the development team.

Why this answer

Utilizing a secrets management service such as HashiCorp Vault or AWS Secrets Manager allows for the injection of short-lived credentials into the environment. This minimizes the risk of long-lived key exposure in logs or source control. By automating secret rotation and access control, organizations can ensure developer productivity remains high without compromising the overall security posture of the infrastructure.

Exam trap

Candidates suggest hardcoding API keys in configuration files or embedding them directly into source control repositories.

4
Multi-Selectmedium

A developer enablement team is building an internal prompt playground so engineers can iterate on Claude prompts without writing API code. They want the playground to reflect production behavior and avoid surprising cost overruns. (Choose two.)

Select 2 answers
A.Disable streaming in the playground to reduce token consumption.
B.Require engineers to file a ticket for every prompt test so usage is reviewed before execution.
C.Give every engineer an unlimited personal API key so experimentation is never blocked.
D.Pin the playground to the same model version and system prompt that production uses, so experiments reflect real behavior.
E.Enforce per-user token quotas and surface usage dashboards so experimentation stays within budget.
AnswersD, E

Experiments only transfer to production if the model version and system prompt match. A playground running a different model or a divergent system prompt produces results that mislead engineers and cause rework. Pinning both keeps the feedback loop honest and makes prompt changes meaningful. It also prevents the common failure where a prompt succeeds in the playground but behaves differently once deployed.

Why this answer

A playground is only useful if its results predict production, so the model version and system prompt must match what ships. Cost safety comes from automated controls: per-user quotas and usage dashboards. Together these give engineers fast, trustworthy iteration without exposing the organization to unbounded spend.

Unlimited keys, disabled streaming, and ticket-gated testing each fail either the parity goal or the velocity goal.

Exam trap

The trap here is optimizing for frictionless access with unlimited keys while ignoring that unbounded experimentation is the exact cost risk the team is trying to avoid.

5
Multi-Selectmedium

Which THREE practices assist in maintaining a robust observability strategy for Anthropic API usage?

Select 3 answers
A.Log the full raw content of every user prompt and model response.
B.Track token usage per request to monitor cost and model efficiency.
C.Monitor request latency distributions to identify performance bottlenecks.
D.Implement structured logging for error codes and request IDs.
E.Rely solely on standard HTTP status codes for all operational monitoring.
AnswersB, C, D

Tracking token usage is the most important metric for cost management and architectural planning. It allows teams to identify high-cost requests and optimize prompt engineering or model choice. This visibility is essential for operational enablement, ensuring that developers can monitor their budget impact in real-time as they iterate.

Why this answer

Observability is critical for identifying bottlenecks and managing costs. By tracking token usage, latency, and error rates, teams gain the visibility needed to optimize performance. Integrating this data into existing monitoring stacks enables proactive alerting and trend analysis.

These practices empower developers to debug issues quickly and make data-driven decisions about infrastructure and model selection, which is essential for operational excellence in AI deployments.

Exam trap

Test-takers sometimes select metrics focused purely on business revenue rather than technical operational indicators like token usage, latency distributions, and error logs.

6
MCQmedium

An organization is conducting a risk assessment for an AI application. They are concerned about 'model drift' over time. Which governance action is most appropriate to manage this risk?

A.Increase the system prompt complexity.
B.Conduct periodic evaluation against a golden dataset.
C.Switch to a different model provider immediately.
D.Automate user feedback loops for real-time retraining.
AnswerB

A golden dataset provides a consistent benchmark to measure the model's performance over time. Comparing current outputs against this baseline allows organizations to quantify drift, identify specific areas of failure, and maintain quality control. This is the most effective way to ensure long-term stability in AI production environments.

Why this answer

Model drift refers to the degradation of model performance as the environment or data distribution changes. Periodic re-evaluation against a baseline 'golden dataset' is a standard governance practice to ensure the model remains reliable. This proactive monitoring allows teams to detect performance drops, adjust system prompts, or re-train/update the model, ensuring that the AI remains safe and effective for its original intended use case.

Exam trap

Candidates often mistake model drift for a security breach or a training data issue. They focus on retraining the model immediately rather than implementing a monitoring process to detect the degradation first.

7
Multi-Selecthard

Your organization is scaling its use of Anthropic API across multiple business units. Which TWO actions should you take to ensure effective lifecycle management and communication of service changes to stakeholders?

Select 2 answers
A.Mandate that all business units use the latest 'claude-3-opus-latest' alias to ensure they always receive the most advanced features.
B.Implement a formal versioning strategy using explicit model IDs and maintain a registry of which services utilize which version.
C.Automate the deprecation of older models by silently switching traffic to new versions to reduce infrastructure overhead.
D.Develop a communication plan that outlines update schedules, deprecation timelines, and required testing windows for all internal consumers.
E.Allow each business unit to independently choose whether to use Anthropic or a competing provider without centralized architectural oversight.
AnswersB, D

Explicit versioning is critical for auditability and risk management. By tracking model versions per service, you gain the ability to rollback instantly if a model update causes regressions. This transparency allows stakeholders to understand exactly what version of the technology is driving their business processes at any time.

Why this answer

Managing an AI lifecycle at scale requires both technical monitoring and proactive communication. By implementing version control for model deployments and establishing a standardized notification pipeline, you ensure stakeholders are prepared for updates. This structured approach prevents downtime and surprises, allowing business units to test new model versions in non-production environments before wide adoption, ensuring continuity and alignment with organizational stability goals.

Exam trap

Candidates often focus on individual model performance, ignoring the need for a global registry and versioning strategy that allows multiple business units to coordinate updates effectively.

8
Multi-Selecthard

Which TWO security measures are most effective at preventing 'Prompt Injection' attacks that target the arguments of tools used by an agent?

Select 2 answers
A.Adding 'Ignore all previous instructions' to the system prompt
B.Strict JSON schema validation for all tool inputs
C.Filtering the assistant's output for the word 'password'
D.Executing tools in a sandboxed, ephemeral environment
E.Using a secondary model to re-write every user query
AnswersB, D

By enforcing a rigid schema, the orchestrator ensures that the arguments passed to the tool conform to expected types and formats. This prevents attackers from injecting malicious scripts or unexpected commands into fields that should only contain simple data, such as numeric IDs or pre-defined string constants.

Why this answer

Prompt injection in tool arguments occurs when an agent processes untrusted data and passes it into a tool call that executes logic. Validating inputs against a strict schema and running the tool in an isolated environment are the primary defenses, ensuring that even if the agent is misled, the impact is contained.

Exam trap

Candidates often rely on 'system prompt instructions' to tell the model not to be hacked. This is ineffective against sophisticated prompt injection that bypasses textual constraints to manipulate tool arguments.

9
MCQhard

A bank is deploying a Claude agent that can call internal tools to move funds between accounts. Risk leadership wants a control that limits the blast radius if the agent is manipulated into performing unauthorized transfers. Which control best addresses this requirement?

A.Route all agent traffic through a proxy that logs every tool invocation and alerts the security operations center after transfers complete.
B.Enforce authorization and transaction limits in the downstream banking APIs the agent calls, independent of anything the model outputs.
C.Add a system prompt instructing the model to never perform transfers above a defined threshold or to accounts not previously seen.
D.Increase the model's temperature to zero so its responses become fully deterministic and cannot be manipulated.
AnswerB

This is correct because placing authorization, per-transaction caps, and velocity limits in the downstream systems means the model's output can never exceed what the API permits, regardless of manipulation. The blast radius is bounded by deterministic server-side policy rather than by model behavior, which is exactly the defense-in-depth posture risk leadership is requesting for a high-impact tool.

Why this answer

When an agent can take consequential actions, enforcement must live outside the model in the systems that hold authority. Server-side authorization, per-transaction caps, and velocity limits ensure that even a fully manipulated agent cannot exceed policy, because the downstream API rejects anything outside its rules. Prompt instructions, sampling settings, and post-hoc monitoring cannot provide that hard boundary.

Exam trap

The trap here is treating a strong system prompt or deterministic sampling as a security boundary when only the downstream authorization layer can actually enforce limits.

10
MCQhard

A developer support team receives repeated reports that Claude responses in an internal tool are truncated mid-sentence. Logs show the stop_reason value is max_tokens on most affected calls. Which change most directly resolves the truncation while preserving response quality?

A.Set temperature to zero so the model produces shorter, more deterministic completions.
B.Enable streaming so partial responses are delivered before the token limit is reached.
C.Add a system prompt instruction telling the model to always finish its sentences.
D.Increase the max_tokens parameter on the affected requests and verify the model's context window can accommodate input plus output.
AnswerD

A stop_reason of max_tokens means generation stopped because the output limit was reached, not because the model finished. Raising max_tokens gives the model room to complete its response, provided the combined input and output still fit within the model's context window. This directly addresses the observed cause and preserves quality because the model continues naturally rather than being forced to compress.

Why this answer

The stop_reason value is the authoritative signal: max_tokens means the output ceiling, not the model, ended the response. Raising max_tokens lets the model finish, and checking that input plus output fit the context window prevents a new failure mode. Sampling settings, prompt wording, and streaming do not change how many tokens the model may emit, so they cannot resolve this truncation.

Exam trap

The trap here is treating truncation as a prompt-quality problem when the API is explicitly reporting a token-limit stop.

11
MCQhard

A developer is building an application that needs to use multiple models (e.g., Haiku for speed, Sonnet for quality). What is the best pattern to handle model selection dynamically?

A.Hardcode the model selection logic into every individual service's controller.
B.Implement a central Model Router service that selects models based on request metadata.
C.Only use the highest quality model for all requests to ensure consistency.
D.Ask the user to manually select the model from a dropdown menu in the UI.
AnswerB

A Model Router provides a centralized point of control for selecting the appropriate model based on criteria like latency requirements or task complexity. This decoupling allows developers to optimize performance and costs dynamically without re-deploying individual application services, significantly improving operational agility and ease of experimentation.

Why this answer

Implementing a Model Router pattern allows for clean separation between business logic and model configuration. By defining a routing layer, developers can easily change model versions or swap models based on context (e.g., latency budget, task complexity) without modifying the core application code. This architectural pattern facilitates A/B testing and performance optimization, which are critical for maintaining developer productivity in complex, multi-model production systems.

Exam trap

Many candidates mistakenly propose hardcoding model selection logic directly inside application modules, missing that a central Model Router service provides the necessary abstraction for dynamic swapping and maintenance.

12
MCQmedium

What is the primary benefit of using a 'System Prompt' to define an agent's persona and constraints compared to embedding these in the user message?

A.System prompts are cached globally, reducing latency to zero.
B.User messages can be easily ignored, while system prompts cannot.
C.It establishes a distinct boundary between the agent's identity and user intent.
D.It is the only way to enable tool-use capabilities.
AnswerC

Separating identity and constraints from user inputs creates a robust architecture. This ensures that the agent's core instructions remain constant regardless of the user's conversational turns. This separation is fundamental to maintaining agent integrity and prevents users from coercing the agent into behaviors that violate defined policies.

Why this answer

System prompts are treated with higher priority by the model and are less susceptible to 'jailbreaking' or prompt injection from user inputs. By defining the agent's core identity and behavioral constraints in the system message, you create a stable, authoritative boundary that defines the model's behavior, ensuring consistent adherence to safety and operational guidelines throughout a long-lived conversation.

Exam trap

Candidates often assume the benefit is purely about token savings or context length, missing the critical security and authoritative boundary benefits of system prompts over user messages.

13
MCQhard

When designing an agent that must perform highly sensitive operations (e.g., deleting records), what architectural pattern is mandatory?

A.Fine-tune the model to never delete records.
B.Implement a mandatory human-in-the-loop confirmation step.
C.Use a low-temperature setting for sensitive tool calls.
D.Log the action to a secure bucket after execution.
AnswerB

The HITL pattern forces an external, verifiable authorization layer into the workflow. This ensures that no destructive or sensitive action can proceed without an explicit, audit-trailed human approval, which is the gold standard for secure agentic systems dealing with sensitive backend operations and data integrity.

Why this answer

A 'Human-in-the-loop' (HITL) gate for sensitive actions is non-negotiable for enterprise safety. The agent generates the plan and the proposed tool call, but the system architecture must intercept this and require explicit human verification before execution. This prevents accidental data loss or malicious exploitation of the agent's capabilities, balancing autonomous productivity with critical risk management in production environments.

Exam trap

Candidates often suggest relying on model-level security or fine-tuning to prevent errors, failing to recognize that human-in-the-loop is the only mandatory safeguard for high-stakes, irreversible actions.

14
MCQmedium

You are the lead architect on a Claude-powered claims triage system that has been in production for six months. A newly appointed VP of Operations asks for a quarterly review of the system's lifecycle status. Which deliverable best satisfies this request while maintaining stakeholder alignment?

A.A lifecycle status report covering adoption metrics, cost trends, incident history, model version currency, and the next-quarter roadmap.
B.A raw export of the last 90 days of API request logs for the VP's team to analyze.
C.A revised statement of work that extends the engagement by another two quarters.
D.A model card update listing the current Claude model version and its evaluation benchmarks.
AnswerA

A lifecycle status report addresses the VP's request directly by consolidating operational, financial, and strategic indicators into one governance artifact. It shows whether the system is healthy, whether it is delivering value, and what decisions are coming, which is exactly what a lifecycle review at the executive level requires. It also creates a recurring cadence for stakeholder alignment.

Why this answer

The VP requested a lifecycle status review, which is a governance activity, not a technical artifact or a contracting action. Consolidating adoption, cost, incidents, version currency, and the roadmap into a single report gives the stakeholder the decision-ready picture they need and establishes a repeatable review cadence. Pointing to a model card, raw logs, or a statement of work extension each answers a different question than the one asked.

Exam trap

The trap here is assuming a technical artifact such as a model card or raw logs automatically doubles as an executive lifecycle status report.

15
MCQmedium

A company is scaling its Claude-powered applications globally. Which strategy best optimizes for both latency and cost?

A.Use the largest available Claude model for every API request globally.
B.Route complex tasks to Sonnet and simpler tasks to Haiku.
C.Cache all API responses indefinitely to eliminate future costs.
D.Deploy all applications in a single region to simplify infrastructure.
AnswerB

Right-sizing models based on task complexity is a highly effective optimization strategy. It reduces operational costs by leveraging smaller models for lightweight tasks while maintaining performance for complex reasoning. This architectural choice improves developer productivity by providing a balanced toolkit that addresses diverse use cases with optimal efficiency and speed.

Why this answer

Selecting the appropriate model based on task complexity (right-sizing) combined with regional deployment of application logic minimizes latency. By routing simpler tasks to smaller, faster models and reserving the most capable models for complex reasoning, the organization maximizes cost-efficiency. This operational strategy ensures that developers can build high-performance applications while staying within budget constraints, which is vital for long-term project sustainability and scalability.

Exam trap

Candidates frequently assume the most powerful model should be used for every task, ignoring cost-efficiency strategies like routing simpler tasks to smaller models.

16
Multi-Selecthard

A security architect is performing red-teaming on a new Claude-powered application. They are specifically testing for 'jailbreaking' attempts where a user tries to bypass safety filters by using roleplay or adversarial framing. Which TWO strategies are most effective for mitigating this specific risk at the architectural level?

Select 2 answers
A.Increasing the temperature parameter to 1.0
B.Implementing a robust, immutable System Prompt
C.Reducing the max_tokens limit for all users
D.Utilizing a separate 'Safety' instance of Claude for output validation
E.Switching from Claude 3.5 Sonnet to Claude 3 Haiku
AnswersB, D

A well-defined system prompt acts as a foundational governance layer that defines the model's persona and safety constraints. By explicitly instructing the model to reject roleplay attempts that violate safety policies, architects can significantly harden the application against common jailbreaking techniques that rely on tricking the model into ignoring its rules.

Why this answer

Mitigating adversarial attacks requires a multi-layered approach that combines model-native features with external validation. Using a strong system prompt sets clear boundaries that the model prioritizes, while implementing an independent moderation layer provides a final check on outputs. These strategies ensure that even if one layer is bypassed, the overall system remains resilient against malicious intent.

Exam trap

Candidates often select client-side or prompt-only solutions, assuming standard instructions are bulletproof. They forget that jailbreaking specifically targets and bypasses text-based prompts, requiring multi-layered architectural safeguards like independent validators.

17
MCQhard

An organization is deploying Claude to provide automated coding assistance. To manage the risk of generating insecure code or violating open-source licenses, which governance step is most effective?

A.Disabling the model's ability to output code blocks entirely.
B.Implementing a mandatory 'Human-in-the-Loop' review and automated SAST scanning.
C.Relying on Claude's internal safety training to prevent all insecure code generation.
D.Requiring all developers to use Claude only for writing documentation, not logic.
AnswerB

Static Application Security Testing (SAST) tools can automatically detect vulnerabilities in the code Claude generates. Combined with human review, this ensures that any AI-driven suggestions are vetted for security flaws and license compliance before being merged into the production codebase, providing a robust governance layer.

Why this answer

Risk management for AI-generated code requires a multi-layered approach. While the model is highly capable, it may occasionally suggest patterns that contain vulnerabilities or mimic copyrighted code. Integrating automated security scanning into the development lifecycle ensures that AI-generated artifacts meet the same standards as human-written code.

Exam trap

Candidates often rely solely on the model's internal safety training to prevent code vulnerabilities, forgetting that external developer workflows require active validation steps.

18
MCQmedium

Refer to the exhibit. The monitoring system logs these safety rejections. How should you communicate this to the product owner?

A.Tell them the system is broken and that we need to stop the development immediately.
B.Provide a report showing the effectiveness of safety guardrails in mitigating potential risk.
C.Keep the logs hidden to avoid making the project look like it has problems.
D.Ask the product owner to manually approve all output that the model generates.
AnswerB

Presenting the data as a validation of safety guardrails demonstrates professional competency and keeps the product owner informed about risk mitigation. It reframes the logs from a 'problem' to a 'success indicator,' which provides the product owner with confidence in the system's security posture and ongoing operational oversight.

Why this answer

This situation is an opportunity to show the product owner the value of the safety guardrails in action. Presenting the logs as evidence of success—that the system is working as intended to prevent potential violations—reassures the product owner while initiating a conversation about potential refinements. This proactive reporting builds trust in the security of the application and keeps the business stakeholder informed about the protective measures in place.

Exam trap

Candidates often treat safety rejections as system failures or bugs, failing to frame them to stakeholders as successful security measures that are actively protecting the organization from risk.

19
MCQmedium

During a project, a stakeholder requests a feature that violates your established safety guidelines. What is the most professional way to handle this?

A.Refuse the request immediately and do not discuss it further.
B.Implement the feature but add a disclaimer to the output.
C.Explain the safety implications and facilitate a session to brainstorm alternative, compliant ways to meet the underlying business need.
D.Ask the stakeholder to sign a waiver accepting responsibility for the potential safety issues.
AnswerC

This approach is the gold standard for professional conflict resolution in architecture. It respects the safety guidelines while acknowledging the stakeholder's business motivation. By facilitating a brainstorming session, you demonstrate value as a partner and problem-solver, ensuring that the final solution is both safe and aligned with the business goals.

Why this answer

Handling requests that violate safety guidelines requires firm adherence to architectural principles, coupled with a collaborative approach to finding alternatives. By explaining the rationale behind the safety rules and then engaging in a discovery process for acceptable use cases, you maintain the integrity of your security posture while still helping the stakeholder achieve their underlying business objective. This demonstrates leadership and alignment with corporate compliance while avoiding unnecessary conflict.

Exam trap

Candidates often choose to blindly reject the request without offering alternatives, or conversely, compromise safety guidelines to satisfy stakeholder demands.

20
Multi-Selectmedium

A developer enablement group is standing up a shared Claude integration library that dozens of internal teams will consume. They want to reduce duplicated work and prevent each team from re-implementing fragile request handling. Which TWO practices best improve developer productivity across those consuming teams? (Choose two.)

Select 2 answers
A.Give every team direct access to the raw HTTP layer and encourage them to call the API without the shared library.
B.Keep the client API undocumented so teams are forced to read the source code and understand every internal detail.
C.Require every consuming team to write its own request wrapper so each can tune retries to its own latency profile.
D.Publish a typed client library that encapsulates authentication, retries, and error normalization behind a small, stable interface.
E.Ship runnable reference examples and integration tests that demonstrate the recommended patterns for common tasks.
AnswersD, E

A typed client centralizes the fragile parts of API integration, so consuming teams get consistent retry and error behavior without re-implementing it. Types catch misuse at compile time, and a small stable surface reduces the learning curve. This directly reduces duplicated effort and prevents each team from shipping subtly different request handling.

Why this answer

A typed client library with a small stable interface centralizes fragile request handling so teams do not each reinvent retries and error normalization. Runnable examples and integration tests lower the cost of adoption and document the recommended patterns. Together these reduce duplicated work and support questions, which is exactly what a developer enablement group is trying to achieve.

Exam trap

The trap here is assuming that more per-team customization improves productivity, when in practice duplicated request handling multiplies maintenance and creates inconsistent behavior across the organization.

21
MCQhard

Which architectural pattern is best suited for long-running, multi-step agentic workflows that require human-in-the-loop intervention?

A.Monolithic synchronous execution from the user's browser.
B.Stateless API calls with all context re-sent in every request.
C.State machine-driven orchestration with persistent task queues.
D.Client-side polling of the Anthropic API directly from the frontend.
AnswerC

State machines allow for robust tracking of long-running processes. By using queues to manage tasks, the system can pause, wait for external input, and reliably resume. This provides the durability required for complex agentic workflows, making it easier for developers to build, test, and maintain sophisticated AI-driven business processes.

Why this answer

The 'Orchestrator-Worker' pattern with a state machine is ideal for complex workflows. By saving the state of the agent at each step in a database, the system can pause for human review and resume seamlessly once input is received. This pattern provides the necessary durability and auditability for production applications, ensuring that developer productivity is not hampered by fragile, monolithic processes that fail on restart.

Exam trap

Many candidates choose simple message queues or naive retry logic, failing to recognize that state machine orchestration is specifically required to maintain context across human-in-the-loop pause points.

22
MCQeasy

A team is building an agent that must call a payment provider's API. The API requires an idempotency key on every charge request, and the agent may retry a charge if a network error occurs mid-call. Which design correctly preserves financial correctness when retries happen?

A.Generate a fresh idempotency key for each retry attempt so the provider can distinguish a retry from a genuinely new charge.
B.Disable automatic retries for charge calls and surface every network error to a human operator for manual reconciliation.
C.Omit the idempotency key and instead check the account balance before each retry to see whether the charge already landed.
D.Derive the idempotency key deterministically from the logical charge (for example, a hash of order ID and amount) and reuse that same key across all retries of that charge.
AnswerD

A deterministic key ties every retry to the same logical charge, so the provider deduplicates repeats and returns the original result. Hashing stable fields such as order ID and amount means the key survives process restarts and agent replanning. This is the standard way to make an at-least-once retry loop behave as exactly-once at the provider boundary.

Why this answer

The idempotency key is the provider's contract for recognizing repeated attempts as one logical charge. Deriving it deterministically from stable charge attributes means every retry, even after a process restart or agent replan, presents the same key, so the provider returns the original outcome instead of charging again. This preserves correctness under at-least-once retry semantics without sacrificing resilience.

Exam trap

The trap here is assuming each retry needs a unique key to be distinguishable, when uniqueness per attempt is precisely what causes duplicate charges and defeats the idempotency contract.

23
MCQmedium

An organization is deploying a customer-facing chatbot using Claude 3.5 Sonnet and needs to ensure the model adheres to ethical guidelines without relying solely on manual moderation. Which core Anthropic safety framework is primarily responsible for the model's ability to self-correct based on a predefined set of principles during its training phase?

A.Reinforcement Learning from Human Feedback (RLHF)
B.Retrieval-Augmented Generation (RAG) Filtering
C.Constitutional AI
D.Differential Privacy Injection
AnswerC

Constitutional AI applies a specific list of rules that the model uses to evaluate its own outputs during the reinforcement learning phase. This method ensures that the model adheres to ethical guidelines and safety standards without requiring constant human oversight, making it a highly scalable and reliable solution for enterprise-grade deployments.

Why this answer

Claude's safety is built on Constitutional AI, which uses a set of principles to guide the model's self-improvement during training. This approach reduces the need for human-annotated safety data and allows for more transparent and steerable AI behavior compared to traditional RLHF. Architects must understand how this foundation impacts model responses to ambiguous or harmful prompts in enterprise production environments.

Exam trap

Candidates often confuse Constitutional AI with RLHF or fine-tuning. They miss the distinction that Constitutional AI is a specific training methodology using principles for self-correction.

24
Multi-Selectmedium

A hospital network is drafting its AI risk register for a Claude-based discharge-summary assistant. The governance lead wants entries that describe residual risk after existing controls are applied, and that can be assigned an owner and a review cadence. Which TWO characteristics must each risk register entry have to meet this standard? (Choose two.)

Select 2 answers
A.A verbatim copy of the vendor's model card and system prompt.
B.A list of every employee who has ever accessed the assistant.
C.A residual risk rating that reflects the effect of the controls already in place.
D.A projected cost saving attributed to deploying the assistant.
E.A named accountable owner and a defined review interval.
AnswersC, E

The governance lead explicitly asked for residual risk, meaning the exposure remaining after existing mitigations. Recording only inherent risk overstates exposure and misdirects investment; recording only that a control exists hides whether it is effective. A residual rating ties the register to the actual decision the network faces about accepting or further reducing exposure.

Why this answer

An operational risk register entry must state the exposure that remains after controls and must have someone accountable for watching it on a schedule. Residual rating captures what is actually at stake post-mitigation, while owner and review interval ensure the entry is revisited and acted upon as the assistant and its clinical context change.

Exam trap

The trap here is padding risk entries with supporting evidence and business-case data, which feels thorough but leaves the register without the ownership and residual-exposure statements that make it actionable.

25
MCQmedium

Which architectural approach is best for handling an agent's failure to retrieve information from a database tool?

A.Simply return the raw error message to the end user.
B.Configure the agent to automatically retry the exact same query 5 times.
C.Provide the error back to the agent as a new observation for re-planning.
D.Default to a hard-coded fallback value to satisfy the user.
AnswerC

Treating an error as an observation is a key principle of agentic architecture. By letting the agent 'see' what went wrong, it can apply its reasoning to correct the error, such as by broadening a query scope or sanitizing an input, leading to much higher success rates.

Why this answer

An 'error-handling wrapper' that catches the tool failure and feeds the error back into the agent's reasoning loop is essential. This allows the model to perform a 'self-correction' phase—perhaps by reformulating the query, searching a different table, or asking the user for clarification. This turns an error into a new data point, making the system significantly more robust and helpful than just reporting a failure.

Exam trap

Candidates often suggest showing the error to the user or simply retrying the same query, failing to leverage the model's ability to self-correct based on feedback.

26
MCQmedium

A development team is integrating Claude into a high-throughput CI/CD pipeline and notices occasional 429 Too Many Requests errors. What is the most effective architectural approach to improve system reliability while maintaining developer speed?

A.Increase the concurrency limit in the AWS/GCP account quota settings immediately.
B.Implement a message queue with a worker pool using exponential backoff and jitter.
C.Modify the application to use synchronous HTTP calls without any retry logic.
D.Switch from Claude 3.5 Sonnet to Haiku to avoid all rate limits.
AnswerB

Decoupling the Claude API interaction from the main pipeline using a queue ensures durability. If a request fails, the worker can retry with backoff, ensuring that transient errors do not crash the pipeline. This pattern is essential for high-throughput systems to maintain operational stability and developer productivity.

Why this answer

Implementing an exponential backoff strategy with jitter in the application layer is the standard architectural pattern for handling rate limits in distributed systems. By staggering retries, the team prevents the 'thundering herd' effect, ensuring that requests are distributed more evenly over time. This approach increases the overall reliability of the pipeline and reduces manual intervention, which is critical for maintaining high velocity in automated workflows.

Exam trap

Candidates often select simple synchronous sleep timers or client-side request throttling, ignoring that robust high-throughput pipelines require message queues coupled with exponential backoff and jitter.

27
MCQmedium

Which approach best aligns with the principle of 'least privilege' when providing API access to internal teams?

A.Create a single organization-wide API key for all departments to share.
B.Use separate API keys for each project with specific rate limits and usage quotas.
C.Provide all developers with unrestricted access to the master organization API key.
D.Rotate all team keys daily to ensure the highest level of security.
AnswerB

Granular, project-specific keys allow for precise control and oversight. By enforcing usage quotas and rate limits, the organization can manage costs and limit the impact of any potential security breach to a single project, directly adhering to the principle of least privilege in a production environment.

Why this answer

Least privilege is best enforced by utilizing scoped API keys and rate limits mapped to specific project requirements. By isolating access, an architect ensures that a compromise in one department does not cascade across the entire organization. This structure allows for granular monitoring and easier revocation, forming a robust foundation for organizational security and minimizing the blast radius of any potential credential leakage or misuse.

Exam trap

Candidates often choose a single shared API key for simplicity, incorrectly assuming organizational convenience overrides the security necessity of granular, project-level scoping and rate limits.

28
MCQmedium

A developer is building a tool that lets engineers query an internal knowledge base through Claude. During testing, Claude sometimes invents plausible but nonexistent document titles when the retrieved context is thin. The team wants a systematic way to detect and reduce these hallucinations before the tool reaches general availability. Which approach is most appropriate?

A.Add an instruction telling the model to be accurate and never hallucinate, then rely on that instruction in production.
B.Raise the temperature setting so responses become more varied and the model is less likely to repeat a fabricated title.
C.Build an evaluation set of questions with known ground-truth answers, score responses for faithfulness to the retrieved context, and iterate on retrieval and prompting until scores meet a defined threshold.
D.Switch the knowledge base queries to return a fixed number of documents regardless of relevance, so the model always has something to cite.
AnswerC

A labeled evaluation set with faithfulness scoring turns a vague symptom into a measurable signal. Iterating on retrieval quality and prompting against that metric addresses the root cause, since thin context drives fabrication. A defined threshold provides an objective release gate, so the team can demonstrate improvement rather than relying on anecdotal spot checks.

Why this answer

Hallucinated titles are a grounding failure, so the fix is to measure faithfulness against retrieved context and improve retrieval and prompting where scores fall short. A labeled evaluation set with ground-truth answers makes the problem quantifiable, and a release threshold converts that measurement into a defensible go or no-go decision.

Exam trap

The trap here is reaching for a sampling parameter or a stern instruction to cure fabrication, when hallucination of this kind stems from insufficient or poorly ranked grounding context.

29
MCQmedium

A fintech company wants to use Claude to generate personalized financial advice for retail customers. The compliance team mandates that all AI-generated advice must be traceable to a specific model version and configuration for audit purposes. Which governance control best satisfies this requirement?

A.Implement a human review step where a compliance officer approves each piece of advice before it is sent.
B.Restrict API access to a whitelist of IP addresses and require multi-factor authentication.
C.Enable logging of all API requests and responses with model version and configuration metadata.
D.Use a single, static prompt template that never changes and is stored in version control.
AnswerC

This control directly provides an audit trail linking each piece of advice to the exact model version and configuration used. By capturing request and response data along with metadata, the company can reconstruct how any advice was generated, satisfying traceability requirements. It is the most direct and reliable method to meet the compliance mandate without altering the model's behavior.

Why this answer

The correct control is logging API requests and responses with model version and configuration metadata. This creates a detailed audit trail that links each output to the exact model version and settings used, enabling full traceability. Other options improve security or oversight but fail to provide the required technical record for auditing AI-generated financial advice.

Exam trap

The trap here is confusing security controls like IP whitelisting or human review with audit traceability, which specifically requires capturing model version and configuration in logs.

30
Multi-Selecthard

A platform team is standardizing how internal teams integrate Claude. Leadership wants faster onboarding, fewer production incidents, and clear accountability for cost. Which TWO practices best support these goals? (Choose two.)

Select 2 answers
A.Publish a golden-path reference implementation with sample prompts, error handling, and a checklist for going to production.
B.Route all Claude traffic through a single shared API key managed by the platform team to simplify billing.
C.Require every team to tag requests with a service identifier and workspace, and surface per-team token usage in a shared dashboard.
D.Let each team choose its own SDK, logging format, and retry strategy to maximize autonomy.
E.Mandate that all teams use the same model version and freeze upgrades for a year to reduce variability.
AnswersA, C

A golden-path reference gives new teams a working starting point with proven error handling and prompt patterns, which speeds onboarding and reduces incidents caused by common mistakes. The checklist makes production readiness explicit. It also spreads a consistent approach without forcing a heavy migration. This directly supports faster onboarding and fewer incidents.

Why this answer

A golden-path reference implementation accelerates onboarding and reduces incidents by giving teams proven patterns, while per-service tagging with a shared usage dashboard makes cost attributable and accountable. Freezing model versions, allowing full tooling autonomy, and centralizing one shared key each work against at least one of the stated goals, so they are not the right practices here.

Exam trap

The trap here is treating a single shared API key as a billing simplification, when it actually removes the per-team attribution that cost accountability depends on.

31
MCQhard

An agent uses a retrieval tool that returns the top 20 chunks for any query. In production, the agent frequently cites irrelevant chunks and sometimes misses the correct answer even when it is present in the corpus. The corpus contains documents with overlapping terminology. Which architectural change most improves answer grounding without increasing the number of retrieved chunks?

A.Embed the entire corpus into the system prompt so the model always has full context available.
B.Increase the retrieval count to 50 so the correct chunk is more likely to be included somewhere in the context.
C.Lower the model's temperature to zero so it sticks more closely to the retrieved text.
D.Add a re-ranking stage that scores the retrieved chunks against the query and passes only the highest-scoring subset to the model.
AnswerD

Re-ranking applies a more precise relevance model to the candidate set, so the chunks most likely to contain the answer are promoted and the noisy ones are dropped. This improves grounding without retrieving more chunks, directly addressing both irrelevant citations and missed answers. It preserves the retrieval tool's contract while adding a precision layer that overlapping terminology makes necessary.

Why this answer

The retrieval step is high-recall but low-precision, which is why irrelevant chunks are cited and correct ones are missed amid overlapping terminology. Adding a re-ranking stage that scores candidates against the query and passes only the top subset improves precision without retrieving more chunks, directly improving grounding and citation quality.

Exam trap

The trap here is assuming that more retrieved context or a lower temperature will fix grounding, when the actual defect is ranking precision within a fixed candidate set.

32
MCQhard

A software company's internal AI review board is defining escalation criteria for its Claude-powered support assistant. The board wants a rule that reliably routes the highest-consequence cases to human specialists rather than relying on the model's own confidence statements. Which escalation design best achieves this?

A.Route cases to specialists based on objective risk signals such as account tier, regulatory keywords, and prior complaint history.
B.Ask the assistant to flag any conversation it finds ambiguous and forward those to specialists for a second opinion.
C.Instruct the assistant to state a confidence percentage with each answer and escalate whenever it reports below ninety percent.
D.Sample five percent of all conversations at random for specialist review after the assistant has already replied.
AnswerA

Objective signals are observable before or independently of the model's output, so routing does not depend on the model judging its own reliability. High-value accounts, regulated topics, and repeat-complaint histories are exactly where errors carry the greatest consequence, making this a deterministic and auditable way to guarantee specialist review.

Why this answer

Reliable escalation must be driven by signals that exist independently of the model's self-assessment. Objective criteria such as account tier, regulatory keywords, and complaint history correlate directly with consequence and can be evaluated deterministically, guaranteeing that the cases the board cares most about reach a specialist before a reply is finalized.

Exam trap

The trap here is trusting the model's own confidence or ambiguity judgments as the routing trigger, when those signals are uncalibrated and can be high precisely when the answer is wrong.

33
MCQeasy

A public-sector agency must demonstrate to an external auditor that its Claude-based citizen inquiry assistant was operated in line with its approved safety policy throughout the prior fiscal year. The agency has no centralized record of which policy text was in force, when it changed, or who approved each change. Which governance practice should the agency institute first?

A.Establish version-controlled policy documents with recorded approvals and effective dates.
B.Publish a citizen-facing FAQ describing how the assistant works and what data it uses.
C.Deploy an additional monitoring dashboard that tracks assistant uptime and query volume.
D.Commission a penetration test of the assistant's public-facing endpoint.
AnswerA

The agency's core gap is knowing which policy was in force at any given time and who authorized it. Version control with approval records and effective dates creates that timeline, allowing the auditor to map any historical period to the exact policy text that governed it. Every other evidence request depends on this foundation being in place first.

Why this answer

Auditors reconstruct conformance by comparing what happened against the rules that were in force at the time, which requires an authoritative, dated policy history with named approvers. Version-controlled policies with effective dates supply that timeline; without it, no amount of monitoring, testing, or public communication can demonstrate year-long adherence.

Exam trap

The trap here is reaching for technical assurance activities when the actual deficiency is the absence of an authoritative record of which policy applied and when.

34
MCQmedium

An enterprise wants to minimize the risk of PII (Personally Identifiable Information) being processed by Claude while maintaining low latency. Which architectural approach provides the best balance of safety and performance?

A.Sending all data to a second LLM for PII scrubbing
B.Using a local PII detection script before the API call
C.Asking Claude to ignore all PII in the system prompt
D.Relying on the base model's default safety filters
AnswerB

A local detection script can quickly scan and redact PII before the data ever leaves the organization's controlled environment. This approach provides a high level of security by ensuring sensitive data is never sent to the API provider, while also maintaining the low latency required for production-grade applications.

Why this answer

Managing PII risk requires a proactive approach that stops sensitive data before it reaches the model. Using a local, lightweight regex or NLP-based scanner for PII detection allows for immediate filtering without the latency of a secondary LLM call. This ensures that the organization maintains its privacy standards while providing a fast, responsive experience for the end user.

Exam trap

Candidates often choose secondary LLM calls for PII detection because they assume AI is required for smart filtering, completely ignoring the severe latency penalty this introduces.

35
MCQhard

An agentic system is designed to run long-lived 'background' tasks that may take several days to complete. How should the architect manage the agent's state to ensure it can resume correctly after a system reboot?

A.Pass the entire history in every API call
B.Use a very long 'max_tokens' setting
C.Externalize state and history to a persistent store
D.Rely on the model's internal memory
AnswerC

By saving the messages array and any orchestrator-level metadata (like step progress or tool IDs) to a database, the system becomes resilient. After a reboot, the orchestrator can reload the state, identify the last completed action, and resume the loop without losing progress or context.

Why this answer

Long-lived agents cannot rely on in-memory state. To ensure durability, the entire conversation history and any relevant internal state (like current goal or pending sub-tasks) must be externalized to a persistent database. This allows any worker instance to reconstruct the agent's 'mind' and continue the task from the last successful turn.

Exam trap

Candidates often assume the agent's session memory is sufficient. They ignore the reality that long-lived background tasks require persistent storage to survive system restarts, crashes, or scaling events.

36
MCQhard

A financial services firm runs a Claude-powered agent that can call internal tools to move funds between accounts. Risk management wants a control that prevents the agent from executing a transfer above a threshold without human sign-off, and that remains effective even if the model is manipulated through injected content in a retrieved document. Which control best meets this requirement?

A.Lower the agent's temperature and restrict its tool list to a single transfer function with a fixed daily cap.
B.Enforce the threshold in the tool-execution layer so that any transfer exceeding the limit is rejected unless a human approval token is presented.
C.Add a system prompt rule stating that transfers above the threshold must be escalated to a human operator.
D.Run a second Claude instance as a reviewer that inspects each proposed transfer and vetoes suspicious ones.
AnswerB

Placing the limit in the layer that actually performs the transfer makes the control independent of model behavior, so a manipulated agent still cannot move funds above the threshold. Requiring a human approval token binds the exception to an accountable person. This is defense in depth: the model may propose, but the execution layer disposes, which is the only design that survives prompt injection.

Why this answer

When an agent can take consequential action, the authorization boundary must live outside the model, in the component that executes the action. Enforcing the threshold and requiring a human approval token at the tool layer means a manipulated model cannot exceed its authority, because the code performing the transfer refuses. Prompt rules, reviewer models, and sampling adjustments modify model behavior but cannot guarantee that a hijacked agent will decline a prohibited action.

Exam trap

The trap here is believing that a system prompt instruction or a second reviewing model constitutes an enforcement control, when only the component that executes the action can reliably refuse it.

37
Multi-Selectmedium

An enterprise client is concerned about the 'black box' nature of LLMs. Which THREE communication strategies should the architect use to build transparency and trust?

Select 3 answers
A.Present the results of red-teaming exercises to show how vulnerabilities are identified and patched.
B.Explain the use of citations and grounding techniques to reduce hallucination risks.
C.Promise stakeholders that the model will have a 100% accuracy rate for all inputs.
D.Detail the human-in-the-loop (HITL) workflows used to validate critical AI-generated outputs.
E.Avoid mentioning technical constraints to keep the stakeholder focused on the positive benefits.
AnswersA, B, D

Sharing red-teaming results demonstrates a proactive security posture and a commitment to rigorous testing. It helps stakeholders understand that the organization is actively looking for flaws and improving the system's resilience, which builds significant confidence in the robustness of the chosen AI solution for sensitive business applications.

Why this answer

Building trust with stakeholders requires moving beyond marketing claims to concrete evidence of system reliability and oversight. By demonstrating how the model reaches decisions, implementing monitoring to catch errors, and providing clear documentation on safety guardrails, the architect demystifies the technology. These strategies collectively address stakeholder anxiety by providing a tangible framework for oversight, control, and accountability within the automated system.

Exam trap

Test-takers frequently select vague marketing assurances or raw performance benchmarks, overlooking concrete operational transparency measures like red-teaming results, grounding, and HITL workflows.

38
Multi-Selectmedium

A large enterprise is setting up its governance framework for Anthropic API usage. Which THREE features provided by the Anthropic Console are essential for maintaining auditability and administrative control?

Select 3 answers
A.Audit Logs for API key usage and console activity
B.Direct access to the model's weight files
C.Single Sign-On (SSO) integration
D.Automatic prompt optimization for all users
E.Member roles and workspace permissions
AnswersA, C, E

Audit logs provide a detailed record of all actions taken within the console and by API keys, which is critical for security investigations and compliance audits. They allow administrators to see exactly when keys were created, used, or deleted, ensuring full accountability for the organization's AI resources.

Why this answer

Governance in an enterprise context requires tools that provide visibility into usage and restrict access to authorized personnel. Features like audit logs, SSO, and granular member roles allow administrators to track who is using the model and for what purpose. These tools are the backbone of a compliant AI strategy, ensuring that all activities are documented and secure.

Exam trap

Test-takers sometimes select operational runtime settings or prompt tuning features as governance tools, confusing general development features with administrative audit and control components.

39
MCQhard

A regulated healthcare client is reviewing your Claude-based patient-intake summarization tool. Their compliance officer asks how you will handle a scenario where the model produces a clinically inaccurate summary. Which response demonstrates the strongest lifecycle and stakeholder communication practice?

A.State that the tool is a decision-support aid and that all liability rests with the clinician who reviews the output.
B.Offer to switch to a smaller, faster model so that summaries are generated quickly enough for clinicians to catch errors in real time.
C.Guarantee that the model will not produce inaccurate summaries because Anthropic's safety training prevents harmful outputs.
D.Describe an incident response process with human-in-the-loop review, error logging, root-cause analysis, and a defined escalation path to the clinical team.
AnswerD

This response acknowledges that errors are possible and demonstrates a controlled, auditable process for detecting, containing, and learning from them. Human review before clinical use, structured logging for traceability, and a clear escalation path are exactly the controls a compliance officer needs to see. It also keeps the clinician accountable for the final decision, which aligns the AI system with existing clinical governance.

Why this answer

The strongest response to a compliance concern about inaccuracy is a documented control process: human review before clinical action, traceable logging, root-cause analysis, and a defined escalation path. This shows the system is governed rather than merely deployed. Guaranteeing perfect accuracy misrepresents LLM behavior, chasing latency misses the point, and shifting all liability to the clinician damages trust without addressing the underlying control gap.

Exam trap

The trap here is treating a compliance question about model error as a request for a guarantee or a liability disclaimer rather than a request for a control process.

40
MCQmedium

Refer to the exhibit. An audit of an Anthropic API configuration reveals the policy shown. What is the primary governance concern with this implementation?

A.The temperature setting is too high for professional administrative tasks.
B.The safety settings expose the organization to content moderation and liability risks.
C.The max_tokens limit is too low, restricting the capability of the system.
D.The system prompt is too generic, leading to poor model performance.
AnswerB

Disabling safety filters for harassment and hate speech ignores the responsibility to provide a safe AI environment. This exposure can lead to the generation of prohibited content, which the organization is legally responsible for, potentially resulting in severe regulatory scrutiny and damage to the corporate reputation.

Why this answer

The configuration explicitly disables core safety filters for harassment and hate speech. In an enterprise context, this creates significant legal and brand risk. Governance frameworks mandate that safety guardrails remain enabled to protect the organization from generating or facilitating harmful content.

By explicitly setting these to 'block_none', the organization bypasses essential protections that are designed to uphold safety standards and ethical AI usage requirements.

Exam trap

Candidates often focus on the 'performance' benefit of disabling filters (faster responses). They overlook that in an enterprise context, the legal and brand liability of unfiltered content is unacceptable.

41
MCQhard

Which THREE strategies are effective for managing bias in AI-driven decision-making systems?

A.Conduct regular testing using diverse datasets to identify performance disparities.
B.Implement a human-in-the-loop review for high-impact decision scenarios.
C.Audit the training data for representative balance and potential historical skew.
D.Disable all feedback loops to prevent users from influencing the model's bias.
E.Use a proprietary model that hides its reasoning to prevent users from detecting bias.
AnswerA, B, C

Regular testing against diverse benchmarks is essential for surfacing hidden biases. By measuring performance across different demographics or scenarios, the organization can identify where the model is failing to be equitable, allowing for timely adjustments and ensuring that decisions remain fair and consistent across all user groups.

Why this answer

Managing AI bias requires a combination of data-level intervention, model validation, and ongoing monitoring. Bias often originates in training data, so evaluating and cleansing datasets is a crucial first step. Continuous monitoring and testing against diverse benchmarks ensure that the model remains fair over time.

These strategies are essential for enterprise governance to ensure that automated decisions are equitable, legally compliant, and aligned with organizational values.

Exam trap

Test-takers often select only algorithmic adjustments while ignoring crucial data-level interventions and human oversight needed for a comprehensive bias management strategy.

42
MCQmedium

An organization requires strict adherence to data residency requirements for PII processed by Claude. Which strategy best ensures that customer prompts and completions remain within a specific geographic boundary while utilizing Anthropic's API?

A.Enable global load balancing across all available Anthropic regions to optimize latency.
B.Encrypt all outgoing prompts using a third-party gateway before sending to Anthropic.
C.Configure the API client to target region-specific endpoints and verify regional data residency settings.
D.Anonymize all data locally and rely on Anthropic's general model training to handle PII.
AnswerC

Targeting region-specific endpoints ensures that the API request is handled by infrastructure physically located within the required jurisdiction. When coupled with data residency settings, this architecture guarantees that neither prompts nor completions are stored in non-compliant regions, effectively addressing both processing and storage governance concerns.

Why this answer

Implementing region-specific API endpoints combined with Data Residency configurations ensures that data processing and storage occur within authorized borders. This approach is critical for regulatory compliance in jurisdictions like the EU. By limiting the scope of model interaction to regional infrastructure, architects prevent the inadvertent cross-border transfer of sensitive data, thereby aligning with global privacy governance frameworks and minimizing legal exposure for the enterprise.

Exam trap

Candidates often assume that simply 'enabling encryption' satisfies data residency. Data residency specifically requires ensuring the data physically stays within defined geographic boundaries, which requires endpoint configuration.

43
MCQmedium

A development team wants to optimize the latency of their prompt engineering workflow using Claude. They currently run evaluations sequentially. Which approach best improves iteration speed?

A.Switch to a smaller model version for all initial prompt testing.
B.Implement a distributed asynchronous evaluation framework for concurrent prompt execution.
C.Reduce the number of test cases to ensure the evaluation suite completes quickly.
D.Cache all user prompts to avoid re-sending identical requests to the API.
AnswerB

Distributing prompts across parallel worker nodes allows for simultaneous evaluation of various prompt structures. This approach maximizes throughput by utilizing the API's concurrency limits, enabling developers to obtain a comprehensive statistical analysis of prompt performance in a fraction of the time required by sequential execution methods.

Why this answer

Parallelizing evaluation pipelines allows developers to test multiple prompt variations concurrently, significantly reducing feedback loops. In the context of LLM development, bottlenecking often occurs during the testing phase where prompt sensitivity to small changes requires broad regression coverage. By integrating asynchronous evaluation frameworks into CI/CD, teams can validate changes rapidly without manual overhead, ensuring that prompt performance remains consistent across diverse input datasets.

Exam trap

Candidates often suggest manual testing or faster hardware, failing to realize that parallel execution is the only architectural way to scale prompt evaluation throughput.

44
MCQmedium

An enterprise client is integrating Claude for automated financial reporting. The stakeholders are concerned about data privacy and the potential for model hallucinations leading to inaccurate balance sheets. As the lead architect, how should you best manage these stakeholder expectations regarding output reliability?

A.Guarantee that the Claude model will achieve 100% accuracy through prompt engineering and iterative fine-tuning.
B.Restrict stakeholder access to the model outputs until the system has completed a six-month stabilization period.
C.Establish a formal verification workflow that routes model-generated reports through a human-in-the-loop audit process.
D.Shift the responsibility for output accuracy entirely to the legal department by requiring a disclaimer on all generated documents.
AnswerC

Introducing a human verification layer directly addresses the risk of hallucination while maintaining stakeholder trust. By formalizing this process, you acknowledge the probabilistic nature of the technology while implementing a concrete control mechanism. This provides stakeholders with a clear risk mitigation strategy that satisfies audit and compliance requirements.

Why this answer

Effective stakeholder management requires transparent communication regarding the limitations of LLMs. By implementing a human-in-the-loop review process and establishing strict system prompts, you mitigate risk while defining clear success metrics. This approach transforms abstract concerns into a structured governance model, ensuring that stakeholders understand the distinction between deterministic software and probabilistic generative AI, thereby aligning project delivery with business expectations for accuracy and compliance.

Exam trap

Candidates often suggest technical fixes like fine-tuning or prompt engineering as a complete solution, ignoring that high-stakes financial environments require human oversight to guarantee accuracy and legal accountability for output.

45
Multi-Selecthard

A fintech company's risk committee is operationalizing a governance program for Claude-powered customer support agents. They must demonstrate to regulators that model behavior changes are tracked, attributable, and reversible. Which TWO practices best satisfy this requirement? (Choose two.)

Select 2 answers
A.Route every request through a gateway that logs the full request and response payloads with the model identifier attached.
B.Enable prompt caching on all production requests so that previously validated prompts are reused verbatim.
C.Run a fixed regression suite against each candidate model version and archive the scored results before promoting it.
D.Pin production deployments to dated model identifiers and maintain a change log linking each identifier to its evaluation results.
E.Store the system prompt in a version-controlled repository and require pull-request review before it is merged.
AnswersC, D

A fixed regression suite produces comparable, dated evidence that a candidate version behaves acceptably on the scenarios the business cares about. Archiving scores before promotion creates the attributable record regulators expect and gives the team an objective basis for approving or rejecting an upgrade. Combined with pinned identifiers, it closes the loop between decision and deployed artifact.

Why this answer

Attributable, reversible model change management requires two things working together: a frozen, dated artifact in production and comparable evidence captured before promotion. Pinning to dated model identifiers supplies the frozen artifact and the rollback target, while a fixed regression suite run against each candidate supplies the before-and-after evidence. Together they let the risk committee answer what changed, who approved it, and how to undo it.

Exam trap

The trap here is assuming that logging, caching, or prompt version control constitutes model change management, when none of them actually freezes or attributes changes to the model artifact itself.

46
MCQhard

A multinational corporation is using Claude to process employee feedback surveys. The data includes sensitive personal opinions. The governance team must ensure that the AI system complies with the EU's General Data Protection Regulation (GDPR). Which control is most critical to address the 'right to explanation' requirement for automated decision-making?

A.Obtain explicit consent from all employees before processing their feedback with AI.
B.Implement a mechanism to provide a human-readable explanation of how the AI arrived at its conclusions for each individual.
C.Anonymize all employee feedback before processing to remove personal identifiers.
D.Store all data within the EU to comply with data residency requirements.
AnswerB

GDPR's right to explanation requires that individuals can obtain meaningful information about the logic involved in automated decisions. Providing a human-readable explanation for each individual directly satisfies this requirement. This control ensures transparency and accountability, which are core to GDPR compliance for automated processing.

Why this answer

The right to explanation under GDPR requires that individuals receive meaningful information about the logic of automated decisions. Implementing a mechanism to generate human-readable explanations for each individual directly fulfills this. Other controls like anonymization, consent, or data residency address different aspects of GDPR but not the specific transparency requirement.

Exam trap

The trap here is confusing other GDPR principles like consent or data residency with the right to explanation, which specifically demands transparency of automated decision logic.

47
MCQeasy

You are the lead architect on a Claude-powered claims triage tool. Six weeks before launch, the executive sponsor asks for a single one-page view of project health that shows schedule variance, cost variance, and scope changes at a glance. Which artifact should you produce?

A.An executive status dashboard summarizing schedule variance, cost variance, and approved scope changes with RAG status.
B.A risk register listing every identified risk with probability and impact scores.
C.A requirements traceability matrix mapping each functional requirement to its test case.
D.A detailed Gantt chart exported from the project management tool showing every task and dependency.
AnswerA

An executive status dashboard is purpose-built for exactly this request: it condenses schedule variance, cost variance, and scope change history into a single page with red/amber/green indicators. It respects the sponsor's time, surfaces exceptions rather than raw task data, and creates a repeatable reporting cadence the steering committee can use through launch and beyond.

Why this answer

The sponsor needs a consolidated, glanceable view of the three classic project health dimensions: schedule, cost, and scope. An executive status dashboard presents variance against baseline with RAG indicators, which supports fast decisions about funding, staffing, or descoping. Task-level artifacts like Gantt charts, traceability matrices, and risk registers serve delivery teams, not executive steering, so they cannot substitute for this rollup.

Exam trap

The trap here is assuming that more detail equals better communication, when executives actually need aggregated variance indicators rather than task-level artifacts.

48
MCQmedium

A document-processing agent must extract structured fields from thousands of PDFs. Some PDFs are scanned images, some are native text, and some are encrypted. The architect wants one pipeline that routes each document to the appropriate extractor and reports per-document confidence. Which design best fits?

A.Pre-classify each PDF by inspecting its structure, route native-text documents to a text extractor, scanned images to OCR, and encrypted files to a decryption step, then validate extracted fields and attach confidence scores.
B.Reject any PDF that is not native text and return an error, since scanned and encrypted documents are out of scope for automated processing.
C.Convert every PDF to images first, then run OCR on all of them, and skip classification because OCR handles any document.
D.Send every PDF to a single vision-language model and ask it to return the structured fields directly, ignoring document type.
AnswerA

Inspecting structure first lets each document take the cheapest and most accurate path: text extraction for native PDFs, OCR for scans, and a decryption stage for encrypted files. Validation and confidence scoring after extraction give the per-document reporting the architect wants. This matches processing to document type instead of forcing one method on all three, improving both cost and accuracy.

Why this answer

The three document classes need different handling, so the pipeline should inspect each file and route it: direct text extraction for native PDFs, OCR for scans, and a decryption stage for encrypted files. After extraction, validating fields and attaching confidence scores gives the per-document reporting the architect requires. This approach uses the cheapest viable method for each class and keeps confidence meaningful, rather than forcing one uniform method that is wrong for at least one class.

Exam trap

The trap here is treating PDF processing as a single uniform step, when routing by document structure is what makes extraction both accurate and cost-effective across mixed inputs.

49
MCQmedium

A stakeholder expresses concern about data privacy regarding the inputs sent to Claude. What is the most appropriate way to address this?

A.Tell them privacy is not an issue because all data is automatically deleted.
B.Provide documentation regarding data handling, encryption, and the Anthropic data processing agreement.
C.Suggest they talk to the legal department and wait for a response.
D.Ask them to sign a waiver stating they are responsible for any data leaks.
AnswerB

This is the most responsible way to address privacy concerns. It provides the stakeholder with objective evidence and legal frameworks that govern the project. This builds professional trust and ensures that the response is legally and technically accurate, satisfying the organization's compliance requirements for data protection and security.

Why this answer

Addressing data privacy requires a formal, evidence-based approach that references established security frameworks and contractual terms. By explaining the data processing agreement and the security controls in place—such as encryption and data retention policies—you provide the stakeholder with the necessary assurance to move forward. This professional response treats privacy as a standard architectural requirement rather than an afterthought, reinforcing trust in the system's design and organizational compliance.

Exam trap

Candidates often suggest vague assurances or technical workarounds like local hosting, failing to realize that formal data processing agreements and official documentation are required.

50
MCQmedium

When planning for the long-term support of an Anthropic API deployment, what is the most important stakeholder alignment activity?

A.Ensure that the IT budget is fixed for the next five years to prevent cost fluctuations.
B.Establish a roadmap for periodic model updates and performance validation.
C.Sign a lifetime contract with the current development team to avoid knowledge loss.
D.Focus on manual documentation of every prompt to keep the system unchanged forever.
AnswerB

A roadmap for updates and validation ensures that the system stays current with the best available models while maintaining the necessary quality standards. This proactive approach helps stakeholders understand that the system is a living product that requires ongoing maintenance, which is key for long-term project success and sustainability.

Why this answer

Regular alignment on the evolving capabilities of the model and the corresponding updates to the application is critical. LLM technology evolves rapidly, and maintaining an alignment between new capabilities and the business roadmap ensures that the organization extracts maximum value over time. This keeps stakeholders engaged and ensures that the system doesn't become stagnant, providing a path for continuous improvement and innovation throughout the lifecycle of the AI integration.

Exam trap

Candidates often focus on 'cost optimization' as the main long-term goal. While important, model evolution requires continuous performance validation to ensure the application remains accurate and safe over time.

51
Multi-Selectmedium

You are designing an agentic workflow that must reliably complete a multi-step refund process across an internal billing service and an external payment gateway. The workflow can fail at any step, and partial completion is unacceptable. Which two architectural mechanisms are required to guarantee that the workflow either completes fully or leaves no partial effect? (Choose two.)

Select 2 answers
A.A compensating action for each forward step, invoked in reverse order when the workflow cannot proceed to completion.
B.A semantic cache that stores prior successful refund transcripts and replays them on similar requests.
C.A higher model temperature on the planner so it can explore alternative step orderings when a step fails.
D.A durable execution log that records each step's intent and outcome so the orchestrator can resume or compensate after a crash.
E.A longer per-step timeout so that slow external calls are given more time to finish before the orchestrator gives up.
AnswersA, D

Because the two services cannot share a single transaction, atomicity must be simulated with compensation. Each forward step needs a defined inverse, such as voiding a charge when a subsequent refund fails. Invoking them in reverse order unwinds partial effects, which is what makes the workflow effectively all-or-nothing despite spanning independent systems.

Why this answer

Atomicity across independent services is achieved with a saga-style pattern: a durable log that records each step's intent and outcome, plus compensating actions that undo committed steps when the workflow cannot finish. The log enables correct recovery after a crash, and the compensations unwind partial effects in reverse order. Together they make the workflow effectively all-or-nothing even though no single transaction spans both systems.

Exam trap

The trap here is reaching for latency or model-tuning knobs such as timeouts and temperature, which change timing and variability but provide no mechanism for undoing a partially committed refund.

52
Multi-Selectmedium

Your organization is preparing to retire a legacy Claude model snapshot that several internal teams still call through their own applications. You must communicate the deprecation without disrupting business operations. (Choose two.)

Select 2 answers
A.Extend the legacy snapshot indefinitely to avoid any possibility of disruption to internal teams.
B.Identify each consuming team, confirm their migration owner, and track migration status against the sunset date.
C.Silently redirect all traffic to the newest available model snapshot so consumers experience no interruption.
D.Publish a deprecation timeline with the exact sunset date, the recommended replacement snapshot, and migration guidance for each consuming team.
E.Require every consuming team to re-run its full evaluation suite and submit results before the sunset date.
AnswersB, D

Knowing who consumes the snapshot and who owns each migration converts a broadcast announcement into an accountable plan. Tracking status against the sunset date lets the platform team intervene early with teams that are stalled, rather than discovering non-compliance on the day of removal. It also produces an accurate impact assessment if the sunset date must move.

Why this answer

Effective deprecation communication combines a clear timeline with a named replacement and per-team migration guidance, plus identification of each consuming team and its migration owner with status tracking. Together these give consumers both the information and the accountability needed to migrate before the sunset date. Silent redirection, indefinite extension, and mandatory evaluation submission each substitute a shortcut for a managed transition.

Exam trap

The trap here is believing that preventing visible disruption, either through silent redirection or indefinite extension, is the same as managing a deprecation well.

53
MCQmedium

A platform team wants every service to call Claude through a single internal gateway that injects the system prompt, enforces token budgets, and emits OpenTelemetry traces. A developer proposes having each service call the Anthropic Messages API directly and centralizing only the API key in a shared vault. What is the strongest architectural reason to reject the developer's proposal?

A.Direct client calls cannot use streaming responses, so latency-sensitive features would break.
B.The Anthropic Messages API rejects requests that originate from more than one client ID per key.
C.Centralizing the API key without centralizing the call path removes the single point where prompt policy, token budgets, and traces can be enforced.
D.A shared vault secret cannot be rotated without redeploying every consuming service.
AnswerC

The gateway exists to be the chokepoint where system prompts are injected, budgets are checked, and spans are emitted. If services call the Messages API directly, those controls fragment across codebases and drift. A shared key protects the credential but does nothing for policy or observability. Centralizing the call path is what makes the controls reliable and auditable across every service.

Why this answer

A shared secret protects only the credential, not the behavior around it. The gateway is valuable because it is the one place where system prompts are injected, token budgets are enforced, and OpenTelemetry spans are produced. Routing every service through it keeps those controls consistent and auditable, whereas direct Messages API calls scatter policy across many codebases and make drift inevitable.

Exam trap

The trap here is assuming that centralizing the API key alone delivers the same governance as centralizing the call path.

54
MCQmedium

Your organization is scaling an internal library that wraps Anthropic API calls. To minimize the cognitive load on developers using this library, what is the most effective pattern to implement?

A.Require developers to write raw HTTP requests to the API in every microservice.
B.Bundle all API interaction logic into a shared, versioned SDK with built-in observability.
C.Mandate that all developers use a specific GUI tool for prompt testing rather than code.
D.Provide only documentation on how to authenticate, leaving all implementation to teams.
AnswerB

A centralized SDK provides a unified, well-tested interface for API consumption. By including observability, retries, and security defaults, it enables developers to integrate Claude quickly and reliably. This approach lowers the barrier to entry, ensures consistent performance, and simplifies maintenance as the organization's usage of the API grows.

Why this answer

Providing a high-level SDK with built-in retry logic, telemetry, and standardized error handling reduces the complexity for individual developers. By abstracting away the boilerplate code required for API connectivity, developers can focus on prompt engineering and business logic. This standardization ensures that all teams use secure, performant, and observable patterns, which significantly enhances organizational productivity and reduces technical debt across various internal projects.

Exam trap

Candidates often suggest building custom wrappers for every project or using raw API calls directly in the codebase, failing to realize that individual implementation leads to inconsistent observability and massive maintenance overhead.

55
MCQmedium

You are architecting a customer-support agent for a SaaS platform. The agent must answer billing questions using a live invoice API and general policy questions using a static knowledge base. The invoice API is fast but occasionally returns stale data; the knowledge base is large and slow to search. Your design gives the agent a dedicated 'billing' sub-agent and a 'policy' sub-agent, coordinated by a router agent. After deployment, you observe the router sending nearly every query to both sub-agents in parallel, causing high latency. Which architectural change best addresses this while preserving answer quality?

A.Merge both sub-agents into a single agent with access to both the invoice API tool and the knowledge base search tool, and let it choose tools freely.
B.Have the router emit a structured routing decision with a confidence score, and only fan out to both sub-agents when the confidence falls below a configured threshold.
C.Cache the router's routing decisions for 24 hours so repeated identical queries skip the router on subsequent requests.
D.Replace the router with a fixed decision tree that maps every query containing the word 'invoice' to the billing sub-agent and all other queries to the policy sub-agent.
AnswerB

This preserves the router's semantic judgment while adding a confidence gate. High-confidence queries go to a single sub-agent, cutting latency; genuinely ambiguous queries still fan out, preserving answer quality. The threshold is tunable per environment, giving you an explicit latency/accuracy dial rather than an all-or-nothing routing rule.

Why this answer

The router is being overly cautious and fanning out on every query, which multiplies latency. Adding an explicit confidence signal to the routing decision lets the orchestrator fan out only when the router is genuinely unsure. High-confidence routing keeps the common case single-hop, while the fallback preserves correctness on ambiguous queries.

This is the only option that changes the router's decision policy rather than replacing it with a brittle rule or an unrelated caching layer.

Exam trap

The trap here is treating 'router fans out too often' as a routing-logic bug to be replaced with deterministic rules, when the real fix is to expose the router's confidence so the orchestrator can gate fan-out.

56
Multi-Selecthard

An agentic system uses a supervisor agent that delegates to specialized worker agents. During a long incident, the supervisor's own context fills with worker transcripts, and it begins losing track of which workers have completed and which are still running. Which TWO architectural changes best preserve the supervisor's ability to coordinate correctly? (Choose two.)

Select 2 answers
A.Restart the supervisor with a fresh context whenever its window fills, and rely on workers to resend their current status on request.
B.Have the supervisor maintain an external task ledger that records each delegated task, its assigned worker, and its status, and read from that ledger instead of retaining full worker transcripts.
C.Allow workers to message each other directly to resolve dependencies, removing the supervisor from the coordination path entirely.
D.Increase the supervisor's context window to the largest available model so it can retain all worker transcripts for the duration of the incident.
E.Have each worker return a compact structured status envelope containing task id, state, and a short result summary, and have the supervisor track tasks by these envelopes.
AnswersB, E

An external ledger keeps authoritative state about what was delegated and what finished, so the supervisor can coordinate without carrying every worker transcript in context. It survives context growth and lets the supervisor query only the fields it needs. This directly addresses losing track of completed versus running workers while keeping the supervisor's context bounded.

Why this answer

The supervisor's difficulty comes from mixing coordination state with raw worker output. Externalizing task state into a ledger and having workers return compact structured envelopes keeps the authoritative record of what is delegated and what has finished separate from bulky transcripts. The supervisor can then coordinate from small, reliable signals and pull detail only when needed, rather than trying to hold the entire incident in its context window.

Exam trap

The trap here is believing that a bigger context window or a restart can substitute for externalizing task state, when both leave the supervisor's coordination record coupled to transcript volume.

57
MCQmedium

Refer to the exhibit. An audit team identifies that the system message lacks specific data handling instructions for sensitive reports. As the architect, what is the best approach to communicate this to the development team?

A.Directly update the production configuration files without consulting the developers to expedite the fix.
B.Schedule a review meeting to explain the requirement for a system message that enforces data classification.
C.Send an email suggesting that they use a different model that has built-in data filtering.
D.Ignore the finding since the model is generally helpful and unlikely to leak information.
AnswerB

A review meeting facilitates shared understanding of the security risk and the proposed technical solution. This collaborative approach allows developers to ask questions and ensures the guardrails are integrated correctly into the application code, which is vital for maintaining consistent policy enforcement across the entire AI service lifecycle.

Why this answer

Addressing the system prompt configuration is a critical technical governance task. By clearly articulating the need for a system-level guardrail, the architect ensures that the development team understands the security implications. This communication approach promotes a culture of security-first development, ensuring that all API interactions are constrained by necessary policy directives before they reach production environments, thereby preventing inadvertent data leakage through improper model interaction.

Exam trap

Test-takers might assume the architect should rewrite the code directly, missing the collaborative and communicative responsibility required to guide development teams on system security configurations.

58
MCQmedium

How should a development team manage sensitive system instructions that they do not want users to see or modify?

A.Obfuscate the prompt by converting it to base64 before sending it to the client.
B.Store the instructions on the server and use them to build the request payload.
C.Ask the model to never repeat its instructions to the user.
D.Hardcode the prompts in the client-side JavaScript to minimize server latency.
AnswerB

Keeping instructions on the server ensures they are never exposed to the client. The server constructs the full request, including the hidden system instructions, and sends it to the API. This is the only way to effectively prevent client-side manipulation and maintain the integrity of the model's operational constraints.

Why this answer

System instructions must be handled server-side to remain protected from client-side interference. By keeping the logic in the backend, developers ensure that users cannot inspect, modify, or inject instructions into the conversation. This pattern is fundamental to security, ensuring the integrity of the model's behavior and protecting the intellectual property of the prompt logic from malicious actors or unauthorized tampering by end-users.

Exam trap

Candidates store sensitive system instructions on the client side where end-users can easily inspect and tamper with them.

59
MCQmedium

When designing a system for high-volume document analysis, what is the best strategy to maximize cost efficiency and developer velocity?

A.Execute all requests synchronously to ensure the user gets an immediate result.
B.Adopt a queue-based architecture with Anthropic's Batch API for bulk tasks.
C.Split large documents into tiny chunks and process them in parallel using individual requests.
D.Deploy a dedicated cluster of GPUs to run an open-source model locally.
AnswerB

The Batch API provides an optimized way to process large volumes of data asynchronously, offering significant cost savings and better reliability than individual synchronous calls. This pattern allows for cleaner, more scalable code, letting developers focus on the document processing logic rather than managing connections and complex retries.

Why this answer

Using the Batch API for asynchronous processing allows the system to operate efficiently at a lower cost while simplifying the architecture. By offloading document processing from the request-response cycle, the system becomes more resilient to traffic spikes. This allows developers to design around throughput rather than latency, leading to cleaner code and fewer infrastructure challenges related to synchronous request timeouts or rate-limiting.

Exam trap

Candidates frequently choose synchronous processing for large volumes, causing massive bottlenecks and timeouts, ignoring the efficiency gains of asynchronous batch processing for non-real-time tasks.

60
MCQmedium

When designing an LLM application, what is the best strategy for managing 'system instructions' (system prompts) to prevent unauthorized alteration?

A.Hardcode the system prompt directly in the client-side JavaScript for speed.
B.Store the system prompt in a centralized, secured configuration service.
C.Allow users to modify the system prompt to customize their experience.
D.Use a public Git repository to store the system prompts for transparency.
AnswerB

Storing system prompts in a secure, backend configuration service keeps them out of the reach of the client, protecting them from tampering. This allows for centralized version control, auditing, and secure distribution, which are critical components of a resilient and well-governed AI application architecture in production.

Why this answer

System instructions should be stored in a secured, backend configuration service that is injected at runtime, rather than being hardcoded or exposed to the client-side. This architecture ensures that the system prompt is immutable from the user's perspective. By centralizing the storage and management of these instructions, architects provide a secure, auditable method for updating behavior without exposing the underlying logic to potential client-side manipulation or injection attacks.

Exam trap

Candidates frequently suggest embedding system prompts directly into client-side code or user-facing payloads, making them vulnerable to tampering and client extraction.

61
MCQmedium

An organization wants to allow non-technical business users to test Claude prompts without exposing them to raw API code. What is the most productive approach to empower these users?

A.Give every user their own API key and a Python IDE to write scripts.
B.Create a secure web-based UI that allows users to test prompts against specific model versions.
C.Require all prompt suggestions to be submitted via a ticket system for developers to code.
D.Provide access to the public Anthropic Console directly for all business users.
AnswerB

A web-based UI provides a safe, intuitive environment for users to experiment without needing code. It allows them to see model responses in real-time, facilitating faster iteration. Centralizing this via a UI also allows the organization to monitor usage, manage costs, and enforce security policies at the entry point.

Why this answer

Building an internal 'Playground' interface that interfaces with the API allows business users to refine prompts in a safe, controlled environment. By abstracting the technical details, the organization enables subject-matter experts to contribute to prompt engineering. This collaborative approach improves productivity by shortening the feedback loop between business needs and technical implementation, ensuring that the final prompts are both effective and aligned with organizational goals.

Exam trap

Candidates frequently suggest giving business users direct access to API keys or IDEs, forgetting that security and ease of use are paramount for non-technical personas.

62
Multi-Selectmedium

You are architecting a Claude agent that orchestrates a long-running approval workflow spanning hours or days, where human reviewers may intervene between steps. Which TWO mechanisms are necessary to keep the workflow correct across these interruptions? (Choose two.)

Select 2 answers
A.Set a very short max_tokens on every turn to reduce the chance of context drift across days.
B.Persist the full message history and tool state to durable storage after each turn so the workflow can resume from the exact checkpoint.
C.Disable tool use entirely during human review windows and rely on the reviewer to re-enter all prior context manually.
D.Implement idempotency keys on all side-effecting tool calls so retries after an interruption do not duplicate approvals or notifications.
E.Use a stateless design where each invocation rebuilds context by summarizing all prior decisions into the system prompt.
AnswersB, D

Durable checkpointing is essential because the process may be suspended or crash between human interventions. Saving message history and tool state lets the workflow resume mid-conversation without re-running side effects or losing prior reasoning. Without it, a restart would either duplicate actions or force the agent to begin again, breaking the approval chain.

Why this answer

Long-running workflows with human-in-the-loop interruptions require durable checkpointing to resume exact state and idempotency keys to make retries safe. Together they ensure that an interruption or crash neither loses the reasoning chain nor duplicates side effects such as approvals or notifications. Stateless summarization, token limits, and disabling tools all fail to provide either guarantee.

Exam trap

The trap here is believing that reconstructing context from summaries is equivalent to persisted state, when only durable checkpoints preserve pending tool calls and exact arguments.

63
Multi-Selecthard

Your team operates a Claude-based internal knowledge assistant. A new compliance officer asks how the platform will handle model deprecations and capability changes over the next 24 months. Which TWO practices should be established now to give stakeholders durable lifecycle assurance? (Choose two.)

Select 2 answers
A.Ask the model provider to guarantee in writing that no model used by the platform will ever be retired.
B.Pin production traffic to a specific model version and maintain a documented, tested upgrade path with a rollback option before any version change.
C.Publish a lifecycle register that maps each Claude model in use to its intended business function, known deprecation signals, owner, and scheduled review date.
D.Freeze all prompt and evaluation changes indefinitely so that the only variable in the system is the underlying model.
E.Always route production requests to the newest available Claude model so the platform is never running outdated versions.
AnswersB, C

Pinning gives deterministic behaviour that compliance can reason about, while a tested upgrade path with rollback converts a disruptive event into a controlled change. Together they provide the predictability the officer is asking about, because the platform can demonstrate that a deprecation will not silently alter outputs in a regulated workflow.

Why this answer

Durable lifecycle assurance rests on two complementary controls: version pinning with a rehearsed upgrade and rollback path, and a documented register that ties every model to its business function, owner, and review cadence. Pinning supplies predictability; the register supplies visibility and accountability. Promises of eternal model availability, constant upgrades to the newest version, and indefinite prompt freezes all fail because they either cannot be honoured or actively prevent the controlled change that deprecation management requires.

Exam trap

The trap here is mistaking maximum freshness or a vendor permanence promise for lifecycle assurance, when assurance actually comes from controlled versioning plus documented ownership.

64
MCQhard

An enterprise wants to deploy an AI-powered customer service agent. What is the most important governance consideration when integrating with internal customer databases?

A.Ensure the model has write access to the database to update customer profiles.
B.Use a service account with scoped, read-only permissions to the necessary database views.
C.Store the database connection string directly in the prompt for ease of access.
D.Configure the agent to query the entire production database to ensure comprehensive answers.
AnswerB

Scoped, read-only permissions minimize risk by ensuring the AI agent can only access exactly what it needs to perform its function. This prevents unauthorized data exposure and ensures that any potential model hallucination or injection cannot result in unintended modifications to the underlying customer databases.

Why this answer

The most important consideration is implementing a robust access control layer between the AI agent and the database. The agent should only have read-only access to a strictly defined, non-sensitive subset of data. This prevents the model from inadvertently surfacing sensitive info or executing unauthorized data modifications, ensuring that the integration adheres to the principle of least privilege and maintains corporate data integrity standards.

Exam trap

Test-takers frequently select broad administrative permissions or full database access for the AI agent, failing to apply the principle of least privilege required for database integrations.

65
MCQhard

A research agent uses Claude with an extended thinking budget to analyze a 200-page regulatory filing. Mid-analysis it must call a `fetch_footnote` tool whose result is essential to the conclusion. The architect wants the tool result to be incorporated without discarding the model's prior reasoning. Which approach best achieves this?

A.Store the footnote in an external vector store and instruct the agent to retrieve it again later if needed.
B.Append the tool result as a new user message and restart the analysis from the beginning of the filing.
C.Summarize the footnote with a separate cheap model and paste the summary into the system prompt before the next call.
D.Insert the tool result into the conversation as a tool_result block associated with the original tool_use, then continue the same turn so the model resumes from its existing reasoning.
AnswerD

Returning the result as a tool_result block tied to the preceding tool_use preserves the message sequence the model already reasoned over, letting it continue the same turn without re-deriving prior steps. This is the canonical way to feed external data into an in-progress reasoning chain. It keeps the thinking budget intact and respects the API's tool-use contract.

Why this answer

Tool results must be returned as tool_result blocks paired with the originating tool_use so the conversation remains valid and the model can continue from its existing reasoning rather than restart. This preserves the thinking budget and the message sequence. Restarting, summarizing through another model, or deferring to retrieval all either discard reasoning or break the API contract, so they fail this scenario.

Exam trap

The trap here is treating a tool result as ordinary text that can be injected anywhere, when the API requires it to be paired with its tool_use block to keep the reasoning chain valid.

66
MCQmedium

An organization is evaluating the safety of an LLM-based agent. What is the 'Red Teaming' process in this context?

A.The process of updating the model's weights.
B.An adversarial testing methodology to find vulnerabilities.
C.The standard unit testing for API latency.
D.The automated deployment of new models.
AnswerB

Red Teaming specifically focuses on simulating adversarial behavior to probe for weaknesses in the AI's safety architecture. By proactively attempting to 'break' the model, teams can discover security flaws, prompt injection vulnerabilities, and other safety concerns before they can be exploited by real-world malicious actors.

Why this answer

Red Teaming is a structured adversarial testing approach where a dedicated team attempts to find weaknesses, bypass safety guardrails, and elicit harmful responses from an AI system. It is a vital component of the development lifecycle because it identifies edge cases and vulnerabilities that automated tests might miss. Conducting regular red teaming ensures that the system is resilient against sophisticated attacks and remains safe for production deployment.

Exam trap

Candidates often confuse Red Teaming with standard unit testing or QA processes. They assume it is about checking if the model works, rather than actively trying to break it.

67
MCQhard

An architect needs to build an agent that handles complex, multi-step data migrations where each step depends on the output of the previous one. Which approach is most robust for ensuring the agent doesn't lose track of the long-term goal during execution?

A.Reactive Tool Calling (ReAct)
B.Zero-shot prompting with all tools
C.Few-shot prompting with migration examples
D.Parallel Sub-tasking
E.Plan-and-Execute Pattern
AnswerE

This pattern involves an initial planning phase to decompose the migration into discrete steps, followed by an execution loop that references and updates the plan. This separation of concerns ensures the agent maintains a stable roadmap, allowing it to verify each dependency before proceeding to the next migration phase.

Why this answer

For multi-dependency tasks, a Plan-and-Execute pattern is superior because it separates the high-level strategy from the low-level tool interactions. By maintaining an explicit plan that is updated after each step, the agent can track its progress against the original goal and adjust its future actions based on the results of completed tasks.

Exam trap

Candidates often suggest a simple ReAct or sequential chain. They fail to realize that complex, multi-dependency tasks require an explicit planning phase to prevent the agent from drifting off-target mid-execution.

68
MCQeasy

A project sponsor asks for a timeline estimation for a project utilizing Claude for sentiment analysis. What is the most appropriate architectural response?

A.Provide a fixed date based on a standard waterfall project model.
B.Give a broad range with a disclaimer that accuracy tuning will dictate the final go-live date.
C.State that the timeline is impossible to predict and refuse to provide any estimates.
D.Ask the sponsor to define the date they want and agree to it immediately.
AnswerB

This response manages expectations by linking the timeline to objective quality standards. It highlights the technical reality that the project's completion depends on achieving specific performance metrics, which is a responsible way to communicate risk and uncertainty to stakeholders while still providing them with a reasonable window for planning.

Why this answer

Providing a timeline for AI projects requires acknowledging the iterative nature of model performance and data integration. By focusing on milestones rather than rigid delivery dates, you maintain flexibility while providing transparency. This approach protects the project team from unreasonable pressure and allows for the necessary R&D time required to refine prompts and safety guardrails, ultimately leading to a more successful and sustainable outcome for the stakeholders.

Exam trap

Candidates often provide a specific date, ignoring the inherent uncertainty of AI model performance and the iterative nature of prompt engineering and accuracy tuning.

69
MCQmedium

You are the lead architect on a Claude-powered contract-analysis platform. Two weeks before go-live, the General Counsel asks for written assurance that the system will not retain client contract text for model training and that all data stays within the EU. Which artifact should you produce first to satisfy this governance obligation?

A.A prompt-engineering guideline telling analysts to strip client names from contract text before submitting it.
B.A data-flow and retention attestation that maps contract text through the Claude API, documents Anthropic's zero-data-retention configuration, and confirms the EU-only processing region.
C.An updated project charter that adds a legal-review milestone to the schedule before the go-live date.
D.A risk-register entry logging 'data residency uncertainty' with a medium severity rating and an owner.
AnswerB

This directly addresses both concerns: it traces where contract text travels and records the contractual zero-data-retention setting plus the EU region selection. A written data-flow attestation is the governance artifact that lets counsel verify the claim independently, rather than accepting a verbal promise, and it becomes the reference document for later audits.

Why this answer

The stakeholder needs verifiable evidence about two distinct properties: retention behaviour and processing location. Only a data-flow and retention attestation both traces the path of contract text through the Claude API and records the concrete configuration choices, such as zero-data-retention and EU-region processing, that make the assurance true. Schedule changes, redaction habits, and risk logs are adjacent activities that do not answer the question.

Exam trap

The trap here is treating a governance request for evidence as a project-management request, and responding with a schedule change or risk log instead of a document that actually verifies the data-handling claims.

70
MCQhard

A production agent uses a ReAct loop and frequently reaches its maximum step budget while still mid-task, then returns a partial answer that looks complete. Telemetry shows the agent often re-reads the same file and re-queries the same database row across consecutive steps. Which change most directly reduces wasted steps while preserving the agent's ability to finish?

A.Raise the maximum step budget from 15 to 60 so the agent always has enough room to finish any task.
B.Replace the ReAct loop with a single-shot prompt that asks the model to plan every step up front and then answer without further tool calls.
C.Add a step-level deduplication cache that returns the prior result when the agent issues an identical tool call with identical arguments, and inject a reminder of remaining budget into the loop.
D.Lower the maximum step budget to 8 so the agent is forced to be more decisive and avoid redundant work.
AnswerC

Caching identical tool calls with identical arguments removes the repeated file reads and database queries that consume steps, directly attacking the observed waste. Surfacing remaining budget lets the model prioritize finishing over redundant exploration. Together they cut wasted steps while leaving the agent free to complete the task, rather than artificially capping its work or raising the ceiling without improving efficiency.

Why this answer

The wasted steps come from repeating identical tool calls, so the most direct fix is to detect and short-circuit duplicates by caching results keyed on the call and its arguments. Pairing that with visibility into remaining budget helps the model spend its steps on novel actions and wrap up cleanly. Adjusting the budget up or down changes when the agent stops but not how much of its work is redundant, and collapsing the loop removes the observation-driven adaptation the agent depends on.

Exam trap

The trap here is treating the step budget as the problem, when the budget is only the boundary that exposes redundant tool calls as the real inefficiency.

71
Multi-Selectmedium

When evaluating the performance of a new 'Orchestrator-Worker' agentic architecture, which THREE metrics provide the most insight into the system's efficiency and reliability?

Select 3 answers
A.Task Success Rate (TSR)
B.Total number of tools defined in the schema
C.Average Steps per Task
D.The model's pre-training data cutoff date
E.Tool Call Accuracy Rate
AnswersA, C, E

TSR is the primary indicator of whether the agent is actually fulfilling its intended purpose. It measures the percentage of sessions where the agent reaches a correct and verified conclusion, providing a high-level view of the architecture's effectiveness across a diverse set of test cases or user queries.

Why this answer

Evaluating agents requires looking beyond simple accuracy to understand the operational characteristics of the loop. Task success rate measures the ultimate goal, while steps-per-task and tool accuracy pinpoint where the logic might be breaking down or becoming inefficient. These metrics collectively guide the architect in refining the agent's reasoning path.

Exam trap

Candidates often select business metrics like 'User Satisfaction' or 'Revenue Impact'. These are lagging indicators that do not provide the granular technical feedback necessary to debug agentic reasoning loops or tool usage.

72
Multi-Selecthard

Which THREE components are essential for a robust 'Stakeholder Communication Plan' during an LLM project rollout?

Select 3 answers
A.Defined channels for feedback and incident reporting.
B.A static, one-time document updated only at project end.
C.Recurring status updates tailored to the stakeholder's role.
D.An escalation matrix for performance or quality issues.
E.Strict secrecy policies to prevent stakeholders from knowing details.
AnswersA, C, D

Structured channels for feedback allow the project team to capture issues as they arise in real-time. This is essential for continuous improvement and maintaining a responsive posture. Without these defined paths, feedback becomes fragmented and often misses the team entirely, preventing effective resolution of critical issues during the project rollout.

Why this answer

A communication plan must be systematic, addressing reporting, escalation, and change management. By establishing clear channels for feedback, regular status updates, and defined procedures for handling model deviations, the architect ensures alignment. This is critical because LLM projects involve high ambiguity, and a well-structured plan prevents the team from feeling isolated or lost as the model behaves in unexpected ways during the production rollout phase.

Exam trap

Candidates often forget the escalation matrix, failing to provide a clear path for when the model inevitably behaves unexpectedly, which leads to confusion and loss of confidence.

73
MCQmedium

A developer support team wants to give engineers a fast way to reproduce and debug failed Claude requests without exposing API keys or requiring them to install the SDK locally. Which approach best balances speed and safety?

A.Provide a web-based request playground that runs server-side, logs sanitized request IDs, and lets engineers replay a failed request with the same parameters.
B.Share a read-only API key in the team wiki so engineers can paste requests into a local script.
C.Ask engineers to file tickets with the failing request body, and have a central team reproduce them manually.
D.Publish a Docker image containing the SDK and a preconfigured key so engineers can run requests locally in an isolated container.
AnswerA

A server-side playground keeps API keys off developer machines while letting engineers replay exact request parameters tied to a logged request ID. Sanitized logging avoids leaking secrets. This gives fast reproduction and debugging without local installation, matching both the speed and safety goals. Replay with identical parameters is what makes failures reproducible.

Why this answer

A server-side playground with sanitized logging and request replay gives engineers self-service debugging at speed while keeping keys on the server. Replay of exact parameters makes failures reproducible, and per-request IDs support tracing. Shared keys, ticket queues, and keyed Docker images either expose credentials or slow engineers down, so they miss one of the two goals.

Exam trap

The trap here is thinking that a read-only key is safe to share, when any shared credential still violates the no-exposure requirement and blocks per-user attribution.

74
MCQhard

A business unit has been running its own Claude-powered tool outside the central platform for four months. It works well, and the unit lead wants it blessed as-is rather than migrated. As the platform architect, what is the most appropriate first step?

A.Escalate the unit lead to the executive sponsor for violating architecture standards.
B.Assess the tool against platform standards for security, data handling, cost visibility, and supportability, then agree on a remediation or integration path with the unit lead.
C.Grant an indefinite exemption so the unit can keep operating the tool without further review.
D.Mandate immediate migration to the central platform to eliminate the unmanaged tool.
AnswerB

Assessing against standards turns an unmanaged asset into a governed one without discarding value. It identifies concrete gaps, such as unclear data retention or invisible spend, and produces a negotiated path: remediate in place, integrate key capabilities into the platform, or migrate with a funded timeline. The unit lead stays a partner rather than a resistor.

Why this answer

Shadow AI is best handled by converting it into governed capability rather than by punishing it. A standards-based assessment produces concrete findings and a negotiated path, preserving the working tool's value while closing security, data-handling, cost, and support gaps. That keeps the unit lead as a partner and strengthens the platform's credibility.

Exam trap

The trap here is treating an unmanaged but successful tool as a compliance failure to be stamped out, when the higher-value move is to evaluate it and fold its strengths into the governed platform.

75
Multi-Selecthard

An engineering organization is building a shared internal Claude gateway used by many product teams. They want to enable rapid experimentation while keeping spend predictable and preventing any single team from starving others. Which TWO controls should the gateway implement to meet these goals? (Choose two.)

Select 2 answers
A.A shared global API key distributed to every team so they can call the Anthropic API directly when the gateway is slow.
B.A single organization-wide rate limit with no per-team differentiation, relying on social norms and team goodwill to prevent overuse.
C.Automatic model downgrades for any team that exceeds its budget, applied silently without notifying the team or recording the change.
D.Per-team token budgets with usage metering and alerts, enforced at the gateway before requests are forwarded to the Anthropic API.
E.Per-team concurrency and rate limits at the gateway, with queueing so bursts are smoothed rather than rejected outright.
AnswersD, E

Per-team budgets with metering let the platform enforce spend limits and notify owners before overruns occur. Enforcing at the gateway means a runaway team cannot consume shared capacity or budget, which directly addresses predictable spend and fair access across product teams.

Why this answer

Per-team budgets with metering address predictable spend and accountability, while per-team concurrency and rate limits with queueing address fair access and burst smoothing. Together they let teams experiment freely within their allocation, prevent any single team from monopolizing shared capacity, and keep the gateway's behavior observable and enforceable.

Exam trap

The trap here is thinking a single global limit is sufficient for fairness, when without per-team attribution one heavy consumer can silently starve every other team.

Page 1 of 4

Page 2

All pages