Be able to read a Messages API request and predict its behavior, then choose the right parameter for a goal. The single most important thing: know that conversation state lives in the messages array you send, and that temperature plus top_p jointly control output consistency.
Start practicing
Claude Model Fundamentals — choose a session length
Free · No account required
Domain overview
This domain covers the core mechanics of the Claude Messages API: how the messages array and roles work, why XML tags structure prompts, how temperature and top_p shape output determinism, and how streaming changes response delivery. Questions are scenario-based, asking you to predict behavior or pick the correct parameter or feature for a stated goal.
Exam objectives
How appending a message to the messages array continues a multi-turn conversation with Claude
Using XML tags to delimit instructions, context, and examples so Claude parses prompts reliably
Adjusting temperature and top_p to make summarization output predictable and consistent
Enabling streaming in the Messages API so tokens render incrementally instead of in one block
Assuming each request is stateless and forgetting that prior turns must be resent in the messages array to preserve context.
Lowering only temperature while leaving top_p high, or raising both, instead of tightening both for deterministic output.
Confusing streaming with a larger max_tokens or a different model, when it is a separate API parameter controlling delivery.
Click any question to see the full explanation and answer options, or start a focused practice session above.
An engineering team is designing a RAG system using Claude 3.5 Sonnet. They need to ensure the model focuses exclusively on the provided context without hallucinating external knowledge. Which architectural approach best ensures high adherence to provided context?
2Which capability is a primary benefit of using Claude 3.5 Sonnet compared to smaller, legacy models when processing complex, multi-step instructions?
3Refer to the exhibit. An application sends the provided JSON payload to the Anthropic API. What is the expected behavior regarding the system prompt?
4Which of the following describes the purpose of 'Role' assignment in the Messages API?
5A user is experiencing 'Model Refusal' when processing a document that contains sensitive (but safe) medical information. What is the most likely cause?
6When fine-tuning or optimizing prompts for Claude, what is the impact of excessive 'System Prompt' length?
7A developer wants Claude to act as a specialized technical support assistant for a specific product. Which component of the API call is most appropriate for defining the assistant's persona and product boundaries?
8Refer to the exhibit. What is the specific purpose of the 'stop_sequences' parameter in this JSON payload?
9When dealing with extremely large documents, what is the best strategy to maximize Claude's accuracy in information extraction?
10Why should developers use the Anthropic Messages API instead of legacy Completions API?
11Which of the following is an effective technique for reducing hallucination in Claude when answering fact-based questions?
12Refer to the exhibit. What will happen if the user adds a new message to the 'messages' array?
13A developer needs to ensure that Claude 3.5 Sonnet consistently follows a specific JSON schema for structured data extraction. Which implementation strategy provides the highest level of deterministic output format control?
14Which TWO of the following statements accurately describe the characteristics of the Claude 3.5 Sonnet model's context window and performance?
15When designing a prompt for Claude, which technique is most effective at reducing the risk of 'hallucination' or factually incorrect information?
16Which THREE factors are primary considerations when calculating the cost of using the Anthropic API in a production environment?
17What is the primary purpose of using XML tags within a prompt when working with Claude models?
18What does the 'Temperature' parameter control when configuring a request to a Claude model?
19Which of the following describes the correct behavior of the Anthropic API regarding the 'system' prompt?
20When evaluating model output, what does the term 'latency' specifically refer to in the context of the Anthropic API?
21In the context of the Anthropic API, what is the primary benefit of using Streaming?
22An insurance firm wants to use Claude to extract data from scanned PDF claim forms that contain both handwritten text and printed tables. Which fundamental capability of the Claude 3 model family makes this workflow possible without using external OCR software?
23Refer to the exhibit. A developer is testing the vision capabilities of Claude 3.5 Sonnet. Based on the provided JSON request, which statement accurately describes how the model will process this input?
24Which THREE components are required when making a successful request to the Anthropic Messages API? (Select THREE)
25Refer to the exhibit. In the context of Claude Model Fundamentals, what is the primary purpose of the 'system' field shown in this API configuration?
26An application requires Claude to analyze a 180,000-token legal document and find a specific clause. Why might Claude 3 Opus be a better choice for this task than a smaller model like Haiku, even though both have a 200,000-token window?
27Which term describes the fundamental unit of text that Claude processes, which can be a single character, a part of a word, or a whole word?
28Refer to the exhibit. A company is building a translation service for live subtitles where latency must be under 200ms per segment. Which Claude 3 model family member is the only viable candidate based on the provided performance and cost data?
29A healthcare provider wants to use Claude to summarize patient-doctor conversations. They are concerned about the model 'hallucinating' or making up medical facts. Which Claude 3 feature or design principle directly addresses this concern by ensuring the model is honest and admits when it doesn't know an answer?
30An analyst is using Claude 3.5 Sonnet to compare two 50-page contracts. They notice that Claude correctly identifies a discrepancy in a small footnote on page 74 of the combined input. This demonstrates which fundamental performance characteristic of Claude?
31An enterprise legal team needs to analyze a 150,000-token collection of merger and acquisition documents to identify conflicting indemnity clauses. Which feature of the Claude 3 model family is most critical for ensuring the entire dataset is processed in a single inference pass without losing context?
32A developer is integrating Claude 3.5 Sonnet into a medical imaging application. Which TWO capabilities of the Claude 3 family make it particularly suited for analyzing diagnostic reports alongside X-ray images?
33Which core design principle is used by Anthropic to ensure that Claude models are helpful, honest, and harmless by training them against a set of written rules or values?
34An AI researcher is concerned about 'hallucinations' when Claude summarizes internal technical specifications. Which approach leverages Claude's fundamental design to minimize the risk of the model inventing non-existent features?
35A company is developing an application that generates long-form creative content. They notice that as the output length increases, the model's speed seems to fluctuate. Which technical factor most significantly impacts the latency of the 'Time to Last Token' (TTLT) in this scenario?
36A data analyst wants to ensure that Claude produces highly predictable and consistent results when summarizing financial reports. Which TWO parameter adjustments should they make to the API request?
37An engineering firm is using Claude to help design a complex micro-architecture for a new processor. The task requires deep logical reasoning, knowledge of hardware description languages (Verilog), and the ability to handle highly abstract concepts. Which model should be used for the highest possible accuracy?
38A UX designer wants the AI assistant to appear more interactive by showing the response as it is being generated, rather than waiting for the entire block of text to be finished. Which API feature should the developer implement?
39A developer is building a RAG application and notices Claude 3.5 Sonnet occasionally hallucinates when provided with a large context window. Which architectural adjustment is most effective for improving factual fidelity?
40What is the primary function of the 'stop_sequences' parameter in the Claude API?
41When designing a prompt for Claude to perform complex reasoning, why is it beneficial to include 'think step-by-step' in the instructions?
42A product team is building a customer-support assistant on Claude. They want Claude to answer only from a fixed set of help-center articles and to refuse any question outside that scope. They also need to update the article set frequently without retraining a model. Which approach best meets these requirements?
43A financial analyst uses Claude 3.5 Sonnet to extract line items from quarterly PDFs. They want the model to output only a JSON array without any conversational filler. Which feature should they configure to enforce this output format most reliably?
44A support team wants Claude to answer customer questions using only the company's internal help articles. They plan to paste the relevant article text into the prompt before the user's question. Which prompting technique does this describe?
45A developer is building a multi-turn chat application with the Anthropic Messages API. After several exchanges, they notice Claude loses track of details mentioned early in the conversation. Which action best addresses this while staying within the API's design?
46A product team is building a customer support assistant on Amazon Bedrock using the Anthropic Claude 3.5 Sonnet model. They need the model to answer questions strictly from a provided knowledge base and to refuse to answer if the information is not present. Which technique should they use to constrain Claude's behavior most reliably?
47A product team is evaluating Claude for a customer-facing assistant that must refuse to give medical diagnoses. Which TWO techniques should they use to make the refusal behavior consistent across many different user phrasings? (Choose two.)
48A developer is choosing a Claude model for a high-volume classification task that must meet a strict monthly budget. The task is straightforward and does not require deep reasoning. Which selection principle best fits this scenario?
49A developer is using the Anthropic Messages API with Claude 3 Opus to build a multi-turn technical support chatbot. The conversation history is growing large, and they want to reduce token usage while preserving the most relevant context. They decide to implement a sliding window that keeps only the last N turns. What is a potential drawback of this approach that they should consider?
50An engineering team wants Claude to classify thousands of support tickets into categories. They need the model to always return one of five exact category labels. Which approach most reliably constrains the output to those labels?
51A team is building a document Q&A feature on Claude and wants to reduce hallucinations when the answer is not present in the supplied documents. Which TWO techniques are appropriate? (Choose two.)
52A data engineer is using Amazon Bedrock to invoke Claude 3 Haiku for classifying customer feedback into categories. They notice that the model sometimes returns categories that are not in the predefined list. Which change to the prompt is most likely to improve adherence to the allowed categories?
53A developer wants Claude to return structured data that another service can parse automatically. They need the output to follow a fixed schema every time, with no conversational text around it. Which approach is most appropriate?
54A developer is building an application that uses the Anthropic Messages API with Claude 3.5 Sonnet to generate structured JSON output for a data pipeline. They need to ensure the output is valid JSON and conforms to a specific schema. Which TWO strategies should they use to maximize reliability? (Choose two.)
55A developer is choosing between Claude 3 Haiku and Claude 3.5 Sonnet for a real-time chat application that requires very low latency and handles simple, short queries. Cost is a primary concern. Which model is most appropriate and why?
56A product manager is evaluating Claude models for a customer-support chatbot that must handle 50,000 conversations per day while keeping inference costs low. The conversations are short, factual, and do not require complex reasoning. Which Claude model is the most appropriate choice?
57A customer support team is building a Claude-powered assistant that must answer questions using a 300-page product manual. They want to avoid sending the entire manual with every request because of latency and cost. Which approach best leverages Claude's capabilities while keeping responses grounded in the manual?
58A developer is building a customer support chatbot using Claude. The chatbot must remember details from earlier in the conversation, such as the customer's order number and issue, to provide coherent responses. The conversation can last for many turns. Which implementation strategy best ensures Claude maintains context without exceeding token limits?
59A developer is integrating Claude into a legal document review system. The system must process lengthy contracts and answer questions about specific clauses. Which TWO of the following techniques help mitigate the risk of Claude hallucinating details not present in the document? (Choose two.)
60A product team is using Claude to generate marketing copy. They want to ensure the output aligns with brand voice and avoids certain topics. Which TWO techniques are most effective for controlling Claude's output in this scenario? (Choose two.)
61A developer wants Claude to always respond in a strict, terse style and never use emojis, regardless of how users phrase their requests. Where should this persistent behavioral instruction be placed in a request to the Anthropic Messages API?
62A machine learning engineer is comparing Claude 3 Opus and Claude 3.5 Sonnet for a complex mathematical reasoning task. The task involves multi-step proofs and requires the highest possible accuracy. Cost is not a primary concern. Which statement accurately describes the trade-off between these models for this use case?
63A team is choosing a Claude model for a nightly batch job that summarizes thousands of internal reports. Cost per token and throughput matter more than peak reasoning ability, and the summaries tolerate minor stylistic variation. Which selection criterion is most appropriate?
64An engineer is choosing between Claude 3.5 Sonnet and Claude 3 Opus for a pipeline that extracts structured fields from scanned invoices. The pipeline processes thousands of documents per hour and must balance accuracy against cost. Which statement best reflects the appropriate model-selection reasoning?
65A developer is using the Anthropic Messages API to build a conversational agent. They want to maintain context across multiple turns and ensure that the model's responses are consistent with the system prompt. Which API feature should they use to provide the system prompt?
66A team is using Claude to summarize legal contracts. They need the summaries to reflect only the contract text, not any outside assumptions. Which prompting technique best reduces the chance that Claude introduces information not present in the source document?
67A developer is designing a Claude-powered assistant that must maintain a coherent conversation across many turns while controlling cost and context limits. Which TWO practices are appropriate according to Claude model fundamentals? (Choose two.)
Be able to read a Messages API request and predict its behavior, then choose the right parameter for a goal. The single most important thing: know that conversation state lives in the messages array you send, and that temperature plus top_p jointly control output consistency.
The Courseiva CCAO-F question bank contains 67 questions in the Claude Model Fundamentals domain. Click any question to see the full explanation and answer breakdown.
Start with a 10-question focused session to identify your baseline accuracy in this domain. Read every explanation — even for questions you answer correctly — to understand the reasoning. Once you score consistently above 80%, move to a 20–30 question session to confirm depth before moving to the next domain.
Yes — the session launcher on this page draws questions exclusively from the Claude Model Fundamentals domain. Choose 10, 20, 30, or 50 questions for a focused session, or click individual questions to review them one by one.
Save your results, see per-domain analytics, and get readiness scores — free, for every certification.
Sign Up FreeFree forever · Every certification included