CCAO-F Claude Model Fundamentals Practice Question
An application requires Claude to analyze a 180,000-token legal document and find a specific clause. Why might Claude 3 Opus be a better choice for this task than a smaller model like Haiku, even though both have a 200,000-token window?
⚠ Common exam trap
Candidates assume that sharing the same maximum token window means smaller models perform retrieval tasks identically to flagship models, ignoring retrieval accuracy disparities.
Answer choices
Why each option matters
Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.
Correct answer & explanation
✓
Opus has higher 'Needle In A Haystack' recall at large context sizes.
Context window size is only one factor in model performance. For very large inputs, the model's ability to maintain focus and accurately retrieve information from the middle of the text (recall) is critical. Higher-tier models like Opus are specifically optimized for better performance at these extreme context limits.
Answer analysis
Option-by-option breakdown
For each option: why learners choose it and why it is or isn't the right answer here.
- ✗
Haiku's context window is only 200,000 tokens for output, not input.
Why it's wrong here
This is a misunderstanding of the context window. All Claude 3 models share the same 200k limit for the total sum of input and output. There is no distinction where Haiku has a smaller input capacity than Opus; the difference lies in processing quality, not raw capacity.
- ✓
Opus has higher 'Needle In A Haystack' recall at large context sizes.
Why this is correct
As the context window fills up, smaller models can sometimes lose 'focus' on information buried in the middle of the text. Opus is engineered to maintain near-perfect recall across the full 200,000 tokens, making it more reliable for finding specific details in very large legal documents.
- ✗
Opus can process 200,000 tokens in less than one second.
Why it's wrong here
Opus is the slowest model in the Claude 3 family due to its size. Processing 180,000 tokens involves significant computation and will take much longer than one second. Speed is actually a disadvantage for Opus compared to Haiku, though it is traded for higher accuracy.
- ✗
Smaller models automatically truncate inputs over 100,000 tokens.
Why it's wrong here
Claude models do not silently truncate inputs that are within their stated 200,000-token limit. If a developer sends 180,000 tokens, Haiku will attempt to process all of them. The difference is that Haiku may be less accurate in its analysis of that data compared to Opus.
About these practice questions
This CCAO-F question is part of Courseiva's 259-question bank — original exam-style content with full explanations and wrong-answer analysis, never real exam questions or exam dumps. Learn why practice questions differ from exam dumps →
JA
Written and reviewed by Johnson Ajibi, MSc IT Security
Senior Network & Security Engineer · founder of Courseiva
Last reviewed September 2026 · checked against the official Anthropic exam blueprint
This CCAO-F practice question is part of Courseiva's free Anthropic certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the CCAO-F exam.