Courseiva
Context and Reliability →mediumMultiple Choice

CCAR-F Context and Reliability Practice Question

An application uses Claude 3.5 Sonnet to summarize technical logs. Users report that the model occasionally ignores specific error codes when logs exceed 50,000 tokens. Which strategy best ensures consistent reliability for long-context tasks?

⚠ Common exam trap

Candidates often assume passing the entire raw log file is sufficient, ignoring the 'lost in the middle' phenomenon where models lose focus on critical data in very long prompts.

Answer choices

Why each option matters

Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.

Correct answer & explanation

✓

Pre-process logs into smaller, overlapping chunks and use a multi-step aggregation approach for summarization.

Long-context retrieval requires managing token density and model attention span. By implementing a sliding window or chunking strategy with overlapping segments, you ensure the model maintains context across boundaries. This prevents the 'lost in the middle' phenomenon where models prioritize information at the beginning or end of a prompt. Reliable context management is critical for technical applications where missing a single error code can lead to incorrect diagnostic conclusions.

Answer analysis

Option-by-option breakdown

For each option: why learners choose it and why it is or isn't the right answer here.

  • ✗

    Increase the temperature setting to 1.2 to encourage more creative scanning of the input text.

    Why it's wrong here

    Increasing temperature beyond 1.0 introduces excessive randomness, which undermines reliability in technical tasks requiring precision. Claude typically performs best with lower temperature settings for extraction and summarization tasks. High temperature increases the likelihood of hallucinations rather than improving the model's ability to attend to specific tokens within a large context.

  • ✗

    Use a system prompt to explicitly instruct the model to pay attention to all tokens regardless of their position.

    Why it's wrong here

    While system prompts provide important instructions, they do not overcome the architectural limitations of attention mechanisms in extremely long sequences. Relying solely on prompts to force attention ignores the inherent signal-to-noise ratio challenges present in massive inputs. Technical reliability must be supported by architectural preprocessing rather than just qualitative instructions.

  • ✓

    Pre-process logs into smaller, overlapping chunks and use a multi-step aggregation approach for summarization.

    Why this is correct

    Chunking with overlap ensures that no information is lost at the boundaries of the input window. This methodology allows the model to process manageable segments while maintaining continuity through the overlap. By aggregating these summaries, you ensure a higher degree of recall and reliability for critical data points like error codes.

  • ✗

    Switch to a smaller model version to ensure faster processing of the large log files.

    Why it's wrong here

    Switching to a smaller model generally decreases the context window capacity and reasoning capability, which exacerbates issues with long-context recall. Smaller models are less effective at maintaining complex dependencies across large datasets compared to the Sonnet or Opus variants. Reliability depends on selecting the model tier that matches the complexity.

About these practice questions

One of 271 original CCAR-F practice questions on Courseiva, each with a full explanation and wrong-answer analysis — not exam dumps or protected exam content. Learn why practice questions differ from exam dumps →

How Courseiva writes practice questions · Editorial policy

JA

Written and reviewed by Johnson Ajibi, MSc IT Security

Senior Network & Security Engineer · founder of Courseiva

Last reviewed September 2026 · checked against the official Anthropic exam blueprint

This CCAR-F practice question is part of Courseiva's free Anthropic certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the CCAR-F exam.