RAG for Proprietary Data Privacy
A startup is building a customer support chatbot using Vertex AI and wants to ground responses in their product documentation to reduce hallucinations. Which approach should they use?
⚠ Common exam trap
Google Cloud often tests the misconception that fine-tuning is the best way to incorporate domain knowledge, but the trap here is that fine-tuning does not provide dynamic, verifiable grounding with citations, whereas Vertex AI Grounding with a custom data store does, making it the correct choice for reducing hallucinations in a retrieval-augmented generation use case.
Answer choices
Why each option matters
Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.
Correct answer & explanation
✓
Enable Vertex AI Grounding with a custom enterprise data store containing the documentation.
Vertex AI Grounding with a custom enterprise data store is the correct approach because it allows the chatbot to retrieve and cite specific chunks from the product documentation in real time, directly reducing hallucinations by constraining responses to verified content. This method uses the underlying grounding service to query a vector-based data store (powered by Vertex AI Search) and append source references to the model's output, ensuring factual accuracy without retraining.
Answer analysis
Option-by-option breakdown
For each option: why learners choose it and why it is or isn't the right answer here.
- ✓
Enable Vertex AI Grounding with a custom enterprise data store containing the documentation.
Why this is correct
Grounding with a custom enterprise data store retrieves passages from the product documentation and injects them into the prompt, so responses cite actual content rather than relying on parametric memory. This directly addresses the hallucination constraint in the stem.
- ✗
Use the Codey API for text generation.
Why it's wrong here
Codey generates and completes code; it cannot retrieve product documentation, so responses remain ungrounded. It is tempting because Codey is a Vertex AI text-generation model, and would be the right choice for code assistance tasks such as completion, generation or code chat, not document-grounded support answers.
- ✗
Use the base model without any grounding to maximize flexibility.
Why it's wrong here
Using the base model alone leaves responses dependent on parametric knowledge, so hallucinations persist and no documentation is consulted. It is tempting because base models offer broad flexibility across topics, and would be correct for open-ended creative or general tasks where grounding in a specific corpus is not required.
- ✗
Fine-tune the model on the documentation and deploy.
Why it's wrong here
Fine-tuning bakes documentation into weights, which cannot cite sources or update without retraining, so it does not ground responses at inference time. It is tempting because fine-tuning does adapt a model to domain content, and would be correct for teaching consistent style or task format rather than retrieving facts.
Go deeper
Related to this question
About these practice questions
One of 1,008 original Generative AI Leader practice questions on Courseiva, each with a full explanation and wrong-answer analysis — not exam dumps or protected exam content. Learn why practice questions differ from exam dumps →
Same concept, more angles
2 more ways this is tested on Generative AI Leader
These questions test the same concept from different angles. Work through them to make sure you can recognise it however the exam phrases it.
Variation 1. A startup is building a customer service chatbot that generates responses in real-time. They want the model to have up-to-date information on the latest product catalog but cannot afford frequent fine-tuning. Which technique should they use to inject current data into the model without retraining?
easy- A.Rely on the model's zero-shot capabilities to infer product details.
- ✓ B.Use retrieval-augmented generation (RAG) to fetch relevant documents from a vector database at inference time.
- C.Craft detailed system prompts that include the entire product catalog in the prompt.
- D.Fine-tune the base model weekly on the latest product catalog.
Why B: Retrieval-Augmented Generation (RAG) is the correct technique because it allows the chatbot to fetch the most current product catalog entries from an external vector database at inference time, without requiring any model retraining. This keeps responses grounded in up-to-date information while avoiding the cost and latency of frequent fine-tuning.
Variation 2. A healthcare company is building a clinical decision support system using Gemini 1.5 Pro on Vertex AI. They need responses that are highly accurate and comply with medical regulations, including traceability to source documents. They have a large corpus of curated medical guidelines stored in PDFs in Cloud Storage. Their team has experience with both fine-tuning and prompt engineering. Which approach best ensures regulatory compliance and accuracy?
medium- ✓ A.Use a combination of grounding to the medical guidelines and prompt engineering with system instructions specifying compliance requirements.
- B.Use prompt engineering with system instructions and few-shot examples, but no grounding.
- C.Use grounding to the medical guidelines but rely on prompt engineering only for compliance instructions.
- D.Fine-tune the model on the medical guidelines corpus to internalize the knowledge.
Why A: Grounding the model to the curated medical guidelines in Cloud Storage ensures responses are directly traceable to source documents, which is critical for medical regulatory compliance. Combining this with system instructions that specify compliance requirements (e.g., HIPAA, FDA guidelines) enforces behavioral constraints without altering the model's weights, maintaining accuracy and auditability.
JA
Written by Johnson Ajibi, MSc IT Security
Senior Network & Security Engineer · founder of Courseiva
This Generative AI Leader practice question is part of Courseiva's free Google Cloud certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the Generative AI Leader exam.