NCA-GENL Software Development Practice Question
A developer is using the NVIDIA API Catalog to experiment with a hosted LLM. They want to send a prompt and receive a completion. Which endpoint should they use?
⚠ Common exam trap
The trap here is assuming any completions endpoint works, but the chat completions endpoint is the correct one for modern conversational LLMs.
Answer choices
Why each option matters
Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.
Correct answer & explanation
✓
/v1/chat/completions
The /v1/chat/completions endpoint is designed for conversational LLMs and accepts a structured messages array. It is the primary interface for interacting with hosted models on NVIDIA API Catalog. Other endpoints serve different purposes like listing models or generating embeddings, and the legacy completions endpoint is less suitable for chat-based models.
Answer analysis
Option-by-option breakdown
For each option: why learners choose it and why it is or isn't the right answer here.
- ✓
/v1/chat/completions
Why this is correct
The /v1/chat/completions endpoint is the standard OpenAI-compatible API for chat-based interactions. It accepts a messages array and returns model-generated responses. NVIDIA API Catalog endpoints follow this convention, making it the correct choice for sending prompts and receiving completions.
- ✗
/v1/models
Why it's wrong here
The /v1/models endpoint lists available models but does not generate completions. It is used for discovery, not inference. Sending a prompt here would not return a generated response, so it does not meet the developer's need.
- ✗
/v1/completions
Why it's wrong here
The /v1/completions endpoint is for legacy text completion models that take a prompt string and return a completion. While it can generate text, NVIDIA's API Catalog primarily promotes chat completions for conversational LLMs. The chat endpoint is more appropriate and widely supported for modern models.
- ✗
/v1/embeddings
Why it's wrong here
The /v1/embeddings endpoint converts text into vector representations. It is used for tasks like semantic search, not for generating text completions. This endpoint would return embeddings, not a natural language response, so it is incorrect for this scenario.
About these practice questions
This NCA-GENL question is part of Courseiva's 367-question bank — original exam-style content with full explanations and wrong-answer analysis, never real exam questions or exam dumps. Learn why practice questions differ from exam dumps →
JA
Written and reviewed by Johnson Ajibi, MSc IT Security
Senior Network & Security Engineer · founder of Courseiva
Last reviewed September 2026 · checked against the official NVIDIA exam blueprint
This NCA-GENL practice question is part of Courseiva's free NVIDIA certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the NCA-GENL exam.