Courseiva
Software Development →easyMultiple Choice

NCA-GENL Software Development Practice Question

A developer is using the NVIDIA API Catalog to experiment with a hosted LLM. They want to send a prompt and receive a completion. Which endpoint should they use?

⚠ Common exam trap

The trap here is assuming any completions endpoint works, but the chat completions endpoint is the correct one for modern conversational LLMs.

Answer choices

Why each option matters

Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.

Correct answer & explanation

✓

/v1/chat/completions

The /v1/chat/completions endpoint is designed for conversational LLMs and accepts a structured messages array. It is the primary interface for interacting with hosted models on NVIDIA API Catalog. Other endpoints serve different purposes like listing models or generating embeddings, and the legacy completions endpoint is less suitable for chat-based models.

Answer analysis

Option-by-option breakdown

For each option: why learners choose it and why it is or isn't the right answer here.

  • ✓

    /v1/chat/completions

    Why this is correct

    The /v1/chat/completions endpoint is the standard OpenAI-compatible API for chat-based interactions. It accepts a messages array and returns model-generated responses. NVIDIA API Catalog endpoints follow this convention, making it the correct choice for sending prompts and receiving completions.

  • ✗

    /v1/models

    Why it's wrong here

    The /v1/models endpoint lists available models but does not generate completions. It is used for discovery, not inference. Sending a prompt here would not return a generated response, so it does not meet the developer's need.

  • ✗

    /v1/completions

    Why it's wrong here

    The /v1/completions endpoint is for legacy text completion models that take a prompt string and return a completion. While it can generate text, NVIDIA's API Catalog primarily promotes chat completions for conversational LLMs. The chat endpoint is more appropriate and widely supported for modern models.

  • ✗

    /v1/embeddings

    Why it's wrong here

    The /v1/embeddings endpoint converts text into vector representations. It is used for tasks like semantic search, not for generating text completions. This endpoint would return embeddings, not a natural language response, so it is incorrect for this scenario.

About these practice questions

This NCA-GENL question is part of Courseiva's 367-question bank — original exam-style content with full explanations and wrong-answer analysis, never real exam questions or exam dumps. Learn why practice questions differ from exam dumps →

How Courseiva writes practice questions · Editorial policy

JA

Written and reviewed by Johnson Ajibi, MSc IT Security

Senior Network & Security Engineer · founder of Courseiva

Last reviewed September 2026 · checked against the official NVIDIA exam blueprint

This NCA-GENL practice question is part of Courseiva's free NVIDIA certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the NCA-GENL exam.