AI Prosumer
EN
nvidia/llama3-chatqa:latest

A model from NVIDIA based on Llama 3 that excels at conversational question answering (QA) and retrieval-augmented generation (RAG).

Input → Output
Text Text
Input / Output · per 1M tokens
Not available in Credits
Context window
8,192 tokens
License: open_weights

From model to application

Quick Start

API reference
  1. Get your API key

    Create a key in the ShareAI Console, then save it as an environment variable on your server.

    Create API key
    Environment variable
    export SHAREAI_API_KEY="your-api-key"
  2. Choose your endpoint

    Choose a tag available in this access view to see its API request.

Pricing

No Credits offer for this tag

Choose another Credits tag or explore Token Exchange in Network.

Explore Network →

A little more detail

Frequently asked questions

What is llama3-chatqa?

A model from NVIDIA based on Llama 3 that excels at conversational question answering (QA) and retrieval-augmented generation (RAG).

How much does llama3-chatqa cost?

A per-token price has not been published for this tag. Check your Console for available access and current pricing.

What inputs and outputs does llama3-chatqa support?

Inputs: Text. Outputs: Text. These capabilities apply to the selected model tag. Check the endpoint documentation for supported request formats.

What is the context length of llama3-chatqa?

This tag has a context window of 8,192 tokens. Context size and pricing thresholds are separate. Request limits can be lower for a particular deployment.

How do I use llama3-chatqa with ShareAI?

Create an API key in the ShareAI Console, choose a model tag, and send an authenticated request to the endpoint in Quick Start. Your key must have access to the chosen model. Keep your key on your server.