docs
models / Meta / Llama Guard 3 8B

Llama Guard 3 8B

textmeta/llama-guard-3-8bZDR supported

Llama Guard 3 is a Llama-3.1-8B pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM inputs (prompt classification) and in LLM responses (response classification). It acts as an LLM – it generates text in its output that indicates whether a given prompt or response is safe or unsafe, and if unsafe, it also lists the content categories violated.

$0.03list priceper 1M output tokens
pricing
unitlist price
per 1M input tokens$0.484
per 1M output tokens$0.03

Everyone pays the list price. Reaching Elite (hold or stake 1,337 $XAVA, or hold 1,337,000 $XINF) unlocks the Elite price and a bigger buyback of your token. How it works

openai sdk (python)
from openai import OpenAI

client = OpenAI(
    base_url="https://zinf.ai/v1",
    api_key="xk_live_...",  # your Xava Inference key
)

r = client.chat.completions.create(
    model="meta/llama-guard-3-8b",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(r.choices[0].message.content)
curl
curl https://zinf.ai/v1/chat/completions \
  -H "Authorization: Bearer $XINF_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"meta/llama-guard-3-8b","messages":[{"role":"user","content":"Hello!"}]}'
limits
providerMeta
typetext
context window131,072 tokens
max output—
prompt storagenone
provider retentionnone ZDR supported
livesoon
requests 24h—
p50 latency—
uptime 7d—