Llama Guard 3 8B
text
meta/llama-guard-3-8bZDR supportedLlama Guard 3 is a Llama-3.1-8B pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM inputs (prompt classification) and in LLM responses (response classification). It acts as an LLM – it generates text in its output that indicates whether a given prompt or response is safe or unsafe, and if unsafe, it also lists the content categories violated.
$0.03list priceper 1M output tokens
pricing
| unit | list → Elite price |
|---|---|
| per 1M input tokens | $0.484 |
| per 1M output tokens | $0.03 |
No Elite discount on this model right now: everyone pays the list price.
Everyone pays the list price. Reaching Elite (hold or stake 1,337 $XAVA, or hold 1,337,000 $XINF) unlocks the Elite price and a bigger buyback of your token. How it works
openai sdk (python)
from openai import OpenAI
client = OpenAI(
base_url="https://zinf.ai/v1",
api_key="xk_live_...", # your Xava Inference key
)
r = client.chat.completions.create(
model="meta/llama-guard-3-8b",
messages=[{"role": "user", "content": "Hello!"}],
)
print(r.choices[0].message.content)curl
curl https://zinf.ai/v1/chat/completions \
-H "Authorization: Bearer $XINF_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"meta/llama-guard-3-8b","messages":[{"role":"user","content":"Hello!"}]}'limits
providerMeta
typetext
context window131,072 tokens
max output—
prompt storagenone
provider retentionnone ZDR supported
livesoon
requests 24h—
p50 latency—
uptime 7d—