docs
models / OpenAI / GPT-6 Astra

GPT-6 Astra

textopenai/gpt-6-astra

GPT-6 Astra is OpenAI's most capable model, built for complex reasoning, coding, computer use, research, and document creation.

$50.00list priceper 1M output tokens (short context)
pricing
unitlist price
per 1M cache write tokens (long context)$25.00
per 1M cache write tokens (short context)$12.00
per 1M cached input tokens (long context)$2.00
per 1M cached input tokens (short context)$1.00
per 1M input tokens (long context)$20.00
per 1M input tokens (short context)$10.00
per 1M output tokens (long context)$75.00
per 1M output tokens (short context)$50.00

Everyone pays the list price. Reaching Elite (hold or stake 1,337 $XAVA, or hold 1,337,000 $XINF) unlocks the Elite price and a bigger buyback of your token. How it works

openai sdk (python)
from openai import OpenAI

client = OpenAI(
    base_url="https://zinf.ai/v1",
    api_key="xk_live_...",  # your Xava Inference key
)

r = client.chat.completions.create(
    model="openai/gpt-6-astra",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(r.choices[0].message.content)
curl
curl https://zinf.ai/v1/chat/completions \
  -H "Authorization: Bearer $XINF_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"openai/gpt-6-astra","messages":[{"role":"user","content":"Hello!"}]}'
limits
providerOpenAI
typetext
context window1,050,000 tokens
max output—
prompt storagenone
provider retentionthe model provider's policy
livesoon
requests 24h—
p50 latency—
uptime 7d—