GPT-6 Luna
gpt-6-luna•Company:OpenAIUltra-fast, high-throughput micro-task model with sub-180ms TTFT for semantic routing, classification, and real-time chat.
Tokens total window
Tokens per response
Per 1,000,000 input tokens
Per 1,000,000 output tokens
Key Capabilities
- • 100% OpenAI-compatible Chat Completions API endpoint
- • Zero prompt data retention — never used for model training
- • Sub-50ms gateway routing overhead with prompt caching support
- • Atomic multi-provider failover ensures zero service interruptions
OpenAI SDK Drop-in Call
from openai import OpenAI
client = OpenAI(
base_url="https://apihundred.com/v1",
api_key="your-api100-key"
)
response = client.chat.completions.create(
model="gpt-6-luna",
messages=[
{"role": "user", "content": "Explain transactional RLS in Postgres"}
]
)
print(response.choices[0].message.content)curl https://apihundred.com/v1/chat/completions \
-H "Authorization: Bearer $API100_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gpt-6-luna",
"messages": [{"role": "user", "content": "Hello!"}]
}'Gateway Performance & Verification Telemetry
Measured across 10,000 requests routed through API100 multi-region edge nodes.
Frequently Asked Questions about GPT-6 Luna
How do I call GPT-6 Luna with the OpenAI SDK?
Initialize your standard OpenAI client with baseURL="https://api.apihundred.com/v1" and use your API100 key. Pass model="gpt-6-luna" into client.chat.completions.create().
What are the token prices for GPT-6 Luna?
On API100, GPT-6 Luna is priced at $0.2 per million input tokens and $0.8 per million output tokens with direct pass-through rates and unified wallet billing.
Does API100 store prompts when calling GPT-6 Luna?
Zero prompts or outputs are stored on disk. All requests are streamed via TLS 1.3 in ephemeral volatile memory. Prompts are never saved to server logs or used for training.
What happens if OpenAI has an outage?
API100 provides automated multi-provider failover. You can configure fallback models or enable our intelligent router to automatically re-route requests without application errors.
Alternative & Related Models
View Full CatalogDeepSeek V3
671B parameter Mixture-of-Experts (MoE) activating 37B parameters per token using Multi-Head Latent Attention (MLA).
GPT-4o
OpenAI flagship omni model with high-speed intelligence across text, vision, and structured schema generation.
GPT-6 Astra
OpenAI's frontier agentic model engineered for autonomous computer use, desktop GUI interaction, and repository refactoring.

