Nano Banana Pro
nano-banana-pro•Company:GoogleHigh-density lightweight foundation model designed for edge devices and fast serverless microservice invocation.
Tokens total window
Tokens per response
Per 1,000,000 input tokens
Per 1,000,000 output tokens
Key Capabilities
- • 100% OpenAI-compatible Chat Completions API endpoint
- • Zero prompt data retention — never used for model training
- • Sub-50ms gateway routing overhead with prompt caching support
- • Atomic multi-provider failover ensures zero service interruptions
OpenAI SDK Drop-in Call
from openai import OpenAI
client = OpenAI(
base_url="https://apihundred.com/v1",
api_key="your-api100-key"
)
response = client.chat.completions.create(
model="nano-banana-pro",
messages=[
{"role": "user", "content": "Explain transactional RLS in Postgres"}
]
)
print(response.choices[0].message.content)curl https://apihundred.com/v1/chat/completions \
-H "Authorization: Bearer $API100_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "nano-banana-pro",
"messages": [{"role": "user", "content": "Hello!"}]
}'Gateway Performance & Verification Telemetry
Measured across 10,000 requests routed through API100 multi-region edge nodes.
Frequently Asked Questions about Nano Banana Pro
How do I call Nano Banana Pro with the OpenAI SDK?
Initialize your standard OpenAI client with baseURL="https://api.apihundred.com/v1" and use your API100 key. Pass model="nano-banana-pro" into client.chat.completions.create().
What are the token prices for Nano Banana Pro?
On API100, Nano Banana Pro is priced at $0.1 per million input tokens and $0.4 per million output tokens with direct pass-through rates and unified wallet billing.
Does API100 store prompts when calling Nano Banana Pro?
Zero prompts or outputs are stored on disk. All requests are streamed via TLS 1.3 in ephemeral volatile memory. Prompts are never saved to server logs or used for training.
What happens if Google has an outage?
API100 provides automated multi-provider failover. You can configure fallback models or enable our intelligent router to automatically re-route requests without application errors.
Alternative & Related Models
View Full CatalogDeepSeek V3
671B parameter Mixture-of-Experts (MoE) activating 37B parameters per token using Multi-Head Latent Attention (MLA).
Gemini 1.5 Pro
Google frontier model featuring industry-leading 2M token context window for massive repos, long docs, and media.
Gemini 3.1 Pro
Google's high-capacity reasoning model featuring 2M tokens context for large-scale code repositories and video archives.

