GPT-5.6 家族轻量款,响应迅捷、成本更低,适合高并发与快速问答。
Model ID: gpt-5.6-luna · Type: chat · Provider: OpenAI
Endpoints: /v1/chat/completions · /v1/responses · /v1/messages
| Input (per 1M tokens) | $0.13 USD |
| Output (per 1M tokens) | $0.78 USD |
| Cache read (per 1M tokens) | $0.013 USD |
| Cache write 5m (per 1M tokens) | $0.1625 USD |
| ≤ 272000 tokens | $0.13 / 1M in | $0.78 / 1M out |
| > 272000 | $0.26 / 1M in | $1.17 / 1M out |
from openai import OpenAI
client = OpenAI(api_key="sk-...", base_url="https://api.router.ai/v1")
resp = client.chat.completions.create(
model="gpt-5.6-luna",
messages=[{"role": "user", "content": "Hello"}],
)
print(resp.choices[0].message.content)GPT-5.6 Luna (`gpt-5.6-luna`) is billed per usage at $0.13/1M in · $0.78/1M out, in USD. Current pricing is always listed at https://ai.yunbaozi.com/models/gpt-5.6-luna.
Send a request to https://api.router.ai/v1/v1/chat/completions with the header `Authorization: Bearer <your API key>` and `"model": "gpt-5.6-luna"`. The API is OpenAI-compatible, so any OpenAI SDK works by changing base_url to https://api.router.ai/v1 — no other code change.
GPT-5.6 Luna can be called on: /v1/chat/completions; /v1/responses; /v1/messages.
GPT-5.6 Luna accepts up to 922,000 input tokens and can return up to 128,000 output tokens. Requests exceeding the input limit are rejected before reaching the model.
GPT-5.6 Luna supports: vision, function_calling, prompt_caching.
GPT-5.6 Luna is a chat model from OpenAI, available through the 国际云宝子(Ai) gateway with the same API key as every other model.
Call it through the 国际云宝子(Ai) OpenAI-compatible endpoint (API base: https://api.router.ai/v1). AI agents can discover and call every model on this gateway through MCP (https://mcp.router.ai/mcp) with no manual integration.