Text generation / OpenAI
gpt-5.6-luna
StreamFunction callingStructured outputs
Overview
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for its price tier.
- Model type: Text generation
- Input: Text / Image
- Output: Text
- Endpoints: chat / chat_completion
Tiered billing details
| Tier | Input | Output | Cache hit | Cache write |
|---|---|---|---|---|
| ≤272K input tokens | $0.0002/1K | $0.0012/1K | $0.0000/1K | $0.0003/1K |
| >272K input tokens | $0.0004/1K | $0.0018/1K | $0.0000/1K | $0.0005/1K |
Billed bylen (input context token count). Coefficient is the $ / 1M tokens price. Default group price (other groups converted by discount).
Code example
Call with gpt-5.6-luna :
curl https://www.starunion.net/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gpt-5.6-luna",
"messages": [{ "role": "user", "content": "Hello" }]
}'