The smallest, lowest-cost tier of OpenAI's GPT-6 generation, at $0.10/$0.50 per million tokens โ built for high-volume, latency-sensitive work such as classification, extraction, routing and short-form generation. Cached inputs at a 90% discount. Prompts above 272K input tokens are billed at 2x input / 1.5x output for the whole request.
Test this model instantly in the Console Playground โ no code required
Copy usage instructions for Claude, ChatGPT, or other AI
| Token Type | Credits | USD Equivalent |
|---|---|---|
| Input Tokens | 186 | $0.12 |
| Output Tokens | 929 | $0.62 |
| Cached Tokens | 18.58 | $0.01 |
* 1,500 credits โ $1 (actual charges may vary based on usage)
curl -X POST "https://api.core.today/llm/openai/v1/chat/completions" \
-H "Content-Type: application/json" \
-H "Authorization: Bearer cdt_your_api_key" \
-d '{
"model": "gpt-6-luna",
"messages": [
{
"role": "system",
"content": "Classify the support ticket into one of: billing, bug, feature_request, account. Reply with the label only."
},
{
"role": "user",
"content": "I was charged twice for my subscription this month."
}
],
"reasoning_effort": "low",
"max_completion_tokens": 200
}'| Parameter | Type | Required | Default | Description |
|---|---|---|---|---|
messages | array | Yes | - | Array of message objects with role and content |
model | string | Yes | gpt-6-luna | Model identifier |
max_completion_tokens | integer | No | 4096 | Maximum tokens in response (up to 128000). Note: use max_completion_tokens, not max_tokens |
reasoning_effort | string | No | medium | Reasoning effort level: low, medium, high, xhigh, or max lowmediumhighxhighmax |
temperature | float | No | 1.0 | Sampling temperature (0-2) |
stream | boolean | No | false | Enable Server-Sent Events streaming |
Low-cost ticket triage with GPT-6 Luna
curl -X POST "https://api.core.today/llm/openai/v1/chat/completions" \
-H "Content-Type: application/json" \
-H "Authorization: Bearer cdt_your_api_key" \
-d '{
"model": "gpt-6-luna",
"messages": [
{
"role": "system",
"content": "Classify the support ticket into one of: billing, bug, feature_request, account. Reply with the label only."
},
{
"role": "user",
"content": "I was charged twice for my subscription this month."
}
],
"reasoning_effort": "low",
"max_completion_tokens": 200
}'Reuse a large system prompt with 90% cached input discount
curl -X POST "https://api.core.today/llm/openai/v1/chat/completions" \
-H "Content-Type: application/json" \
-H "Authorization: Bearer cdt_your_api_key" \
-d '{
"model": "gpt-6-luna",
"messages": [
{
"role": "system",
"content": "<large repeated system prompt or codebase context>"
},
{
"role": "user",
"content": "Summarize the open TODOs and rank them by risk."
}
],
"temperature": 0.3,
"max_completion_tokens": 4000
}'POST /llm/openai/v1/chat/completions