Skip to main content
Core.Today
|
OpenAIMediumUltra

GPT-5.3 Codex

Codex model on the GPT-5.3 base, optimized for agentic coding harnesses and served via the OpenAI Responses API. 400K token context window, 128K max output tokens, cached inputs at a 90% discount.

3,252/26,012credits
input / output ยท per 1M tokens
Codex model on the GPT-5.3 base
Optimized for agentic coding harnesses (Codex-style)
Served via the Responses API - this gateway routes Codex models to /v1/responses, not chat completions
400K context window
128K max output tokens
Cached input tokens for cost savings on repeated context

Run it right now

Test this model instantly in the Console Playground โ€” no code required

Sign in to try

Use with AI Assistant

Copy usage instructions for Claude, ChatGPT, or other AI

Model Specifications

Context Window
400K
tokens
Max Output
128K
tokens
Training Cutoff
Not published
Compatible SDK
OpenAI

Capabilities

Vision
Function Calling
Streaming
JSON Mode
System Prompt

Token Pricing (per 1M tokens)

Token TypeCreditsUSD Equivalent
Input Tokens3,252$2.17
Output Tokens26,012$17.34
Cached Tokens325.15$0.22

* 1,500 credits โ‰ˆ $1 (actual charges may vary based on usage)

Quick Start

curl -X POST "https://api.core.today/llm/openai/v1/responses" \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer cdt_your_api_key" \
  -d '{
  "model": "gpt-5.3-codex",
  "input": "Diagnose why this async job queue occasionally drops tasks under load, then propose and implement a fix:\n\n<code snippet>",
  "max_output_tokens": 6000
}'

Parameters

ParameterTypeRequiredDefaultDescription
modelstringYesgpt-5.3-codexModel identifier
inputstring | arrayYes-Responses API input: a plain string or an array of input items (messages, tool results)
instructionsstringNo-System-level instructions for the run
max_output_tokensintegerNo-Maximum tokens in the response, including internal reasoning tokens (Responses API field; not max_tokens)
streambooleanNofalseEnable Server-Sent Events streaming

Examples

Complex Coding Task

Responses API request using GPT-5.3 Codex

curl -X POST "https://api.core.today/llm/openai/v1/responses" \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer cdt_your_api_key" \
  -d '{
  "model": "gpt-5.3-codex",
  "input": "Diagnose why this async job queue occasionally drops tasks under load, then propose and implement a fix:\n\n<code snippet>",
  "max_output_tokens": 6000
}'

Tips & Best Practices

1Works with the OpenAI SDK - set base_url to https://ai.api.core.today/llm/openai/v1 and use client.responses.create()
2Codex models are served via the Responses API (POST /llm/openai/v1/responses), not chat completions
3Cached input tokens are billed at a 90% discount ($0.175/M) - reuse long repository context across turns

Use Cases

Coding agents and autonomous dev loops
Complex debugging and architecture-level changes
Large refactors across a repository
Code review automation