Skip to main content
OpenAI빠름높음

GPT-6 Luna

OpenAI GPT-6 세대에서 가장 작고 저렴한 티어로, 백만 토큰당 $0.10/$0.50입니다. 분류·추출·라우팅·짧은 생성 같은 대량·저지연 작업용이며 90% 할인 캐시 입력을 지원합니다. 입력 272K 토큰 초과 시 요청 전체가 입력 2배·출력 1.5배로 과금됩니다.

186/929크레딧
입력 / 출력 · 100만 토큰당
GPT-6 세대 최저가 티어
1M 토큰 컨텍스트 윈도우
128K 최대 출력 토큰
학습 기준일: 2026년 5월 18일
캐시 입력 가격 (90% 할인)
조절 가능한 추론 노력 수준 (low~max)
함수 호출 및 네이티브 비전 지원

지금 바로 실행해보세요

콘솔의 Playground에서 별도 코드 없이 이 모델을 즉시 테스트할 수 있어요

로그인 후 사용해보기

AI 어시스턴트에서 사용하기

이 모델의 사용법을 Claude, ChatGPT 등에 복사

모델 상세 사양

컨텍스트 윈도우
1M
토큰
최대 출력
128K
토큰
학습 데이터
2026-05
호환 SDK
OpenAI

기능 지원

비전
함수 호출
스트리밍
JSON 모드
시스템 프롬프트

토큰별 가격 (1M 토큰당)

토큰 종류크레딧달러 환산
입력 토큰186$0.12
출력 토큰929$0.62
캐시된 토큰18.58$0.01

* 1,500 크레딧 ≈ $1 (실제 요금은 사용량에 따라 달라질 수 있습니다)

빠른 시작

curl -X POST "https://api.core.today/llm/openai/v1/chat/completions" \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer cdt_your_api_key" \
  -d '{
  "model": "gpt-6-luna",
  "messages": [
    {
      "role": "system",
      "content": "Classify the support ticket into one of: billing, bug, feature_request, account. Reply with the label only."
    },
    {
      "role": "user",
      "content": "I was charged twice for my subscription this month."
    }
  ],
  "reasoning_effort": "low",
  "max_completion_tokens": 200
}'

파라미터

파라미터타입필수기본값설명
messagesarrayYes-role과 content를 포함한 메시지 객체 배열
modelstringYesgpt-6-luna모델 식별자
max_completion_tokensintegerNo4096응답의 최대 토큰 수 (최대 128000). 주의: max_tokens 대신 max_completion_tokens 사용
reasoning_effortstringNomedium추론 노력 수준: low, medium, high, xhigh, max
lowmediumhighxhighmax
temperaturefloatNo1.0샘플링 온도 (0-2)
streambooleanNofalseServer-Sent Events 스트리밍 활성화

예제

대량 분류

GPT-6 Luna를 활용한 저비용 티켓 분류

curl -X POST "https://api.core.today/llm/openai/v1/chat/completions" \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer cdt_your_api_key" \
  -d '{
  "model": "gpt-6-luna",
  "messages": [
    {
      "role": "system",
      "content": "Classify the support ticket into one of: billing, bug, feature_request, account. Reply with the label only."
    },
    {
      "role": "user",
      "content": "I was charged twice for my subscription this month."
    }
  ],
  "reasoning_effort": "low",
  "max_completion_tokens": 200
}'

캐시된 반복 컨텍스트

큰 시스템 프롬프트를 90% 할인된 캐시 가격으로 재사용

curl -X POST "https://api.core.today/llm/openai/v1/chat/completions" \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer cdt_your_api_key" \
  -d '{
  "model": "gpt-6-luna",
  "messages": [
    {
      "role": "system",
      "content": "<large repeated system prompt or codebase context>"
    },
    {
      "role": "user",
      "content": "Summarize the open TODOs and rank them by risk."
    }
  ],
  "temperature": 0.3,
  "max_completion_tokens": 4000
}'

팁 & 모범 사례

1Luna($0.10/$0.50 per M)는 GPT-6 최저가 티어 — 더 깊은 추론이 필요하면 GPT-6 Sol($2/$10)이나 Astra($10/$50)로 승격
2반복되는 시스템 프롬프트와 RAG 컨텍스트에는 캐시 입력($0.01/M, 90% 할인) 활용
3가능하면 입력 272K 토큰 이하로 — 초과 시 요청 전체가 입력 2배·출력 1.5배로 과금
4분류·추출 작업은 reasoning_effort를 low로 두어 지연과 출력 토큰 최소화
5128K 출력으로 한 번에 장문 생성 가능
6코딩 및 분석 작업에는 낮은 온도(0.2-0.5)
7캐시 쓰기는 입력 단가의 1.25배로 과금됩니다 — 암묵 캐싱이 기본으로 쓰기를 만들기 때문에 큰 프롬프트의 첫 요청은 대부분 캐시 쓰기로 과금되고, 이후 읽기는 90% 할인됩니다
8Fast mode(service_tier: fast 또는 priority) 요청은 토큰 요금이 2배입니다 — 응답이 보고한 티어 기준으로 정산하므로 표준으로 처리된 요청은 표준 요금입니다

사용 사례

대량 분류 및 라우팅
구조화 데이터 추출
짧은 생성 및 요약
에이전트 파이프라인의 서브 에이전트·도구 호출 단계
캐시 입력 활용 반복 컨텍스트 (RAG, 코드베이스)