// gpt
GPT-5.6 Luna
The small GPT-5.6. Cheap enough for high-volume routing and classification.
Official list
llmrelay
Input / M tokens
$1.00
$0.50
Output / M tokens
$6.00
$3.00
Context window
—
1,050,000 tokens
Max output
—
128,000 tokens
What it costs you per month
Real-world budget scenarios. Numbers are simple sums — official list price vs llmrelay's 50% off tier.
Usage scenario
Official
llmrelay
Light coding (1M in / 200K out per month)
$2.20
$1.10save $1.10
Heavy Cursor / Cline user (50M in / 5M out per month)
$80.00
$40.00save $40.00
Production RAG (500M in / 20M out per month)
$620.00
$310.00save $310.00
What it's good at
- +Low latency
- +Cheap at volume
- +Solid at classification
Best for
- — Classification
- — Routing
- — Guardrails
- — Batch tagging
Try GPT-5.6 Luna at half the price
Free to create an account, no subscription, $10 minimum top-up — enough to run this model against a real task and judge quality yourself.
Get API key →