llm-bill

API cost calculator · prices checked 2026-10-11 · runs measured 2026-10-10 to 2026-10-11

Same price, different bill Claude Haiku 5.5 vs GPT-6 Luna

Claude Haiku 5.5 and GPT-6 Luna have the same list price, but their tokenizers split text differently: for the same English text, Haiku counts ×1.60 as many tokens as Luna. Enter your task to see what each would actually bill.

List price, both models
$0.10 in · $0.50 outper million tokens
Haiku tokens per Luna token, same text
EN ×1.60 · JA ×1.26code and JSON ×1.56–×1.62
Higher price tier starts at
100K · 272Kinput tokens per request (Haiku · Luna)

Your task

Example scenarios

Approximates our measured run. Measured: $0.237 (Claude Haiku 5.5) and $0.159 (GPT-6 Luna) per 1,000 requests.

≈ Haiku 217 · Luna 172 input tokens per request

Both APIs add schema instructions to the input: about 162 + 26 per field for Haiku, 14 + 9 per field for Luna (string fields, measured).

Output characters per request (the two models may write different lengths)

≈ Haiku 402 · Luna 215 output tokens per request, before reasoning

Reasoning tokens per request (billed as output)

The scenarios use reasoning tokens from our medium-effort runs: Luna reports them, Haiku's are estimated (assumption). Your own logs are the best source.

Your monthly bill

Claude Haiku 5.5 costs 49% more than GPT-6 Luna: $7.29 vs $4.91 a month

The main reason is reply length: Claude Haiku 5.5 writes about 442 characters per reply, GPT-6 Luna about 297, and output costs more ($0.50 per million output tokens against $0.10 for input).

Claude Haiku 5.5

Input$0.10/M217 tok$0.650
Cache reads$0.01/M0 tok$0
Output$0.50/M402 tok$6.02
Reasoning$0.50/M41 tok$0.620

$7.29 a month $0.243 per 1,000 requests

GPT-6 Luna

Input$0.10/M172 tok$0.515
Cache reads$0.01/M0 tok$0
Output$0.50/M215 tok$3.23
Reasoning$0.50/M77 tok$1.16

$4.91 a month $0.164 per 1,000 requests

How the gap changes as each input gets longerClaude Haiku 5.5's bill divided by GPT-6 Luna's, with every other setting fixed. Dashed lines mark where each model's higher price tier starts; the solid line is your current input length. The curve stops where a model's input limit is reached.With your other settings, the higher price tier starts for Claude Haiku 5.5 from about 109,898 characters and for GPT-6 Luna from about 375,260 characters of input per request.
×0×2×4×61001K10K100K1Minput characters per request →same costClaude Haiku 5.5 tier ≈109,898GPT-6 Luna tier ≈375,260Haiku ÷ Lunayou: ×1.49
Show the curve as a table
Input characters per requestHaiku ÷ Luna
1,000×1.42
10,000×1.30
50,000×1.26
100,000×1.26
200,000×6.29
500,000×3.14
1,000,000×3.14

What three real tasks cost

Same prompts, medium effort on both models, costs recomputed from the billed tokens at current list prices. The scenarios above approximate these runs. Why the gaps differ

Cost per 1,000 requests; accuracy and latency listed Haiku / Luna
TaskRequestsClaude Haiku 5.5GPT-6 LunaHaiku ÷ LunaAccuracy / blind judgingMedian latencyMeasured
Ticket classification (English and Japanese)50$0.033$0.033×0.98100% / 100% correct0.90 s / 1.45 s2026-10-11
Invoice extraction (structured output, 6 fields)50$0.147$0.090×1.64100% / 100% correct1.72 s / 2.23 s2026-10-10
Japanese business email25$0.237$0.159×1.49GPT-6 Luna preferred: 22/25, 25/252.44 s / 3.56 s2026-10-10

Email judging: Claude Opus 5.5 and GPT-6.1 Sol each compared the two replies blind, in both orders.

How the calculator works

Characters per token by content type
ContentClaude Haiku 5.5GPT-6 LunaHaiku ÷ Luna tokens95% intervalSamples
English2.604.17×1.6041.552–1.68324
Japanese1.101.38×1.2611.235–1.28624
Chinese1.001.39×1.3931.375–1.41324
JavaScript / TypeScript2.203.56×1.6161.559–1.68124
Python2.654.14×1.5611.527–1.59224
JSON2.113.39×1.6081.517–1.69015