₩ Cost & build-vs-buy
Korean LLM cost calculator — API vs self-host
Estimate monthly cost by call volume, compare API per-token pricing against self-hosting open-weight models, and see where the crossover is for your workload.
✦Independent · zero vendor funding·Data verified 2026-07-13·Methodology & sources ↗
API models — estimated monthly cost
| Model | $/1M in | $/1M out | Monthly |
|---|---|---|---|
| Solar Pro 3✓검증cheapest | ₩63,000 / $45.00 | ||
| Solar Pro 4✓검증 | ₩126,000 / $90.00 | ||
| Gemini (참고)참고 | ₩525,000 / $375 | ||
| GPT (참고)참고 | ₩1,050,000 / $750 | ||
| Claude (Sonnet급 참고)참고 | ₩1,470,000 / $1,050 |
Self-host (open weights) — infra cost, volume-independent
- ▪A.X 4.0 (72B) — Open weights · self-host GPU infra (check the official license before commercial use)
- ▪EXAONE 4.0 32B — Open weights · verify license terms · self-host infra
- ▪Trillion Tri-7B — Lightweight open weights · minimal infra
Solar Pro 4 and Solar Pro 3 standard list prices were verified against Upstage on 2026-08-23; temporary promotions are excluded. Others are editable reference defaults — confirm on official pages. HyperCLOVA X is console-priced (enter your rate). A.X reports ~33% better Korean token efficiency than GPT-4o, so adjust token counts to your logs rather than assuming it. Self-host trades per-token cost for fixed GPU infra; the crossover depends on volume.
How to read it
- ›API cost scales with volume; self-host is a fixed GPU infra cost — high volume favors self-host, low/spiky volume favors API.
- ›Korean token efficiency matters: Korea-tuned models can use ~30% fewer Korean tokens, lowering real cost below the headline per-token price.
- ›Open-weight terms are model/version specific; A.X 4.0 and EXAONE 4.0 require an official license check before commercial or self-hosted adoption.