Models
Three aliases, one endpoint. Each converts tokens to credits at its own weight, so the model you pick is the price you pay.
The three models
luna
Fast drafts, classification, summaries and high-volume chat.
BEST FOR
- Classification
- Data extraction
- High-volume automation
terra
Everyday assistant work — RAG, agents, tools and coding.
BEST FOR
- Coding
- Data analysis
- RAG & documents
sol
Deep reasoning, hard coding, Codex-style tasks and research.
BEST FOR
- Deep research
- Hard coding & Codex tasks
- Complex agent workflows
Credit weights
Short context
| MODEL | INPUT | CACHE READ | CACHE WRITE | OUTPUT + REASONING |
|---|---|---|---|---|
| luna | 1 | 0.55 | 1.25 | 6 |
| terra | 10 | 5.5 | 12.5 | 60 |
| sol | 25 | 13.75 | 31.25 | 150 |
Long context
| MODEL | INPUT | CACHE READ | CACHE WRITE | OUTPUT + REASONING |
|---|---|---|---|---|
| luna | 2 | 1.1 | 2.5 | 9 |
| terra | 20 | 11 | 25 | 90 |
| sol | 50 | 27.5 | 62.5 | 225 |
A request with more than 272,000 input tokens converts at the long-context rate. Plans and pricing · Your first request