multimodal MoE reasoning.
Inkling beside its stablemates and nearest rivals. The ◆ marks the best value in each column across every row shown.
| Model | Intelligence | Coding | Speed | Input $/M | Output $/M | Blended $/M | Context |
|---|---|---|---|---|---|---|---|
| GPT-5.6 Luna | 51.2◆ | 71.4◆ | 184.4◆ | $0.10◆ | $0.60◆ | $0.23◆ | 1.05M◆ |
| Kimi K2.7 Code | 41.9◆ | 60.8◆ | 44.4◆ | $0.95◆ | $4◆ | $1.71◆ | 256K◆ |
| Tencent Hy3 | 41.2◆ | 58.8◆ | 65.2◆ | $0.14◆ | $0.58◆ | $0.25◆ | 262K◆ |
| Inkling ◆ | 40.7◆ | 52.1◆ | 56.5◆ | $1◆ | $4.05◆ | $1.76◆ | 256K◆ |
| Step 3.7 Flash | 30.3◆ | 39.6◆ | 399.5◆ | $0.20◆ | $1.15◆ | $0.44◆ | 256K◆ |
| Inkling Small | —◆ | —◆ | —◆ | $0.50◆ | $1.20◆ | $0.68◆ | 1M◆ |
| DeepSeek V4 Propinned | 44.3◆ | 59.4◆ | 70.9◆ | $0.43◆ | $0.87◆ | $0.54◆ | 1M◆ |
| DeepSeek V4 Flashpinned | 40.3◆ | 56.2◆ | 122◆ | $0.14◆ | $0.28◆ | $0.18◆ | 1M◆ |
The Intelligence Index and its sub-scores, ranked against every scored model in the catalog. A metric that has not been measured for Inkling has been left empty.
How far a month of credits goes on Inkling.
Real coding tasks priced end to end on Inkling, from a quick lookup to a full-repo agent run.
| Task | Tokens in · out | Inkling | GPT-5.6 Luna | Claude Haiku 4.5 |
|---|---|---|---|---|
| Quick lookup / one-liner | 8K · 1K | $0.0061 | $0.0008 | $0.0065 |
| Review a 500-line PR | 60K · 4K | $0.04 | $0.0046 | $0.04 |
| Fix a bug (agent loop) | 180K · 12K | $0.12 | $0.01 | $0.12 |
| Refactor a module | 320K · 20K | $0.21 | $0.02 | $0.21 |
| Full-repo agent run | 900K · 45K | $0.50 | $0.05 | $0.49 |
$1/M input and $4.05/M output, cache reads $0.17/M. In an agent loop most input is cache-read, so the effective input rate is about $0.42/M.cmd --model thinkingmachines/inkling, or type /model in a session and pick it. You can switch mid-session without losing context.Command Code is the AI coding agent that continuously learns your taste. Start for $1.