# Thinking Machines: Inkling Small — price and measured latency by host

Updated August 21, 2026. Served by 3 providers; 3 of them measured here.

Latency and uptime are measured by llmlatency.dev probes. Prices are published by the providers and read via OpenRouter on 2026-08-21 — not measured.

| Host | Input $/M | Output $/M | Context | Fastest p50 | From | Uptime |
| --- | --- | --- | --- | --- | --- | --- |
| deepinfra | $0.450 | $1.200 | 524,288 | 234 ms | us-central | 100% |
| baseten | $0.500 | $1.200 | 1,048,576 | 57 ms | us-central | 100% |
| together | $0.500 | $1.200 | 524,288 | 124 ms | us-central | 100% |

Source: https://llmlatency.dev/model/thinkingmachines-inkling-small · CC BY 4.0
