Skip to content
Dashboard

Inkling Small

Inkling-Small is a lighter-weight model with 12B active parameters, trained with a similar recipe, to Inkling that achieves strong performance with even lower cost and latency. Your use subject to Thinkingmachines's Terms & Privacy Policies.

ReasoningTool UseVision (Image)File InputImplicit Caching
import { streamText } from 'ai'
const result = streamText({
model: 'thinkingmachines/inkling-small',
prompt: 'Why is the sky blue?'
})
Read docs
Latency

P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds. View the docs for more info.