Thinking Machines: Inkling Small (free): what the free tier actually gets you

Thinking Machines: Inkling Small (free) ranks #11 on the AI Hippo board with a 1.0M context window and a free tier in this snapshot. Specs, cost math, and how it compares.

The quick read

Thinking Machines: Inkling Small (free) from thinkingmachines lands #11 largely because it ships a free tier: a 1.0M window at no token cost is unbeatable on a context-per-dollar board. The real questions are throughput limits and licensing, not sticker price.

Spec sheet at a glance

By the numbers: 1.0M context window; multimodal (text, image, and file input); 11 exposed API parameters, including tool calling with reasoning support. Use these to judge fit for agentic, long-context, or structured-output workloads.

Pricing & how it compares

Pricing: a free tier in this snapshot. It carries a free tier, so token cost is not the deciding factor. On AI Hippo the composite board rewards cheap long-context models such as Meta: Llama 4 Scout, so a lower rank here means weaker value-per-token, not weaker capability.

Head-to-head

Its closest board neighbor is Google: Lyria 3 Pro Preview (1.0M ctx, free tier). Side by side, Thinking Machines: Inkling Small (free) offers 1.0M ctx at free tier: context is similar and, on blended token price, it is level. Pick between them on whichever axis your workload is bound by.

Where it fits

Best-fit workloads: whole-repo & long-document work, screenshot / PDF / image understanding, tool-calling agents, multi-step reasoning, high-volume, cost-sensitive batch.

Cost in practice & verdict

A representative 100K-input + 20K-output task costs about $0.00 on Thinking Machines: Inkling Small (free). Verdict: a low-risk pick for prototyping and high-volume batch work — validate rate limits and licensing terms before production.

Insights