Inkling-Small matches or exceeds Inkling on reasoning and agentic tasks, the company said.
Nvidia-backed Thinking Machines Lab has unveiled a new open-weights model that performs “comparabl[y]” to Inkling, but at one-quarter of its size.
The company launched its first AI model Inkling earlier this month following a mega partnership with Nvidia that gave Thinking Machines access to GB300 NVL72 systems for training. The chipmaker also made a significant investment into the AI company.
Inkling-Small comes in at 276bn total parameters, with 12bn active. Compared to its bigger predecessor, the new model achieves comparable performance with much less compute, said Thinking Machines. It matches or exceeds Inkling on reasoning and agentic tasks, it added.
Artificial Analysis scores Inkling-Small 40pc on the Intelligence Index, placing it “well above average” relative to other comparable models. Inkling hits 41pc while DeepSeek V4 Flash matches Inkling-Small at 40pc. At 93 tokens per second, the model is also faster than the average, it said.








