Thinking Machines: Inkling Small (batch) · Thinkingmachines
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of the family, ideal for applications requiring fast processing without compromising quality.
About Thinking Machines: Inkling Small (batch)
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of the family, ideal for applications requiring fast processing without compromising quality.
Where it shines: Efficient multimodal processing · Optimized 12B active parameters · Mixture-of-experts architecture.
How to use Thinking Machines: Inkling Small (batch)
-
1
Type or upload
Type what you want in the box above — or upload the file if the tool asks for one.
-
2
Generate
Click the main button. Wait 2-30 seconds depending on the model and input size.
-
3
Download or share
Download the result or share the direct link. No watermark, ready to use.
Frequently asked questions
Thinking Machines: Inkling Small (batch) - 100% Free - 100% Free
Thinking Machines: Inkling Small (batch) is a Pro model: it costs 10 tokens per use (~$0.05 real cost for us). You need a Pro plan ($9/month → 15,000 tokens) or a one-shot pack. If you already have tokens in the free account, you can also spend them directly.
How many Thinking Machines: Inkling Small (batch) uses are included in the Pro plan?
Pro ($9/month) gives you 15,000 recurring tokens. At 10 tokens per use of Thinking Machines: Inkling Small (batch), that's ~1,500 full uses per cycle. If you run out, one-shot packs (5,000 / 25,000 / 80,000 tokens) add to the balance without expiring before one year.
Thinking Machines: Inkling Small (batch) - 100% Original - 100% Original
Its specific strengths are: efficient multimodal processing, optimized 12b active parameters and mixture-of-experts architecture.It accepts images as input (multimodal) as well as text — useful for describing captures, reading graphs or solving problems from a photo.
Thinking Machines: Inkling Small (batch) - Thinking Machines: Inkling Small (batch)
Thinking Machines: Inkling Small (batch) has a balanced speed: 5-15 seconds per response — neither the fastest nor the slowest in the catalog. The actual time also depends on the length of the prompt and the load of the datacenter — models with huge context take longer when you enter very long texts.
How do I use Thinking Machines: Inkling Small (batch) in ia.gratis?
You can use Thinking Machines: Inkling Small (batch) from /chat/ by selecting Thinking Machines: Inkling Small (batch) in the picker, or via the REST API with `model=inkling-small-batch` in the POST body. Quick summary: inkling Small: efficient multimodal model of 12B active parameters out of 276B total. The internal model identifier is `inkling-small-batch` — useful when integrating by API.