LLM Reference

Inkling on Tinker

Inkling · Thinking Machines Lab

Open Source

Last refreshed 2026-07-15. Next refresh: weekly.

Why use Inkling on Tinker?

Tinker offers Inkling with pay-as-you-go pricing at $1.87/1M input tokens. Tinker is Thinking Machines Lab's managed API platform for fine-tuning and sampling supported foundation models.

Compare Inkling across 2 providers to find the best fit for your use case
Input / 1M
$1.87
Output / 1M
$4.68
Cache
read $0.374
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: thinkingmachines/Inkling
Model ID
thinkingmachines/Inkling

Request example

Curated snippets for this provider are not sourced yet. Use Tinker documentation with model ID thinkingmachines/Inkling.

Gotchas

  • Use provider model ID "thinkingmachines/Inkling", not the LLMReference slug "inkling".

Compare Inkling Across Providers

ProviderInput (per 1M)Output (per 1M)
Tinker$1.87$4.68
Baseten API$1.00$4.05

Pricing

TypePrice (per 1M)
Input tokens$1.87
Output tokens$4.68

Capabilities

VisionMultimodalReasoningJSON / Tool useAudioFine-tuning

About Inkling

Inkling is Thinking Machines Lab's Apache-2.0 open-weight general-purpose multimodal model. It accepts text, image, and audio inputs and generates text with a 1M-token context window. Compare it for Coding, RAG, Agents, Long context, Vision, and JSON / Tool use.

Get Started

Model Specs

Released2026-07-15
Parameters975B total, 41B active
Context1m
ArchitectureMixture of Experts