Inkling-Small

Multimodal
Thinking Machines Lab

Inkling-Small is Thinking Machines Lab's efficient open-weights MoE multimodal model (276B total / 12B active) released under Apache 2.0. It accepts text, image, and audio inputs and generates text, with native reasoning, variable thinking effort, and a context window up to 1M tokens. It scores 80.2% on SWE-Bench Verified and is served through Tinker at $0.30 / $1.20 per 1M input/output tokens.

Key Specifications

Parameters
276.0B
Context
256.0K
Release Date
July 30, 2026
Average Score
60.3%

Timeline

Key dates in the model's history
Announcement
July 30, 2026
Last Update
August 31, 2026

Technical Specifications

Parameters
276.0B
Training Tokens
-
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval

Pricing & Availability

Input (per 1M tokens)
$0.30
Output (per 1M tokens)
$1.20
Max Input Tokens
256.0K
Max Output Tokens
256.0K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning

Benchmark Results

Model performance metrics across various tests and benchmarks

Programming

Programming skills tests
SWE-Bench Verified
Bash-only harness (same as Inkling 77.6%)Self-reported
80.2%

Reasoning

Logical reasoning and analysis
GPQA
GPQA DiamondSelf-reported
89.5%

Other Tests

Specialized benchmarks
AA-Briefcase
Elo score; Artificial Analysis (max_score 3000)Self-reported
30.6%
AIME 2026
Self-reported
95.5%
ARC-AGI
ARC-AGI-1Self-reported
84.0%
ARC-AGI v2
ARC-AGI-2Self-reported
40.1%
Artificial Analysis
Artificial Analysis Intelligence Index v4.1. Score 40 (vendor model card).Self-reported
40.0%
BrowseComp
With context managementSelf-reported
77.4%
CharXiv-R
CharXiv RQ originalSelf-reported
77.4%
CritPT
Self-reported
8.3%
GDPval-AA
GDPval-AA v2 Elo (max_score 3000)Self-reported
42.3%
Global-MMLU-Lite
Self-reported
86.7%
Humanity's Last Exam
Text onlySelf-reported
31.6%
IFBench
Self-reported
82.2%
MCP Atlas
Public setSelf-reported
79.6%
MMAU
Self-reported
77.0%
MMMU-Pro
MMMU Pro Standard 10Self-reported
74.0%
SciCode
Self-reported
48.7%
SimpleQA Verified
Self-reported
20.6%
SWE-Bench Pro
SWE-Bench Pro publicSelf-reported
55.9%
Tau3 Banking
Self-reported
15.5%
Terminal-Bench 2.1
Vendor best harness (model card)Self-reported
64.7%
Toolathlon
Toolathlon VerifiedSelf-reported
54.4%
VoiceBench Avg
Self-reported
90.1%

License & Metadata

License
apache_2_0
Announcement Date
July 30, 2026
Last Updated
August 31, 2026

Compare Inkling-Small

All comparisons

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.