Inkling-Small
MultimodalInkling-Small is Thinking Machines Lab's efficient open-weights MoE multimodal model (276B total / 12B active) released under Apache 2.0. It accepts text, image, and audio inputs and generates text, with native reasoning, variable thinking effort, and a context window up to 1M tokens. It scores 80.2% on SWE-Bench Verified and is served through Tinker at $0.30 / $1.20 per 1M input/output tokens.
Key Specifications
Timeline
Technical Specifications
Pricing & Availability
Benchmark Results
Model performance metrics across various tests and benchmarks
Programming
Reasoning
Other Tests
License & Metadata
Compare Inkling-Small
All comparisonsSimilar Models
All ModelsCommand A+
Cohere
Step-3.5-Flash
StepFun
Kimi K3
Moonshot AI
Kimi K2.5
Moonshot AI
Qwen3.8 Flash Next
Alibaba
Qwen3.8 Flash
Alibaba
GLM-4.6
Zhipu AI
MiniMax M2.5
MiniMax
Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.