Thinkingmachines
Available in Shortly AI
Inkling Small
Released July 30, 2026
ReasoningVision (Image)Tool Use
Inkling-Small is a lighter-weight model with 12B active parameters, trained with a similar recipe, to Inkling that achieves strong performance with even lower cost and latency.
- Context window
- 1M tokens
- Maximum output
- 1M tokens
- Input pricing
- $0.45/M
- Output pricing
- $1.20/M
FAQ
Similar AI models
DeepSeek V4.1 FlashDeepSeekDeepSeek V4.1 Flash is an AI model from DeepSeek designed for fast, efficient multimodal interactions. It combines native multimodal capabilities with a new architecture focused on stronger performance, faster responses, and lower computational cost.Gemini 3.8 FlashGoogleGemini 3.8 Flash is the high-efficiency, cost-effective powerhouse of the Gemini 3 family. It delivers Pro-level agentic capabilities, major leaps in code generation and terminal execution. 3.8 Flash serves as the primary agentic workhorse in the Gemini 3 family, bridging the gap between deep-reasoning Pro models and high-throughput Flash-Lite models while delivering high token efficiency and multi-step multimodal processing.Qwen3.8 27BAlibabaBuilt on the architectural foundation of Qwen3.5, Qwen3.8 delivers substantial gains across coding, professional work, research, and long-horizon agentic tasks. Qwen3.8-27B brings these advances to a compact, deployment-friendly dense model: a native vision-language model that understands images and videos, with flexible thinking control, designed to carry complex, multi-step tasks through to completion with greater reliability.
Available with one subscription
Start with Inkling Small in Shortly AI
Keep your conversations, files and model choices together in one workspace.