Model Overview & Capabilities
Inkling is a frontier artificial intelligence model engineered by Thinking Machines, officially verified in February 2026 and evaluated on the Artificial Analysis Intelligence Index v4.3 with an overall score of 26.0. Architected as a proprietary API service, the system operates across a context window of 1000k tokens, allowing enterprise developer agents and autonomous code evaluation tools to ingest comprehensive repository structures, cross-language modules, and multi-layered documentation hierarchies without degradation of semantic coherence.
Under standardized software engineering benchmarks—including SWE-bench Verified, Terminal-Bench 2.0, Aider Polyglot, and LiveCodeBench—Inkling exhibits rigorous multi-step problem solving, deterministic SEARCH/REPLACE diff compliance, and robust tool-use navigation in sandboxed execution environments. Its reasoning capabilities are calibrated to minimize hallucination rates while handling complex syntax transformations, boundary conditions, and continuous integration workflows.
Economically and operationally, Inkling generates output tokens at a median throughput of 81 tokens per second, with an observed initial latency of 2480 milliseconds. Artificial Analysis rates its estimated cost per task at $0.61 USD, backed by an API token pricing structure of $0.5 per million input tokens and $1.5 per million output tokens (resulting in an effective blended rate of $0.75 per million tokens).