Thinking Machines: Inkling
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
ARCHITECTURE & LIMITS
BENCHMARK EVALUATIONS & CODE METRICS
Verified scores from SWE-bench and Artificial Analysis.
Independent coding benchmark evaluating syntax, algorithms, and code completion
Broad multi-domain reasoning, science, and knowledge retrieval
Autonomous multi-step tool use, API calling, and task execution
TOKEN COST CALCULATOR
Estimate production API costs for Thinking Machines: Inkling.
COMMUNITY ENGINEERING EVALUATIONS
Verified ratings across 4 core software engineering dimensions.
Refactoring accuracy, syntax correctness, and adherence to design patterns.
Zero hallucinations, strict constraint compliance, and negative prompt respect.
Time-to-first-token (TTFT) and throughput tokens per second.
Token pricing economics relative to intelligence quality and output volume.
No text reviews submitted yet. Be the first developer to review Thinking Machines: Inkling!
RECOMMENDED AGENT SKILLS FOR Thinking Machines: Inkling
boardroom
/cs:boardroom <brief> — 6-phase multi-role deliberation across the C-suite with Phase 2 isolation, critic pre-screen, and synthesis. Outputs a board memo. Use when a decision spans multiple executive domains — e.g. a pricing change touching finance, positioning, and product, or a raise-vs-cut runway call.
FREQUENTLY ASKED QUESTIONS ABOUT Thinking Machines: Inkling
What are the input and output token prices for Thinking Machines: Inkling?+
Thinking Machines: Inkling costs $1.00 per 1M input tokens and $4.05 per 1M output tokens. Cached input reads cost $0.17 per 1M tokens.
Does Thinking Machines: Inkling support tool calling and agent execution?+
Yes, Thinking Machines: Inkling natively supports tool calling, parallel function execution, and structured agent workflows.
What is the context window of Thinking Machines: Inkling?+
Thinking Machines: Inkling features an active context window of 1,048,576 tokens with a maximum single-turn output of 471,859 tokens.
What AI Agent Skills work best with Thinking Machines: Inkling?+
Because Thinking Machines: Inkling is optimized for agentic tool calling and multimodal vision and code reasoning, it pairs well with Agent Skills like Boardroom, SEO Audit, and UI/UX Pro Max.
AgenticMarket