Skip to content
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Apps
  • Discover
  • Models
  • Providers
  • Pricing
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Support
  • Works With OR
  • Data

Developer

  • Documentation
  • API Reference
  • SDK
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for thinkingmachines

thinkingmachines

Access 2 thinkingmachines models through the OpenRouter unified API including Inkling Small and Inkling. Compare pricing, context windows, benchmarks, and capabilities between different thinkingmachines models.

thinkingmachines tokens processed on OpenRouter

  • Favicon for thinkingmachines
    Thinking Machines: Inkling SmallInkling Small
    4.31B tokens

    Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of the Inkling family and is suited for reasoning, coding, agentic workflows, retrieval-augmented generation, instruction following, and multilingual conversation.

    by thinkingmachinesJul 30, 2026524K context$0.45/M input tokens$1.20/M output tokens
  • Favicon for thinkingmachines
    Thinking Machines: InklingInkling
    24.7B tokens

    Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems, retrieval-augmented generation, instruction following, and multilingual conversational applications. Its native image and audio understanding supports multimodal analysis alongside text.

    by thinkingmachinesJul 17, 20261.05M context$0.95/M input tokens$4.05/M output tokens