Unlimited Open Source Models

Get Plan
Skip to main content

Thinking Machines: Inkling (batch)

Thinking Machines: Inkling (batch)

Choose a prompt below to get started or type your own message

About Thinking Machines: Inkling (Batch)

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

Technical Specifications

Model ID
thinkingmachines-inkling-batch
Category
LLM Models
Task
Text Generation
Price
$1.26 per million tokens
Added
August 6, 2026

Key Features

  • Chat completion and multi-turn conversation API
  • Streaming response with token-by-token output
  • Function calling and tool use support
  • System prompts and role-based messaging
  • JSON mode and structured output

Quick Start

Integrate Thinking Machines: Inkling (Batch) into your application with a single API call. Get your API key from the pricing page to get started.

import requests
import json
url = "https://modelslab.com/api/v7/llm/chat/completions"
headers = {
"Content-Type": "application/json"
}
data = {
"model_id": "thinkingmachines-inkling-batch",
"messages": [
{
"role": "user",
"content": "Hello!"
}
],
"max_tokens": 1000,
"key": "YOUR_API_KEY"
}
try:
response = requests.post(url, headers=headers, json=data)
response.raise_for_status() # Raises an HTTPError for bad responses (4XX or 5XX)
result = response.json()
print("API Response:")
print(json.dumps(result, indent=2))
except requests.exceptions.HTTPError as http_err:
print(f"HTTP error occurred: {http_err} - {response.text}")
except Exception as err:
print(f"Other error occurred: {err}")

Pricing

Thinking Machines: Inkling (Batch) API costs $1.26 per million tokens. Pay only for what you use with no minimum commitments. View pricing plans

Use Cases

  • AI chatbots and virtual assistants
  • Code generation and developer tools
  • Content writing and copywriting automation
  • Data analysis, summarization, and extraction

Thinking Machines: Inkling (Batch) FAQ

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

You can integrate Thinking Machines: Inkling (Batch) into your application with a single API call. Sign up on ModelsLab to get your API key, then use the model ID "thinkingmachines-inkling-batch" in your API requests. We provide SDKs for Python, JavaScript, and cURL examples in the API documentation.

Thinking Machines: Inkling (Batch) costs $1.26 per million tokens. ModelsLab plans start at $21/month (Basic), and the $149/month Open Source plan includes unlimited generation on open-source models.

The model ID for Thinking Machines: Inkling (Batch) is "thinkingmachines-inkling-batch". Use this ID in your API requests to specify this model.

Yes. ModelsLab is subscription-based with no free tier — plans start at $21/month (Basic, 3,250 API calls). The $149/month Open Source Unlimited plan includes unlimited generation across open-source models; premium third-party models are billed per call from your wallet.