---
title: Thinking Machines: Inkling API | Text Generation | ModelsLab
description: Build with Thinking Machines: Inkling API by Thinking Machines. Text generation, chat, code, and reasoning. Plans from $21/month.
url: https://modelslab.com/models/thinkingmachines/thinkingmachines-inkling.md
canonical: https://modelslab.com/models/thinkingmachines/thinkingmachines-inkling
type: product
component: Playground/LLM/Index
generated_at: 2026-09-11T04:50:14.455644Z
---

Thinking Machines: Inkling
---

 [LLMs.txt](https://modelslab.com/models/thinkingmachines/thinkingmachines-inkling/llms.txt) [.md](https://modelslab.com/models/thinkingmachines/thinkingmachines-inkling.md)

Thinking Machines: Inkling
---

Choose a prompt below to get started or type your own message

 Explain quantum computing in simple terms

 Write a Python function to sort a list

 Create a marketing email for a SaaS product

 Compare REST vs GraphQL APIs

images and text files

Send

### Thinking Machines: Inkling

thinkingmachines thinkingmachines-inkling

Copy model ID

PricingInput $1.00 / 1M tokens

Output $4.05 / 1M tokens

API EndpointsOpenAI Compatible

`https://modelslab.com/api/v7/llm/chat/completions`Endpoint

Anthropic Compatible

`https://modelslab.com/api/v7/llm/v1/messages`Messages

`https://modelslab.com/api/v7/llm/v1/messages/count_tokens`Count Tokens

`https://modelslab.com/api/v7/llm/v1/models`Models

Use with Claude Code

cURL Example

ParametersSystem MessageYou are a helpful AI assistant specialized in providing accurate and detailed responses.

Temperature0.7

Max Tokens4000

Top P0.9

Frequency Penalty0

Presence Penalty0

Thinkinglow

lowmediumhigh

Deeper thinking spends more of the Max Tokens budget before the answer starts.

Model Info

Support

Looking for something different?
---

Models that trade off differently against the one you are viewing

### Better quality

Newer flagship models that generally produce stronger results.

[![GPT 5.2](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/2bdd00f4-1692-49d8-a562-72cb1753d0ff.png)](https://modelslab.com/models/openai/gpt-5.2)[OpenAI](https://modelslab.com/models/openai)

 [GPT 5.2](https://modelslab.com/models/openai/gpt-5.2)

[![GPT 5.1](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/fad83520-47a8-4059-b4b9-b32adf4d0be7.png)](https://modelslab.com/models/openai/gpt5.1)[OpenAI](https://modelslab.com/models/openai)

 [GPT 5.1](https://modelslab.com/models/openai/gpt5.1)

[![Gemini 3 Pro Preview](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/1846e76f-9d5c-4a35-a43a-82db45b7bb8b.png)](https://modelslab.com/models/google/gemini-3-pro-preview)[Google](https://modelslab.com/models/google)

 [Gemini 3 Pro Preview](https://modelslab.com/models/google/gemini-3-pro-preview)

[![GPT-5.1-Codex-Max](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/06356c4c-df09-4ccb-a5c0-629da670f75b.png)](https://modelslab.com/models/openai/gpt-5.1-codex-max)[OpenAI](https://modelslab.com/models/openai)

 [GPT-5.1-Codex-Max](https://modelslab.com/models/openai/gpt-5.1-codex-max)

### Lower cost per run

Open-source models we host, unlimited on the $149/month plan.

[Modelslab: Custom LLM

Popular](https://modelslab.com/models/modelslab/uncensored-chat)[ModelsLab](https://modelslab.com/models/modelslab)

 [Modelslab: Custom LLM

Open Source Model](https://modelslab.com/models/modelslab/uncensored-chat)

[Qwen: Qwen3 235B A22B Instruct 2507](https://modelslab.com/models/qwen/qwen-qwen3-235b-a22b-2507)[Qwen](https://modelslab.com/models/open_router)

 [Qwen: Qwen3 235B A22B Instruct 2507

Closed Source Model](https://modelslab.com/models/qwen/qwen-qwen3-235b-a22b-2507)

[![Gemini 2.5 Flash](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/4fc6b09e-c7ae-4301-aed1-0b180b88157b.png)](https://modelslab.com/models/google/gemini-2.5-flash)[Google](https://modelslab.com/models/google)

 [Gemini 2.5 Flash

Closed Source Model](https://modelslab.com/models/google/gemini-2.5-flash)

[![gpt-oss-20b](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/8b2f0440-52cc-4f41-93eb-d31006ad71e2.webp)](https://modelslab.com/models/openai/gpt-oss-20b)[OpenAI](https://modelslab.com/models/openai)

 [gpt-oss-20b

Closed Source Model](https://modelslab.com/models/openai/gpt-oss-20b)

[![DeepSeek: DeepSeek V4 Flash 0423](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/267a7228-f392-4d3e-abb6-915fef640c49.png)](https://modelslab.com/models/deepseek/deepseek-deepseek-v4-flash)[DeepSeek](https://modelslab.com/models/open_router)

 [DeepSeek: DeepSeek V4 Flash 0423

Closed Source Model](https://modelslab.com/models/deepseek/deepseek-deepseek-v4-flash)

[![Google: Gemini 3.1 Flash Lite](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/05dab216-0dbb-4160-8e12-6c2984e61eea.png)](https://modelslab.com/models/google/google-gemini-3.1-flash-lite)[Google](https://modelslab.com/models/open_router)

 [Google: Gemini 3.1 Flash Lite

Closed Source Model](https://modelslab.com/models/google/google-gemini-3.1-flash-lite)

Related Models
---

Discover similar models you might be interested in

 [View all LLM Models](https://modelslab.com/models?feature=llmaster)

[#### Google/Google: Gemini 2.5 Pro

google-gemini-2.5-pro

From $5.63/M tokens](https://modelslab.com/models/google/google-gemini-2.5-pro)

[#### Qwen/Qwen: Qwen3.7 Flash

qwen-qwen3.7-flash

From $0.08/M tokens](https://modelslab.com/models/qwen/qwen-qwen3.7-flash)

[#### Poolside/Poolside: Laguna XS 2.1

poolside-laguna-xs-2.1

From $0.09/M tokens](https://modelslab.com/models/poolside/poolside-laguna-xs-2.1)

[#### DeepSeek/DeepSeek: DeepSeek V4 Flash 0731 (batch)

deepseek-deepseek-v4-flash-0731-batch

From $0.22/M tokens](https://modelslab.com/models/deepseek/deepseek-deepseek-v4-flash-0731-batch)

[#### OpenAI/OpenAI: gpt-oss-120b (batch)

openai-gpt-oss-120b-batch

From $0.38/M tokens](https://modelslab.com/models/openai/openai-gpt-oss-120b-batch)

[#### Gryphe/MythoMax 13B

gryphe-mythomax-l2-13b

From $0.06/M tokens](https://modelslab.com/models/gryphe/gryphe-mythomax-l2-13b)

[#### Qwen/Qwen2.5 72B Instruct Turbo

Qwen-Qwen2.5-72B-Instruct-Turbo

From $1.20/M tokens](https://modelslab.com/models/qwen/Qwen-Qwen2.5-72B-Instruct-Turbo)

[#### Google/Google: Gemini 2.5 Pro Preview 06-05

google-gemini-2.5-pro-preview

From $5.63/M tokens](https://modelslab.com/models/google/google-gemini-2.5-pro-preview)

[#### Refuel AI/Refuel LLM V2 Small

togethercomputer-Refuel-Llm-V2-Small

From $0.24/M tokens](https://modelslab.com/models/refuel_ai/togethercomputer-Refuel-Llm-V2-Small)

[#### DeepSeek/DeepSeek: DeepSeek V3.2 Exp

deepseek-deepseek-v3.2-exp

From $0.34/M tokens](https://modelslab.com/models/deepseek/deepseek-deepseek-v3.2-exp)

[#### Qwen/Qwen: Qwen3.5-Flash

qwen-qwen3.5-flash-02-23

From $0.16/M tokens](https://modelslab.com/models/qwen/qwen-qwen3.5-flash-02-23)

[#### Google/Gemini 2.5 Pro

gemini-2.5-pro

From $5.63/M tokens](https://modelslab.com/models/google/gemini-2.5-pro)

[#### Nous Research/Nous: Hermes 3 70B Instruct

nousresearch-hermes-3-llama-3.1-70b

From $0.70/M tokens](https://modelslab.com/models/nousresearch/nousresearch-hermes-3-llama-3.1-70b)

[#### Relace/Relace: Relace Search

relace-relace-search

From $2.00/M tokens](https://modelslab.com/models/relace/relace-relace-search)

[#### Qwen/Qwen: Qwen3.6 Plus

qwen-qwen3.6-plus

From $1.14/M tokens](https://modelslab.com/models/qwen/qwen-qwen3.6-plus)

[#### Mistral AI/Mistral: Mistral Large 3 2512

mistralai-mistral-large-2512

From $1.00/M tokens](https://modelslab.com/models/mistralai/mistralai-mistral-large-2512)

[#### Qwen/Qwen: Qwen3 Max

qwen-qwen3-max

From $2.34/M tokens](https://modelslab.com/models/qwen/qwen-qwen3-max)

[#### Qwen/Qwen: Qwen3.5 397B A17B

qwen-qwen3.5-397b-a17b

From $2.02/M tokens](https://modelslab.com/models/qwen/qwen-qwen3.5-397b-a17b)

About Thinking Machines: Inkling
---

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

### Technical Specifications

Model IDthinkingmachines-inkling

CategoryLLM Models

TaskText Generation

Price$2.52 per million tokens

AddedJuly 17, 2026

### Key Features

- Chat completion and multi-turn conversation API
- Streaming response with token-by-token output
- Function calling and tool use support
- System prompts and role-based messaging
- JSON mode and structured output

### Quick Start

Integrate Thinking Machines: Inkling into your application with a single API call. Get your API key from the [pricing page](https://modelslab.com/pricing) to get started.

PythonJavaScriptcURLPHP

```
<code>import requests
import json

url = "https://modelslab.com/api/v7/llm/chat/completions"

headers = {
    "Content-Type": "application/json"
}

data = {
        "model_id": "thinkingmachines-inkling",
        "messages": [
            {
                "role": "user",
                "content": "Hello!"
            }
        ],
        "max_tokens": 1000,
        "key": "YOUR_API_KEY"
    }

try:
    response = requests.post(url, headers=headers, json=data)
    response.raise_for_status()  # Raises an HTTPError for bad responses (4XX or 5XX)
    result = response.json()
    print("API Response:")
    print(json.dumps(result, indent=2))
except requests.exceptions.HTTPError as http_err:
    print(f"HTTP error occurred: {http_err} - {response.text}")
except Exception as err:
    print(f"Other error occurred: {err}")</code>
```

### Pricing

Thinking Machines: Inkling API costs $2.52 per million tokens. Pay only for what you use with no minimum commitments. [View pricing plans](https://modelslab.com/pricing)

### Use Cases

- AI chatbots and virtual assistants
- Code generation and developer tools
- Content writing and copywriting automation
- Data analysis, summarization, and extraction

[Browse LLM Models](https://modelslab.com/models?feature=llmaster) [More from Thinking Machines](https://modelslab.com/models/open_router) [View Pricing](https://modelslab.com/pricing)

Thinking Machines: Inkling FAQ
---

### What is Thinking Machines: Inkling?

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

### How do I use the Thinking Machines: Inkling API?

You can integrate Thinking Machines: Inkling into your application with a single API call. Sign up on ModelsLab to get your API key, then use the model ID "thinkingmachines-inkling" in your API requests. We provide SDKs for Python, JavaScript, and cURL examples in the API documentation.

### How much does Thinking Machines: Inkling cost?

Thinking Machines: Inkling costs $2.52 per million tokens. ModelsLab plans start at $21/month (Basic), and the $149/month Open Source plan includes unlimited generation on open-source models.

### What is the Thinking Machines: Inkling model ID?

The model ID for Thinking Machines: Inkling is "thinkingmachines-inkling". Use this ID in your API requests to specify this model.

### Do I need a paid plan to use Thinking Machines: Inkling?

Yes. ModelsLab is subscription-based with no free tier — plans start at $21/month (Basic, 3,250 API calls). The $149/month Open Source Unlimited plan includes unlimited generation across open-source models; premium third-party models are billed per call from your wallet.

---

*This markdown version is optimized for AI agents and LLMs.*

**Links:**
- [Website](https://modelslab.com)
- [API Documentation](https://docs.modelslab.com)
- [Blog](https://modelslab.com/blog)

---
*Generated by ModelsLab - 2026-09-11*