---
title: Qwen: Qwen3 8B API | Text Generation | ModelsLab
description: Dense 8.2B parameter LLM with 128K context, hybrid thinking modes for reasoning/math/coding, trained on 36T tokens across 119 languages; excels in efficient inference on consumer GPUs (16GB FP16).[1][2][3]
url: https://modelslab.com/models/qwen/qwen-qwen3-8b.md
canonical: https://modelslab.com/models/qwen/qwen-qwen3-8b
type: product
component: Playground/LLM/Index
generated_at: 2026-09-11T05:17:42.833328Z
---

Qwen: Qwen3 8B
---

 [LLMs.txt](https://modelslab.com/models/qwen/qwen-qwen3-8b/llms.txt) [.md](https://modelslab.com/models/qwen/qwen-qwen3-8b.md)

Qwen: Qwen3 8B
---

Choose a prompt below to get started or type your own message

 Explain quantum computing in simple terms

 Write a Python function to sort a list

 Create a marketing email for a SaaS product

 Compare REST vs GraphQL APIs

text files

Send

### Qwen: Qwen3 8B

qwen qwen-qwen3-8b

Copy model ID

PricingInput $0.117 / 1M tokens

Output $0.455 / 1M tokens

API EndpointsOpenAI Compatible

`https://modelslab.com/api/v7/llm/chat/completions`Endpoint

Anthropic Compatible

`https://modelslab.com/api/v7/llm/v1/messages`Messages

`https://modelslab.com/api/v7/llm/v1/messages/count_tokens`Count Tokens

`https://modelslab.com/api/v7/llm/v1/models`Models

Use with Claude Code

cURL Example

ParametersSystem MessageYou are a helpful AI assistant specialized in providing accurate and detailed responses.

Temperature0.7

Max Tokens4000

Top P0.9

Frequency Penalty0

Presence Penalty0

Thinkinglow

lowmediumhigh

Deeper thinking spends more of the Max Tokens budget before the answer starts.

Model Info

Support

Looking for something different?
---

Models that trade off differently against the one you are viewing

### Better quality

Newer flagship models that generally produce stronger results.

[![GPT 5.2](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/2bdd00f4-1692-49d8-a562-72cb1753d0ff.png)](https://modelslab.com/models/openai/gpt-5.2)[OpenAI](https://modelslab.com/models/openai)

 [GPT 5.2](https://modelslab.com/models/openai/gpt-5.2)

[![GPT 5.1](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/fad83520-47a8-4059-b4b9-b32adf4d0be7.png)](https://modelslab.com/models/openai/gpt5.1)[OpenAI](https://modelslab.com/models/openai)

 [GPT 5.1](https://modelslab.com/models/openai/gpt5.1)

[![Gemini 3 Pro Preview](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/1846e76f-9d5c-4a35-a43a-82db45b7bb8b.png)](https://modelslab.com/models/google/gemini-3-pro-preview)[Google](https://modelslab.com/models/google)

 [Gemini 3 Pro Preview](https://modelslab.com/models/google/gemini-3-pro-preview)

[![GPT-5.1-Codex-Max](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/06356c4c-df09-4ccb-a5c0-629da670f75b.png)](https://modelslab.com/models/openai/gpt-5.1-codex-max)[OpenAI](https://modelslab.com/models/openai)

 [GPT-5.1-Codex-Max](https://modelslab.com/models/openai/gpt-5.1-codex-max)

### Lower cost per run

Open-source models we host, unlimited on the $149/month plan.

[Modelslab: Custom LLM

Popular](https://modelslab.com/models/modelslab/uncensored-chat)[ModelsLab](https://modelslab.com/models/modelslab)

 [Modelslab: Custom LLM

Open Source Model](https://modelslab.com/models/modelslab/uncensored-chat)

[![gpt-oss-20b](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/8b2f0440-52cc-4f41-93eb-d31006ad71e2.webp)](https://modelslab.com/models/openai/gpt-oss-20b)[OpenAI](https://modelslab.com/models/openai)

 [gpt-oss-20b

Closed Source Model](https://modelslab.com/models/openai/gpt-oss-20b)

[![DeepSeek: DeepSeek V4 Flash 0423](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/267a7228-f392-4d3e-abb6-915fef640c49.png)](https://modelslab.com/models/deepseek/deepseek-deepseek-v4-flash)[DeepSeek](https://modelslab.com/models/open_router)

 [DeepSeek: DeepSeek V4 Flash 0423

Closed Source Model](https://modelslab.com/models/deepseek/deepseek-deepseek-v4-flash)

[Qwen: Qwen3.5-9B](https://modelslab.com/models/qwen/qwen-qwen3.5-9b)[Qwen](https://modelslab.com/models/open_router)

 [Qwen: Qwen3.5-9B

Closed Source Model](https://modelslab.com/models/qwen/qwen-qwen3.5-9b)

[![Qwen3.5 9B FP8](https://images.modelslab.sh/?Image=https://assets.modelslab.ai/generations/475e7d51-3e69-4cbc-adb0-64e01535ac29.png)](https://modelslab.com/models/Qwen/Qwen-Qwen3.5-9B)[Qwen](https://modelslab.com/models/open_router)

 [Qwen3.5 9B FP8

Closed Source Model](https://modelslab.com/models/Qwen/Qwen-Qwen3.5-9B)

[Meta: Llama Guard 4 12B](https://modelslab.com/models/meta/meta-llama-llama-guard-4-12b)[Meta](https://modelslab.com/models/open_router)

 [Meta: Llama Guard 4 12B

Closed Source Model](https://modelslab.com/models/meta/meta-llama-llama-guard-4-12b)

Related Models
---

Discover similar models you might be interested in

 [View all LLM Models](https://modelslab.com/models?feature=llmaster)

[#### Google/Google: Gemma 3 27B

google-gemma-3-27b-it

From $0.27/M tokens](https://modelslab.com/models/google/google-gemma-3-27b-it)

[#### Qwen/Qwen: Qwen3.5-122B-A10B

qwen-qwen3.5-122b-a10b

From $1.17/M tokens](https://modelslab.com/models/qwen/qwen-qwen3.5-122b-a10b)

[#### NVIDIA/NVIDIA: Nemotron 3 Nano 30B A3B

nvidia-nemotron-3-nano-30b-a3b

From $0.13/M tokens](https://modelslab.com/models/nvidia/nvidia-nemotron-3-nano-30b-a3b)

[#### Cohere/Cohere: Command R7B (12-2024)

cohere-command-r7b-12-2024

From $0.09/M tokens](https://modelslab.com/models/cohere/cohere-command-r7b-12-2024)

[#### Meta/Llama 3.1 Nemotron 70B Instruct HF

nvidia-Llama-3.1-Nemotron-70B-Instruct-HF

From $0.88/M tokens](https://modelslab.com/models/meta/nvidia-Llama-3.1-Nemotron-70B-Instruct-HF)

[#### Mistral AI/Mistral Small (24B) Instruct 25.01

mistralai-Mistral-Small-24B-Instruct-2501

From $0.20/M tokens](https://modelslab.com/models/mistralai/mistralai-Mistral-Small-24B-Instruct-2501)

[#### OpenAI/OpenAI: GPT-3.5 Turbo Instruct

openai-gpt-3.5-turbo-instruct

From $1.75/M tokens](https://modelslab.com/models/openai/openai-gpt-3.5-turbo-instruct)

[#### DeepSeek/DeepSeek V3-0324

deepseek-ai-DeepSeek-V3

From $0.48/M tokens](https://modelslab.com/models/deepseek/deepseek-ai-DeepSeek-V3)

[#### Mistral AI/Mistral: Saba

mistralai-mistral-saba

From $0.40/M tokens](https://modelslab.com/models/mistralai/mistralai-mistral-saba)

[#### Morph/Morph: Morph V3 Large

morph-morph-v3-large

From $1.40/M tokens](https://modelslab.com/models/morph/morph-morph-v3-large)

[#### MiniMax/MiniMax: MiniMax M3

minimax-minimax-m3

From $0.75/M tokens](https://modelslab.com/models/minimax/minimax-minimax-m3)

[#### Google/Google: Gemini 2.5 Pro Preview 06-05

google-gemini-2.5-pro-preview

From $5.63/M tokens](https://modelslab.com/models/google/google-gemini-2.5-pro-preview)

[#### OpenAI/OpenAI: GPT-4o (2024-11-20)

openai-gpt-4o-2024-11-20

From $6.25/M tokens](https://modelslab.com/models/openai/openai-gpt-4o-2024-11-20)

[#### Inception/Inception: Mercury 2.5

inception-mercury-2.5

From $0.10/M tokens](https://modelslab.com/models/inception/inception-mercury-2.5)

[#### Anthropic/Anthropic: Claude Sonnet 4.5

anthropic-claude-sonnet-4.5

From $9.00/M tokens](https://modelslab.com/models/anthropic/anthropic-claude-sonnet-4.5)

[#### Qwen/Qwen: Qwen3 VL 235B A22B Thinking

qwen-qwen3-vl-235b-a22b-thinking

From $2.20/M tokens](https://modelslab.com/models/qwen/qwen-qwen3-vl-235b-a22b-thinking)

[#### Qwen/Qwen3 30B A3b

Qwen-Qwen3-30B-A3B

Free](https://modelslab.com/models/Qwen/Qwen-Qwen3-30B-A3B)

[#### Qwen/Qwen3-VL-8B-Instruct

Qwen-Qwen3-VL-8B-Instruct

From $0.43/M tokens](https://modelslab.com/models/Qwen/Qwen-Qwen3-VL-8B-Instruct)

About Qwen: Qwen3 8B
---

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

### Technical Specifications

Model IDqwen-qwen3-8b

CategoryLLM Models

TaskText Generation

Price$0.286 per million tokens

AddedFebruary 20, 2026

### Key Features

- Chat completion and multi-turn conversation API
- Streaming response with token-by-token output
- Function calling and tool use support
- System prompts and role-based messaging
- JSON mode and structured output

### Quick Start

Integrate Qwen: Qwen3 8B into your application with a single API call. Get your API key from the [pricing page](https://modelslab.com/pricing) to get started.

PythonJavaScriptcURLPHP

```
<code>import requests
import json

url = "https://modelslab.com/api/v7/llm/chat/completions"

headers = {
    "Content-Type": "application/json"
}

data = {
        "model_id": "qwen-qwen3-8b",
        "messages": [
            {
                "role": "user",
                "content": "Hello!"
            }
        ],
        "max_tokens": 1000,
        "key": "YOUR_API_KEY"
    }

try:
    response = requests.post(url, headers=headers, json=data)
    response.raise_for_status()  # Raises an HTTPError for bad responses (4XX or 5XX)
    result = response.json()
    print("API Response:")
    print(json.dumps(result, indent=2))
except requests.exceptions.HTTPError as http_err:
    print(f"HTTP error occurred: {http_err} - {response.text}")
except Exception as err:
    print(f"Other error occurred: {err}")</code>
```

### Pricing

Qwen: Qwen3 8B API costs $0.286 per million tokens. Pay only for what you use with no minimum commitments. [View pricing plans](https://modelslab.com/pricing)

### Use Cases

- AI chatbots and virtual assistants
- Code generation and developer tools
- Content writing and copywriting automation
- Data analysis, summarization, and extraction

[Browse LLM Models](https://modelslab.com/models?feature=llmaster) [More from Qwen](https://modelslab.com/models/open_router) [View Pricing](https://modelslab.com/pricing)

Qwen: Qwen3 8B FAQ
---

### What is Qwen: Qwen3 8B?

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

### How do I use the Qwen: Qwen3 8B API?

You can integrate Qwen: Qwen3 8B into your application with a single API call. Sign up on ModelsLab to get your API key, then use the model ID "qwen-qwen3-8b" in your API requests. We provide SDKs for Python, JavaScript, and cURL examples in the API documentation.

### How much does Qwen: Qwen3 8B cost?

Qwen: Qwen3 8B costs $0.286 per million tokens. ModelsLab plans start at $21/month (Basic), and the $149/month Open Source plan includes unlimited generation on open-source models.

### What is the Qwen: Qwen3 8B model ID?

The model ID for Qwen: Qwen3 8B is "qwen-qwen3-8b". Use this ID in your API requests to specify this model.

### Do I need a paid plan to use Qwen: Qwen3 8B?

Yes. ModelsLab is subscription-based with no free tier — plans start at $21/month (Basic, 3,250 API calls). The $149/month Open Source Unlimited plan includes unlimited generation across open-source models; premium third-party models are billed per call from your wallet.

---

*This markdown version is optimized for AI agents and LLMs.*

**Links:**
- [Website](https://modelslab.com)
- [API Documentation](https://docs.modelslab.com)
- [Blog](https://modelslab.com/blog)

---
*Generated by ModelsLab - 2026-09-11*