Models / Qwen3 8B

Qwen3 8B

Tier 1Apache-2.0

Lightweight dense model for local deployment and edge inference. Supports thinking mode.

qwen/qwen3-8b
Context Window
131K
Max Output
8K
Providers
1
Released
2025-04

Capabilities

ChatCodeThinkingToolsStreaming

Pricing by Provider

Prices shown are the official rate published by the model provider, before any platform discounts. Your actual rate may be lower based on your account's discount tier.

ProviderPriceLatency p50Latency p95Status
alibaba$0.18 /M in · $0.70 /M out180ms450ms

Quick Start

Python
import magicrouter

mr = magicrouter.Client(
    provider_keys={"alibaba": "your-api-key"}
)

response = mr.chat(
    "qwen/qwen3-8b",
    "Your prompt here"
)
print(response.choices[0].message.content)
TypeScript
import { MagicRouter } from "magicrouter";

const mr = new MagicRouter({
  providerKeys: { alibaba: "your-api-key" }
});

const response = await mr.chat({
  model: "qwen/qwen3-8b",
  messages: [{ role: "user", content: "Your prompt here" }]
});
console.log(response.choices[0].message.content);
cURL
curl https://dashscope-intl.aliyuncs.com/compatible-mode/v1/chat/completions \
  -H "Authorization: Bearer your-api-key" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen3-8b",
    "messages": [{"role": "user", "content": "Your prompt here"}]
  }'

Use this model

Sign up for free and test Qwen3 8B in the playground

Get Started