Skip to main content

Qwen: Qwen3 8B Coding Benchmark

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

Context131,072tokens
Max Output8,192tokens
Inputmodality
Price$0.12/1M input

Try Qwen: Qwen3 8B in Kilo Code

Experience this model with the most popular open source coding agent. Free to start, pay only for AI usage. Use in popular IDEs like VS Code, JetBrains, command line, or cloud agents.

5M+

Downloads

500+

models supported

Free

to Start

Access 500+ models including Qwen: Qwen3 8B and many more in Kilo Code

Coding Performance

Coding benchmarks and performance metrics for development tasks

Security

Enkrypt AI red-team scores for Qwen: Qwen3 8B. Each value is the share of successful attacks on a 0–100 scale — lower is safer.

Overall risk
39.5

Lower is safer

Safety

42.9/ 100

Composite safety risk from Enkrypt red-team evaluations. Lower is safer.

NIST

39.0/ 100

Average attack success across NIST-mapped tests: bias, harm, toxicity, CBRN, and insecure code.

OWASP

44.0/ 100

Weighted average of the same tests using OWASP Top 10 for LLMs 2025 risk rankings.

Attack categories

Percentage of successful attacks in each Enkrypt red-team category.

Jailbreak25.9/ 100

Share of jailbreak tests that bypassed the model's safety constraints.

Bias97.7/ 100

Share of tests that elicited biased responses.

Harmful content37.8/ 100

Share of tests that produced dangerous, violent, or hateful content.

Toxicity1.8/ 100

Share of tests that produced toxic or abusive content.

CBRN33.0/ 100

Share of tests that elicited chemical, biological, radiological, or nuclear assistance.

Insecure code27.1/ 100

Share of tests that produced vulnerable or malicious code.

Score
Value
Overall risk
39.5 / 100
Safety
42.9 / 100
NIST
39.0 / 100
OWASP
44.0 / 100
Jailbreak
25.9 / 100
Bias
97.7 / 100
Harmful content
37.8 / 100
Toxicity
1.8 / 100
CBRN
33.0 / 100
Insecure code
27.1 / 100

Security scores from the Enkrypt AI Safety Leaderboard · Last checked Sep 17, 2026

Real-World Usage

Real-world usage statistics from the Kilo Code community

Weekly Token Usage

No ranking data available for this model yet.

Real-world metrics from the Kilo Code Leaderboard

Pricing

Cost per 1 million tokens

Input Tokens
$0.12
per 1M tokens
Output Tokens
$0.46
per 1M tokens

Example Cost

Analyzing a 10,000 line codebase (≈40k input tokens, 10k output tokens) costs approximately $0.0092

Coding Capabilities

Features and parameters relevant to coding tasks

Coding Features

Function Calling
Can call external functions/APIs
Tool Choice
Control over function selection
Structured Outputs
JSON schema validation
Reasoning Tokens
Extended thinking for complex problems

Pricing details from OpenRouter

Technical Details

Architecture and implementation specifications

Specifications
Model ID
qwen/qwen3-8b
Created
April 28, 2025
Tokenizer
Qwen3
Input Modalities
Text
Context Window
131,072 tokens
Max Completion Tokens
8,192 tokens
Input Price
$0.12 per 1M tokens
Output Price
$0.46 per 1M tokens
Content Moderation
Disabled

Ready to try Qwen: Qwen3 8B?

Install Kilo Code and start using Qwen: Qwen3 8B for your coding projects today. Choose from 500+ AI models with complete freedom.

  1. Install Kilo Code

    Get the extension from VS Code Marketplace, JetBrains Plugin Repository, or the CLI.

  2. Open the model selector

    Click the model name in the Kilo Code chat panel to open the selector.

  3. Choose your model

    Search or browse to find and select your preferred model.

  4. Start coding

    Use Code, Ask, Debug, or Plan mode — the model is ready immediately.