Bright

BRIGHT EVIDENCE PACK / Deployed

Choosing when a model thinks longer.

The Qwen team released and described Qwen3, a family of dense and mixture-of-experts language models with selectable thinking and non-thinking modes.

Canonical Bright record · JSON evidence pack · Key-facts embed

Dates and assessment

Source published
2025-04-29
Bright published
2026-09-07
Substantive update
None recorded
Evidence state
Deployed
Independent verification
Not established by this source review
Last source review
2026-09-07

The claim in context

The human problem

People building AI systems often need to balance response speed and cost against extra time for complex reasoning, including in multilingual work.

The prior constraint

A system's reasoning budget is often fixed or hidden from the person using it.

AI’s actual role

Qwen3 can use a slower step-by-step Thinking Mode or a faster Non-Thinking Mode; the team also reports support for 119 languages and dialects.

The documented result

Qwen announced models from 0.6B to 32B dense sizes and 30B-A3B/235B-A22B mixture-of-experts sizes, with pre- and post-trained variants available via Hugging Face, ModelScope, and Kaggle. Its source describes the two operating modes and reports multilingual support.

Why it may matter

Giving developers a stated control over reasoning time could make it easier to reserve slower processing for harder tasks and use faster responses for straightforward work.

Limitations

Original evidence

Attribution

Credit Bright AI Future and link the canonical Bright record.

Linked source material, quotations, trademarks and media remain subject to their owners’ terms. No reuse right is granted for third-party media.