qwen3-4b

apache-2.04B
Context
256K

Capabilities

vision
tool call
structured output
reasoning
json mode
streaming
fine tuning
batch

Details

Provider Qwen
Creator qwen
Familyqwen3
Licenseapache-2.0
Parameters4B
Status
Input modalitiestext
Output modalitiestext
Architecture
Knowledge cutoff
Training data cutoff
Release date
Deprecation date
Typechat
Reasoning tokensYes
Max input
Open weightYes
Sourceofficial
Last updated

Tools

Function Callingfunction_calling
Call external functions and APIs

Family Comparison: qwen3

ModelContextMax outPricing
qwen3-0.6b256K
qwen3-1.7b256K
qwen3-14b256K
qwen3-235b-a22b-instruct256K
qwen3-235b-a22b-thinking256K
qwen3-235b-a22b256K
qwen3-30b-a3b-instruct256K
qwen3-30b-a3b-thinking256K
qwen3-30b-a3b256K
qwen3-32b256K
qwen3-4b256K
qwen3-8b256K
qwen3-asr-flash-filetrans
qwen3-asr-flash-realtime
qwen3-asr-flash
qwen3-coder-30b-a3b-instruct256K
qwen3-coder-480b-a35b-instruct256K
qwen3-coder-flash1M
qwen3-coder-next256K66K
qwen3-coder-plus1M66K
qwen3-livetranslate-flash-realtime53K4K
qwen3-livetranslate-flash53K4K
qwen3-max-preview256K
qwen3-max256K33K
qwen3-next-80b-a3b-instruct256K
qwen3-next-80b-a3b-thinking256K
qwen3-omni-flash-0915
qwen3-omni-flash-realtime66K16K
qwen3-omni-flash66K33K
qwen3-omni-realtime-flash
qwen3-rerank
qwen3-tts-flash-realtime
qwen3-tts-flash
qwen3-tts-instruct-flash-realtime
qwen3-tts-instruct-flash
qwen3-vl-235b-a22b-instruct129K
qwen3-vl-30b-a3b-instruct129K
qwen3-vl-32b-instruct129K
qwen3-vl-8b-instruct129K
qwen3-vl-embedding
qwen3-vl-flash82K
qwen3-vl-plus262K82K
qwen3-vl-rerank
qwen3
qwen3-vl-32b-thinking131K82K$0.16$0.64
qwen3-vl-8b-thinking127K$0.18$2.1
qwen3-vl-30b-a3b-thinking127K$0.2$2.4
qwen3-vl-235b-a22b-thinking127K$0.4$4
qwen3-omni-30b-a3b-captioner66K33K$3.81$3.06

API

GET/v1/models/qwen/qwen3-4b