The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...
Tokens processed for this model through Reroute, per day, over the last 30 days.
Daily token volume shows up here as soon as the first request for this model completes.