Qwen 3.5/3.6
Qwen 3.5/3.6
Qwen is Mixlayer’s general-purpose model family. It spans small, low-latency models through large mixture-of-experts models for difficult reasoning and agentic work.
Model details
- Context window: 131K tokens
- Capabilities: Text · Vision · Reasoning · Tools
Available models
Thinking modes
Qwen supports a thinking mode, with reasoning returned in reasoning_content, and a non-thinking mode for faster direct answers. See Reasoning for request examples.
Mixlayer platform defaults
When these parameters are omitted from a request, Mixlayer applies the following defaults:
Recommended sampling
Qwen also recommends min_p=0.0. Mixlayer does not currently expose
min_p, so omit it and use the gateway default.
For the mixture-of-experts variants, active parameters primarily govern inference cost and latency while total parameters contribute to capability.