Description
Qwen3.8-Flash-Next is an experimental preview of the architecture that will underpin Qwen4, built around a fundamental rethinking of how the core components of modern large language models interact at scale.
Stats
103 Downloads
Capabilities
Minimum system memory
Tags
Description
Qwen3.8-Flash-Next is an experimental preview of the architecture that will underpin Qwen4, built around a fundamental rethinking of how the core components of modern large language models interact at scale.
Stats
103 Downloads
Capabilities
Minimum system memory
Tags
Last updated
Updated on August 28byREADME
Qwen3.8-Flash-Next is an experimental preview of the architecture that will underpin Qwen4, built around a fundamental rethinking of how the core components of modern large language models interact at scale.
xhigh, medium, and low reasoning-effort levels. Reasoning from previous messages is preserved by default for continuity in multi-turn agentic workflows.Custom Fields
Special features defined by the model author
Reasoning Effort
: select
(default=xhigh)
Controls how much reasoning the model should perform.
Enable Thinking
: boolean
(default=true)
Controls whether the model will think before replying.
Preserve Thinking
: boolean
(default=true)
Preserves reasoning content in all prior assistant turns instead of only the most recent one.
Add Vision IDs
: boolean
(default=false)
Labels images with sequential identifiers in the prompt.
Parameters
Custom configuration options included with this model
Sources
The underlying model files this model uses
Based on
Description
Qwen3.8-Flash-Next is an experimental preview of the architecture that will underpin Qwen4, built around a fundamental rethinking of how the core components of modern large language models interact at scale.
Stats
103 Downloads
Capabilities
Minimum system memory
Tags
Description
Qwen3.8-Flash-Next is an experimental preview of the architecture that will underpin Qwen4, built around a fundamental rethinking of how the core components of modern large language models interact at scale.
Stats
103 Downloads
Capabilities
Minimum system memory
Tags
Last updated
Updated on August 28byREADME
Qwen3.8-Flash-Next is an experimental preview of the architecture that will underpin Qwen4, built around a fundamental rethinking of how the core components of modern large language models interact at scale.
xhigh, medium, and low reasoning-effort levels. Reasoning from previous messages is preserved by default for continuity in multi-turn agentic workflows.Custom Fields
Special features defined by the model author
Reasoning Effort
: select
(default=xhigh)
Controls how much reasoning the model should perform.
Enable Thinking
: boolean
(default=true)
Controls whether the model will think before replying.
Preserve Thinking
: boolean
(default=true)
Preserves reasoning content in all prior assistant turns instead of only the most recent one.
Add Vision IDs
: boolean
(default=false)
Labels images with sequential identifiers in the prompt.
Parameters
Custom configuration options included with this model
Sources
The underlying model files this model uses
Based on