Qwen

Qwen: Qwen3.5-Flash

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

ReasoningTool callingVisionStreamingJSON mode
1M token context Released Feb 25, 2026

Intelligence overview

1M tok

Context window

65.5K tok

Max output

Capabilities

What Qwen: Qwen3.5-Flash can do.

Reasoning

Extended reasoning

Works through problems step by step before answering, trading latency for accuracy on harder tasks.

Multimodal

Image understanding

Reads and reasons over images alongside text in the same prompt.

Video input

Takes video as an input modality alongside text and images.

Development

Tool calling

Calls external functions mid-response, then continues reasoning from the result.

Structured output

Returns responses constrained to a JSON schema, ready to parse without cleanup.

Streaming

Streams tokens as they're generated instead of waiting on the full response.

Pricing

Pay for what you use.

Every model in UltraGPT runs behind one flat subscription — this is the underlying per-token cost the catalog tracks.

INPUT

$0.065 / 1M tok

OUTPUT

$0.26 / 1M tok

Use cases

Built for serious work.

Build

Generate and refactor code, with the ability to run and verify tool calls and return structured, parseable output.

Think

Works through multi-step problems with extended reasoning before it answers.

Research

Holds up to 1M tokens of source material in context without losing the thread.

Create

Reads images alongside a prompt to ground creative and design work in what's actually on screen.

Automate

Chains tool calls across multi-step workflows, reasoning between each one.

Preview

Ask Qwen: Qwen3.5-Flash anything.

A preview of the real UltraGPT composer — the model you see here is the model you get.

Qwen: Qwen3.5-Flash Ready

Continues in the UltraGPT app — one login, every model.

Comparison

How Qwen: Qwen3.5-Flash stacks up.

The closest models in the catalog by provider, capability and score — real entries, pulled from the same directory.

Specification

Technical details.

Specification

Context window
1M tokens
Max output
65.5K tokens
Input modalities
Text + Image + Video
Output modalities
Text
Release date
Feb 25, 2026
Streaming
Supported
Tool calling
Supported
JSON mode
Supported

Model identity

Provider
Qwen
Provider ID
qwen

Model ID

Directory

More from Qwen, and everything else.

This model is one entry in the full UltraGPT catalog — browse the rest, compare pricing and context windows, or filter by capability.

Qwen: Qwen3.5-Flash, one subscription away.

UltraGPT puts this model — and every other one in the directory — behind a single login and one flat monthly price.