Qwen: Qwen3 30B A3B Instruct 2507
Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...
Intelligence overview
262.1K tok
Context window
32K tok
Max output
What Qwen: Qwen3 30B A3B Instruct 2507 can do.
Development
Tool calling
Calls external functions mid-response, then continues reasoning from the result.
Structured output
Returns responses constrained to a JSON schema, ready to parse without cleanup.
Streaming
Streams tokens as they're generated instead of waiting on the full response.
Pay for what you use.
Every model in UltraGPT runs behind one flat subscription — this is the underlying per-token cost the catalog tracks.
$0.048 / 1M tok
$0.193 / 1M tok
Built for serious work.
Build
Generate and refactor code, with the ability to run and verify tool calls and return structured, parseable output.
Research
Holds up to 262.1K tokens of source material in context without losing the thread.
Ask Qwen: Qwen3 30B A3B Instruct 2507 anything.
A preview of the real UltraGPT composer — the model you see here is the model you get.
Continues in the UltraGPT app — one login, every model.
How Qwen: Qwen3 30B A3B Instruct 2507 stacks up.
The closest models in the catalog by provider, capability and score — real entries, pulled from the same directory.
Technical details.
Specification
- Context window
- 262.1K tokens
- Max output
- 32K tokens
- Input modalities
- Text
- Output modalities
- Text
- Knowledge cutoff
- Jun 30, 2025
- Release date
- Jul 29, 2025
- Streaming
- Supported
- Tool calling
- Supported
- JSON mode
- Supported
More from Qwen, and everything else.
This model is one entry in the full UltraGPT catalog — browse the rest, compare pricing and context windows, or filter by capability.
