Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and agentic tool use. Post-trained on instruction data, it demonstrates competitive performance across reasoning (AIME, ZebraLogic), coding (MultiPL-E, LiveCodeBench), and alignment (IFEval, WritingBench) benchmarks. It outperforms its non-instruct variant on subjective and open-ended tasks while retaining strong factual and coding performance.

Key specifications

Capability
44.4
Context window
131.1K
Max output
65.5K
Input $/1M
$0.20
Output $/1M
$0.80

Inputs and outputs

Recorded modality support shows what this model can accept and produce. An undocumented modality is shown as unknown rather than assumed to be unsupported.

Accepts

  • Text input
  • Image input

Produces

  • Text output

Capabilities

Function callingVisionStreamingJSON mode

What is Qwen3 30b A3b Instruct 2507?

Catalogue record

Qwen3 30b A3b Instruct 2507 is listed in Requesty’s public model catalogue under the exact routing identifier alibaba/qwen3-30b-a3b-instruct-2507. Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and agentic tool use. Post-trained on instruction data, it demonstrates competitive performance across reasoning (AIME, ZebraLogic), coding (MultiPL-E, LiveCodeBench), and alignment (IFEval, WritingBench) benchmarks. It outperforms its non-instruct variant on subjective and open-ended tasks while retaining strong factual and coding performance.

Recorded interface details

At the time this draft was collected, the catalogue records a context window of 131,072 tokens and a maximum output of 65,536 tokens. It records no reasoning-support flag, vision input support, tool-calling support, and no image-generation flag. These fields describe the routing catalogue entry and should be rechecked against the current provider documentation before production use.

Pricing and availability

The source catalogue reports input and output prices when available. Pricing can depend on the route, region, token tier, cache use, account, and date, so this record does not infer a commercial commitment from a single snapshot. Confirm current pricing, data handling, regional controls, rate limits, and availability with the current upstream documentation.

Editorial status

This is a source-backed draft imported from Requesty’s public catalogue. It is not published automatically. An editor must verify the exact model identifier, the material claims, and the chosen deployment path before approving it for public discovery or sitemap inclusion.

How to evaluate Qwen3 30b A3b Instruct 2507

Use the recorded facts as a shortlist, then validate the model against representative inputs and production constraints.

Workload fit

Qwen3 30b A3b Instruct 2507 is categorized for Image-text-to-text, Text-to-text AI models. Its current record accepts text, image and produces text. Confirm file formats, preprocessing, and provider-specific request schemas before implementation.

Capacity and cost

The directory records 131.1K context and 65.5K maximum output. Listed token prices are $0.20 input and $0.80 output per 1M tokens. Treat missing values as unknown and recheck current commercial terms.

Operational behavior

Recorded output speed is — and time to first token is —. Hosting route, region, prompt length, concurrency, and provider load can materially change both measurements.

Evidence boundary

This page separates sourced model facts from incomplete fields. Benchmark evidence is displayed only when its source is linked and verified. Before choosing Qwen3 30b A3b Instruct 2507, test task quality, tool reliability, safety behavior, data controls, rate limits, and total cost on the exact route you intend to use.

Strengths

  • + Requesty catalogue source recorded

Limitations

  • Editorial approval required before publication

Frequently asked questions

Qwen3 30b A3b Instruct 2507 is recorded under the Requesty routing identifier alibaba/qwen3-30b-a3b-instruct-2507. This source-backed directory entry remains a draft until editorial approval.

Related approved models in Text-to-text AI models. Compare specifications and verify fit for your workload.

Anthropic

Claude 3.5 Sonnet

The 2024 Sonnet that set the standard for practical coding work — now well behind the current Claude line at identical pricing.

200K contextCompare

OpenAI

GPT-4o

OpenAI's omni-era workhorse — text and image in, text out, with a cached-input rate that halves the cost of repeated context.

128K contextCompare

Meta

Llama 3.1 405B

Meta's largest open-weights model, competitive with frontier closed models.

128K contextCompare

DeepSeek

DeepSeek V4 Flash 0731

Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding

1M contextCompare