Developer Dashboard

Back to Model Explorer

DashscopeAvailable on AvalAI

Qwen3 VL 32B Instruct API on AvalAI

qwen3-vl-32b-instruct

Qwen3 VL 32B Instruct is an AI model from Alibaba that is available through AvalAI's OpenAI-compatible API.

Price display scale
Input price$1.60 / 1M tokens
Output price$0.64 / 1M tokens
Context window126,000 tokens
Maximum output32,000 tokens
Release date2025-10-23

Use Qwen3 VL 32B Instruct with the AvalAI API

Keep the API key in an environment variable and send requests from server-side code.

bash
curl https://api.avalai.ir/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $AVALAI_API_KEY" \
  -d '{"model": "qwen3-vl-32b-instruct", "messages": [{"role": "user", "content": "Give a concise, practical solution."}]}'
python
import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["AVALAI_API_KEY"],
    base_url="https://api.avalai.ir/v1",
)

response = client.chat.completions.create(
    model="qwen3-vl-32b-instruct",
    messages=[{"role": "user", "content": "Give a concise, practical solution."}],
)

print(response.choices[0].message.content)
javascript
import OpenAI from "openai";

const client = new OpenAI({
  apiKey: process.env.AVALAI_API_KEY,
  baseURL: "https://api.avalai.ir/v1",
});

const response = await client.chat.completions.create({
  model: "qwen3-vl-32b-instruct",
  messages: [{ role: "user", content: "Give a concise, practical solution." }],
});

console.log(response.choices[0].message.content);

Create an API keyRead the quickstart

Capabilities and endpoints

  • Function calling
  • System messages
  • Tool choice
  • Vision

Third-party model parameters

Third-party reference data does not change AvalAI endpoint support; use the capabilities and endpoints above.

  • frequency_penalty
  • logprobs
  • max_tokens
  • presence_penalty
  • response_format
  • seed
  • stop
  • structured_outputs
  • temperature
  • tool_choice
  • tools
  • top_k
  • top_logprobs
  • top_p

Qwen3 VL 32B Instruct rate limits

TierRPMTPM
Basic140,000
Tier 110200,000
Tier 2251,000,000
Tier 350800,000
Tier 4751,200,000
Tier 55002,000,000

AvalAI and reference marketplace pricing

Token roleAvalAI final price0% platform feeOpenRouter cost breakdown
Model rateCredit purchase feeEffective OpenRouter cost after fee
Input price$1.60 / 1M tokens$0.10 / 1M tokens5.5%$0.11 / 1M tokens
Output price$0.64 / 1M tokens$0.42 / 1M tokens5.5%$0.44 / 1M tokens

AvalAI pricing is final with a 0% platform fee. The effective OpenRouter prices add its 5.5% credit purchase fee to the highest reported model rate, including long-context tiers. OpenRouter charges a 5.5% fee when credits are purchased ($0.80 minimum), so small credit purchases can have a higher effective cost. OpenRouter FAQ

Third-party model data

Source: OpenRouter

Independent sources supplement this context; AvalAI remains authoritative for pricing, availability, and endpoint support.

Reviewed snapshot 2026-07-31 · This snapshot was checked more than 14 days ago.

Provider count
Context window range
131,072
Maximum output range
32,768
OpenRouter model input rate range
$0.104 / 1M
OpenRouter model output rate range
$0.416 / 1M
OpenRouter fee
5.5% credit purchase fee · OpenRouter fee details
Effective input price range after fee
$0.1097 / 1M
Effective output price range after fee
$0.4389 / 1M

Frequently asked questions

What is the Qwen3 VL 32B Instruct API on AvalAI?

Create an AvalAI API key and send the model ID qwen3-vl-32b-instruct to one of the supported endpoints shown on this page.

How much does the Qwen3 VL 32B Instruct API cost on AvalAI?

The Qwen3 VL 32B Instruct API on AvalAI costs $1.60 per 1M input tokens and $0.64 per 1M output tokens.

What is the token cost for Qwen3 VL 32B Instruct?

Current AvalAI pricing for Qwen3 VL 32B Instruct is $1.60 / 1M tokens for input and $0.64 / 1M tokens for output.

What is the Qwen3 VL 32B Instruct context window?

The recorded maximum input context for Qwen3 VL 32B Instruct is 126,000 tokens.

What are the rate limits for Qwen3 VL 32B Instruct on AvalAI?

On tier 5, Qwen3 VL 32B Instruct on AvalAI supports up to 500 requests per minute and 2,000,000 tokens per minute.

What features does Qwen3 VL 32B Instruct support on AvalAI?

Qwen3 VL 32B Instruct on AvalAI supports Function calling, System messages, Tool choice, Vision.

What account tier is required to use Qwen3 VL 32B Instruct on AvalAI?

Calling Qwen3 VL 32B Instruct on AvalAI requires the Basic tier (tier 0) or higher.

Third-party data and freshness

  • Pricing, availability, endpoints, and rate limits: AvalAI structured data (authoritative)
  • OpenRouter marketplace · Last checked: 2026-07-31