Developer Dashboard

Back to Model Explorer

DashscopeAvailable on AvalAI

qwen3-vl-flash API on AvalAI

qwen3-vl-flash

qwen3-vl-flash is an AI model from Alibaba that is available through AvalAI's OpenAI-compatible API.

Price display scale
Input price
Output price
Context window252,000 tokens
Maximum output32,000 tokens

Use qwen3-vl-flash with the AvalAI API

Keep the API key in an environment variable and send requests from server-side code.

bash
curl https://api.avalai.ir/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $AVALAI_API_KEY" \
  -d '{"model": "qwen3-vl-flash", "messages": [{"role": "user", "content": "Give a concise, practical solution."}]}'
python
import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["AVALAI_API_KEY"],
    base_url="https://api.avalai.ir/v1",
)

response = client.chat.completions.create(
    model="qwen3-vl-flash",
    messages=[{"role": "user", "content": "Give a concise, practical solution."}],
)

print(response.choices[0].message.content)
javascript
import OpenAI from "openai";

const client = new OpenAI({
  apiKey: process.env.AVALAI_API_KEY,
  baseURL: "https://api.avalai.ir/v1",
});

const response = await client.chat.completions.create({
  model: "qwen3-vl-flash",
  messages: [{ role: "user", content: "Give a concise, practical solution." }],
});

console.log(response.choices[0].message.content);

Create an API keyRead the quickstart

Capabilities and endpoints

  • Function calling
  • System messages
  • Tool choice
  • Vision

qwen3-vl-flash rate limits

TierRPMTPM
Basic140,000
Tier 125200,000
Tier 250400,000
Tier 375800,000
Tier 42501,000,000
Tier 57502,000,000

Frequently asked questions

What is the qwen3-vl-flash API on AvalAI?

Create an AvalAI API key and send the model ID qwen3-vl-flash to one of the supported endpoints shown on this page.

How much does the qwen3-vl-flash API cost on AvalAI?

The qwen3-vl-flash API on AvalAI costs — per 1M input tokens and — per 1M output tokens.

What is the token cost for qwen3-vl-flash?

Current AvalAI pricing for qwen3-vl-flash is — for input and — for output.

What is the qwen3-vl-flash context window?

The recorded maximum input context for qwen3-vl-flash is 252,000 tokens.

What are the rate limits for qwen3-vl-flash on AvalAI?

On tier 5, qwen3-vl-flash on AvalAI supports up to 750 requests per minute and 2,000,000 tokens per minute.

What features does qwen3-vl-flash support on AvalAI?

qwen3-vl-flash on AvalAI supports Function calling, System messages, Tool choice, Vision.

What account tier is required to use qwen3-vl-flash on AvalAI?

Calling qwen3-vl-flash on AvalAI requires the Basic tier (tier 0) or higher.

Third-party data and freshness

  • Pricing, availability, endpoints, and rate limits: AvalAI structured data (authoritative)