Qwen3 Coder Flash API on AvalAI
dashscope/qwen3-coder-flashQwen3 Coder Flash is an AI model from Alibaba that is available through AvalAI's OpenAI-compatible API.
Use Qwen3 Coder Flash with the AvalAI API
Keep the API key in an environment variable and send requests from server-side code.
curl https://api.avalai.ir/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $AVALAI_API_KEY" \
-d '{"model": "dashscope/qwen3-coder-flash", "messages": [{"role": "user", "content": "Give a concise, practical solution."}]}'import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["AVALAI_API_KEY"],
base_url="https://api.avalai.ir/v1",
)
response = client.chat.completions.create(
model="dashscope/qwen3-coder-flash",
messages=[{"role": "user", "content": "Give a concise, practical solution."}],
)
print(response.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
apiKey: process.env.AVALAI_API_KEY,
baseURL: "https://api.avalai.ir/v1",
});
const response = await client.chat.completions.create({
model: "dashscope/qwen3-coder-flash",
messages: [{ role: "user", content: "Give a concise, practical solution." }],
});
console.log(response.choices[0].message.content);Create an API keyRead the quickstart
Capabilities and endpoints
- Function calling
- Reasoning
- Tool choice
Third-party model parameters
Third-party reference data does not change AvalAI endpoint support; use the capabilities and endpoints above.
frequency_penaltylogprobsmax_tokenspresence_penaltyresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_logprobstop_p
Qwen3 Coder Flash rate limits
| Tier | RPM | TPM |
|---|---|---|
| Basic | 1 | 40,000 |
| Tier 1 | 25 | 200,000 |
| Tier 2 | 50 | 400,000 |
| Tier 3 | 75 | 800,000 |
| Tier 4 | 250 | 1,000,000 |
| Tier 5 | 750 | 2,000,000 |
Third-party model data
Source: OpenRouterIndependent sources supplement this context; AvalAI remains authoritative for pricing, availability, and endpoint support.
Reviewed snapshot 2026-07-31 · This snapshot was checked more than 14 days ago.
- Provider count
- —
- Context window range
- 1,000,000
- Maximum output range
- 65,536
- OpenRouter model input rate range
- $0.195 / 1M
- OpenRouter model output rate range
- $0.975 / 1M
- OpenRouter fee
- 5.5% credit purchase fee · OpenRouter fee details
- Effective input price range after fee
- $0.2057 / 1M
- Effective output price range after fee
- $1.0286 / 1M
Related models
- qwen3-coder-flash-2025-07-28
- qwen3-vl-plus-2025-12-19
- qwen3-vl-flash-2026-01-22
- qwen3-vl-flash-2025-10-15
Frequently asked questions
What is the Qwen3 Coder Flash API on AvalAI?
Create an AvalAI API key and send the model ID dashscope/qwen3-coder-flash to one of the supported endpoints shown on this page.
How much does the Qwen3 Coder Flash API cost on AvalAI?
The Qwen3 Coder Flash API on AvalAI costs — per 1M input tokens and — per 1M output tokens.
What is the token cost for Qwen3 Coder Flash?
Current AvalAI pricing for Qwen3 Coder Flash is — for input and — for output.
What is the Qwen3 Coder Flash context window?
The recorded maximum input context for Qwen3 Coder Flash is 997,952 tokens.
What are the rate limits for Qwen3 Coder Flash on AvalAI?
On tier 5, Qwen3 Coder Flash on AvalAI supports up to 750 requests per minute and 2,000,000 tokens per minute.
What features does Qwen3 Coder Flash support on AvalAI?
Qwen3 Coder Flash on AvalAI supports Function calling, Reasoning, Tool choice.
What account tier is required to use Qwen3 Coder Flash on AvalAI?
Calling Qwen3 Coder Flash on AvalAI requires the Basic tier (tier 0) or higher.
Third-party data and freshness
- Pricing, availability, endpoints, and rate limits: AvalAI structured data (authoritative)
- OpenRouter marketplace · Last checked: 2026-07-31