gpt-oss-120b API on AvalAI
gpt-oss-120bgpt-oss-120b is an AI model from OpenAI that is available through AvalAI's OpenAI-compatible API.
Use gpt-oss-120b with the AvalAI API
Keep the API key in an environment variable and send requests from server-side code.
curl https://api.avalai.ir/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $AVALAI_API_KEY" \
-d '{"model": "gpt-oss-120b", "messages": [{"role": "user", "content": "Give a concise, practical solution."}]}'import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["AVALAI_API_KEY"],
base_url="https://api.avalai.ir/v1",
)
response = client.chat.completions.create(
model="gpt-oss-120b",
messages=[{"role": "user", "content": "Give a concise, practical solution."}],
)
print(response.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
apiKey: process.env.AVALAI_API_KEY,
baseURL: "https://api.avalai.ir/v1",
});
const response = await client.chat.completions.create({
model: "gpt-oss-120b",
messages: [{ role: "user", content: "Give a concise, practical solution." }],
});
console.log(response.choices[0].message.content);Create an API keyRead the quickstart
Capabilities and endpoints
- Function calling
- Parallel function calling
- Reasoning
- Structured output
- Tool choice
- Vision
- Web search
Third-party model parameters
Third-party reference data does not change AvalAI endpoint support; use the capabilities and endpoints above.
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_atop_ktop_logprobstop_p
Supplemental reasoning effort values
high, low, medium
gpt-oss-120b rate limits
| Tier | RPM | TPM |
|---|---|---|
| Basic | 1 | 40,000 |
| Tier 1 | 250 | 450,000 |
| Tier 2 | 500 | 800,000 |
| Tier 3 | 750 | 1,200,000 |
| Tier 4 | 1,500 | 2,000,000 |
| Tier 5 | 2,500 | 4,000,000 |
AvalAI and reference marketplace pricing
| Token role | AvalAI final price0% platform fee | OpenRouter cost breakdown | ||
|---|---|---|---|---|
| Model rate | Credit purchase fee | Effective OpenRouter cost after fee | ||
| Input price | $0.30 / 1M tokens | $0.03 / 1M tokens | 5.5% | $0.03 / 1M tokens |
| Output price | $2.50 / 1M tokens | $0.17 / 1M tokens | 5.5% | $0.18 / 1M tokens |
AvalAI pricing is final with a 0% platform fee. The effective OpenRouter prices add its 5.5% credit purchase fee to the highest reported model rate, including long-context tiers. OpenRouter charges a 5.5% fee when credits are purchased ($0.80 minimum), so small credit purchases can have a higher effective cost. OpenRouter FAQ
Peer benchmark performance
Independent, sourced scores with no composite rating.
Artificial Analysis Agentic Index
| Score | Peer rank | Direction | Source and date |
|---|---|---|---|
| 13.4 index | No comparable data | Higher is better | OpenRouter · 2026-08-17 |
Artificial Analysis Coding Index
| Score | Peer rank | Direction | Source and date |
|---|---|---|---|
| 30.4 index | No comparable data | Higher is better | OpenRouter · 2026-08-17 |
Artificial Analysis Intelligence Index
| Score | Peer rank | Direction | Source and date |
|---|---|---|---|
| 24.1 index | No comparable data | Higher is better | OpenRouter · 2026-08-17 |
Third-party model data
Source: OpenRouterIndependent sources supplement this context; AvalAI remains authoritative for pricing, availability, and endpoint support.
Reviewed snapshot 2026-08-17
- Provider count
- —
- Context window range
- 131,072
- Maximum output range
- 131,072
- OpenRouter model input rate range
- $0.03 / 1M
- OpenRouter model output rate range
- $0.17 / 1M
- OpenRouter fee
- 5.5% credit purchase fee · OpenRouter fee details
- Effective input price range after fee
- $0.0317 / 1M
- Effective output price range after fee
- $0.1793 / 1M
Related models
Frequently asked questions
What is the gpt-oss-120b API on AvalAI?
Create an AvalAI API key and send the model ID gpt-oss-120b to one of the supported endpoints shown on this page.
How much does the gpt-oss-120b API cost on AvalAI?
The gpt-oss-120b API on AvalAI costs $0.30 per 1M input tokens and $2.50 per 1M output tokens.
What is the token cost for gpt-oss-120b?
Current AvalAI pricing for gpt-oss-120b is $0.30 / 1M tokens for input and $2.50 / 1M tokens for output.
What is the gpt-oss-120b context window?
The recorded maximum input context for gpt-oss-120b is 128,000 tokens.
What are the rate limits for gpt-oss-120b on AvalAI?
On tier 5, gpt-oss-120b on AvalAI supports up to 2,500 requests per minute and 4,000,000 tokens per minute.
What features does gpt-oss-120b support on AvalAI?
gpt-oss-120b on AvalAI supports Function calling, Parallel function calling, Reasoning, Structured output, Tool choice, Vision, Web search.
What account tier is required to use gpt-oss-120b on AvalAI?
Calling gpt-oss-120b on AvalAI requires the Basic tier (tier 0) or higher.
Third-party data and freshness
- Pricing, availability, endpoints, and rate limits: AvalAI structured data (authoritative)
- OpenRouter marketplace · Last checked: 2026-08-17