Developer Dashboard

Back to Model Explorer

AzureAvailable on AvalAI

gpt-oss-120b API on AvalAI

gpt-oss-120b

gpt-oss-120b is an AI model from OpenAI that is available through AvalAI's OpenAI-compatible API.

Price display scale
Input price$0.30 / 1M tokens
Output price$2.50 / 1M tokens
Context window128,000 tokens
Maximum output128,000 tokens
Release date2025-08-05
Knowledge cutoff2024-06-30

Use gpt-oss-120b with the AvalAI API

Keep the API key in an environment variable and send requests from server-side code.

bash
curl https://api.avalai.ir/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $AVALAI_API_KEY" \
  -d '{"model": "gpt-oss-120b", "messages": [{"role": "user", "content": "Give a concise, practical solution."}]}'
python
import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["AVALAI_API_KEY"],
    base_url="https://api.avalai.ir/v1",
)

response = client.chat.completions.create(
    model="gpt-oss-120b",
    messages=[{"role": "user", "content": "Give a concise, practical solution."}],
)

print(response.choices[0].message.content)
javascript
import OpenAI from "openai";

const client = new OpenAI({
  apiKey: process.env.AVALAI_API_KEY,
  baseURL: "https://api.avalai.ir/v1",
});

const response = await client.chat.completions.create({
  model: "gpt-oss-120b",
  messages: [{ role: "user", content: "Give a concise, practical solution." }],
});

console.log(response.choices[0].message.content);

Create an API keyRead the quickstart

Capabilities and endpoints

  • Function calling
  • Parallel function calling
  • Reasoning
  • Structured output
  • Tool choice
  • Vision
  • Web search

Third-party model parameters

Third-party reference data does not change AvalAI endpoint support; use the capabilities and endpoints above.

  • frequency_penalty
  • include_reasoning
  • logit_bias
  • logprobs
  • max_tokens
  • min_p
  • presence_penalty
  • reasoning
  • reasoning_effort
  • repetition_penalty
  • response_format
  • seed
  • stop
  • structured_outputs
  • temperature
  • tool_choice
  • tools
  • top_a
  • top_k
  • top_logprobs
  • top_p

Supplemental reasoning effort values

high, low, medium

gpt-oss-120b rate limits

TierRPMTPM
Basic140,000
Tier 1250450,000
Tier 2500800,000
Tier 37501,200,000
Tier 41,5002,000,000
Tier 52,5004,000,000

AvalAI and reference marketplace pricing

Token roleAvalAI final price0% platform feeOpenRouter cost breakdown
Model rateCredit purchase feeEffective OpenRouter cost after fee
Input price$0.30 / 1M tokens$0.03 / 1M tokens5.5%$0.03 / 1M tokens
Output price$2.50 / 1M tokens$0.17 / 1M tokens5.5%$0.18 / 1M tokens

AvalAI pricing is final with a 0% platform fee. The effective OpenRouter prices add its 5.5% credit purchase fee to the highest reported model rate, including long-context tiers. OpenRouter charges a 5.5% fee when credits are purchased ($0.80 minimum), so small credit purchases can have a higher effective cost. OpenRouter FAQ

Peer benchmark performance

Independent, sourced scores with no composite rating.

Artificial Analysis Agentic Index

13.4index
ScorePeer rankDirectionSource and date
13.4 indexNo comparable dataHigher is betterOpenRouter · 2026-08-17

Artificial Analysis Coding Index

30.4index
ScorePeer rankDirectionSource and date
30.4 indexNo comparable dataHigher is betterOpenRouter · 2026-08-17

Artificial Analysis Intelligence Index

24.1index
ScorePeer rankDirectionSource and date
24.1 indexNo comparable dataHigher is betterOpenRouter · 2026-08-17

Third-party model data

Source: OpenRouter

Independent sources supplement this context; AvalAI remains authoritative for pricing, availability, and endpoint support.

Reviewed snapshot 2026-08-17

Provider count
Context window range
131,072
Maximum output range
131,072
OpenRouter model input rate range
$0.03 / 1M
OpenRouter model output rate range
$0.17 / 1M
OpenRouter fee
5.5% credit purchase fee · OpenRouter fee details
Effective input price range after fee
$0.0317 / 1M
Effective output price range after fee
$0.1793 / 1M

Frequently asked questions

What is the gpt-oss-120b API on AvalAI?

Create an AvalAI API key and send the model ID gpt-oss-120b to one of the supported endpoints shown on this page.

How much does the gpt-oss-120b API cost on AvalAI?

The gpt-oss-120b API on AvalAI costs $0.30 per 1M input tokens and $2.50 per 1M output tokens.

What is the token cost for gpt-oss-120b?

Current AvalAI pricing for gpt-oss-120b is $0.30 / 1M tokens for input and $2.50 / 1M tokens for output.

What is the gpt-oss-120b context window?

The recorded maximum input context for gpt-oss-120b is 128,000 tokens.

What are the rate limits for gpt-oss-120b on AvalAI?

On tier 5, gpt-oss-120b on AvalAI supports up to 2,500 requests per minute and 4,000,000 tokens per minute.

What features does gpt-oss-120b support on AvalAI?

gpt-oss-120b on AvalAI supports Function calling, Parallel function calling, Reasoning, Structured output, Tool choice, Vision, Web search.

What account tier is required to use gpt-oss-120b on AvalAI?

Calling gpt-oss-120b on AvalAI requires the Basic tier (tier 0) or higher.

Third-party data and freshness

  • Pricing, availability, endpoints, and rate limits: AvalAI structured data (authoritative)
  • OpenRouter marketplace · Last checked: 2026-08-17