GPT-6 Luna API on AvalAI
gpt-6-lunaGPT-6 Luna is OpenAI's cost-efficient GPT-6 model for high-volume reasoning, coding, and everyday professional workflows. AvalAI supports Chat Completions, Messages, and Responses; hosted tools require separate route support.
Use GPT-6 Luna with the AvalAI API
Keep the API key in an environment variable and send requests from server-side code.
curl https://api.avalai.ir/v1/responses \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $AVALAI_API_KEY" \
-d '{"model": "gpt-6-luna", "input": "Give a concise, practical solution."}'import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["AVALAI_API_KEY"],
base_url="https://api.avalai.ir/v1",
)
response = client.responses.create(
model="gpt-6-luna",
input="Give a concise, practical solution.",
)
print(response.output_text)import OpenAI from "openai";
const client = new OpenAI({
apiKey: process.env.AVALAI_API_KEY,
baseURL: "https://api.avalai.ir/v1",
});
const response = await client.responses.create({
model: "gpt-6-luna",
input: "Give a concise, practical solution.",
});
console.log(response.output_text);Create an API keyRead the quickstart
Best suited for
- High-volume support and summarization
- Cost-sensitive coding workflows
- Repeated-context tasks with prompt caching
Capabilities and endpoints
- Function calling
- Streaming
- Parallel function calling
- PDF input
- Prompt caching
- Reasoning
- Structured output
- System messages
- Tool choice
- Vision
- Web search
Endpoints
/v1/chat/completions/v1/messages/v1/responses
Third-party model parameters
Third-party reference data does not change AvalAI endpoint support; use the capabilities and endpoints above.
include_reasoningmax_completion_tokensmax_tokensreasoningreasoning_effortresponse_formatseedstructured_outputstool_choicetools
Supplemental reasoning effort values
high, low, max, medium, none, xhigh
GPT-6 Luna rate limits
| Tier | RPM | TPM |
|---|---|---|
| Basic | 1 | 10,000 |
| Tier 1 | 50 | 500,000 |
| Tier 2 | 150 | 2,000,000 |
| Tier 3 | 250 | 4,000,000 |
| Tier 4 | 1,500 | 8,000,000 |
| Tier 5 | 10,000 | 20,000,000 |
AvalAI and reference marketplace pricing
| Token role | AvalAI final price0% platform fee | OpenRouter cost breakdown | ||
|---|---|---|---|---|
| Model rate | Credit purchase fee | Effective OpenRouter cost after fee | ||
| Input price | $0.10 / 1M tokens | $0.10 / 1M tokens | 5.5% | $0.11 / 1M tokens |
| Input price (over 272,000 tokens) | $0.20 / 1M tokens | $0.20 / 1M tokens | 5.5% | $0.21 / 1M tokens |
| Output price | $0.50 / 1M tokens | $0.50 / 1M tokens | 5.5% | $0.53 / 1M tokens |
| Output price (over 272,000 tokens) | $0.75 / 1M tokens | $0.75 / 1M tokens | 5.5% | $0.79 / 1M tokens |
AvalAI pricing is final with a 0% platform fee. The effective OpenRouter prices add its 5.5% credit purchase fee to the highest reported model rate, including long-context tiers. OpenRouter charges a 5.5% fee when credits are purchased ($0.80 minimum), so small credit purchases can have a higher effective cost. OpenRouter FAQ
Peer benchmark performance
Independent, sourced scores with no composite rating.
Artificial Analysis Intelligence Index
| Score | Peer rank | Direction | Source and date |
|---|---|---|---|
| 37.3 index | No comparable data | Higher is better | OpenRouter · 2026-09-24 |
Third-party model data
Source: OpenRouterIndependent sources supplement this context; AvalAI remains authoritative for pricing, availability, and endpoint support.
Reviewed snapshot 2026-09-24
- Provider count
- —
- Context window range
- 1,050,000
- Maximum output range
- 128,000
- OpenRouter model input rate range
- $0.1 / 1M
- OpenRouter model output rate range
- $0.5 / 1M
- OpenRouter fee
- 5.5% credit purchase fee · OpenRouter fee details
- Effective input price range after fee
- $0.1055 / 1M
- Effective output price range after fee
- $0.5275 / 1M
Related models
Frequently asked questions
What is the GPT-6 Luna API on AvalAI?
Create an AvalAI API key and send the model ID gpt-6-luna to one of the supported endpoints shown on this page.
How much does the GPT-6 Luna API cost on AvalAI?
The GPT-6 Luna API on AvalAI costs $0.10 per 1M input tokens and $0.50 per 1M output tokens.
What is the token cost for GPT-6 Luna?
Current AvalAI pricing for GPT-6 Luna is $0.10 / 1M tokens for input and $0.50 / 1M tokens for output.
What is the GPT-6 Luna context window?
The recorded maximum input context for GPT-6 Luna is 922,000 tokens.
What are the rate limits for GPT-6 Luna on AvalAI?
On tier 5, GPT-6 Luna on AvalAI supports up to 10,000 requests per minute and 20,000,000 tokens per minute.
What features does GPT-6 Luna support on AvalAI?
GPT-6 Luna on AvalAI supports Function calling, Streaming, Parallel function calling, PDF input, Prompt caching, Reasoning, Structured output, System messages, Tool choice, Vision, Web search.
What account tier is required to use GPT-6 Luna on AvalAI?
Calling GPT-6 Luna on AvalAI requires the Basic tier (tier 0) or higher.
Third-party data and freshness
- Pricing, availability, endpoints, and rate limits: AvalAI structured data (authoritative)
- OpenRouter marketplace · Last checked: 2026-09-24