GPT-4o-mini (2024-07-18) API on AvalAI
gpt-4o-mini-2024-07-18GPT-4o-mini (2024-07-18) is an AI model from OpenAI that is available through AvalAI's OpenAI-compatible API.
Use GPT-4o-mini (2024-07-18) with the AvalAI API
Keep the API key in an environment variable and send requests from server-side code.
curl https://api.avalai.ir/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $AVALAI_API_KEY" \
-d '{"model": "gpt-4o-mini-2024-07-18", "messages": [{"role": "user", "content": "Give a concise, practical solution."}]}'import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["AVALAI_API_KEY"],
base_url="https://api.avalai.ir/v1",
)
response = client.chat.completions.create(
model="gpt-4o-mini-2024-07-18",
messages=[{"role": "user", "content": "Give a concise, practical solution."}],
)
print(response.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
apiKey: process.env.AVALAI_API_KEY,
baseURL: "https://api.avalai.ir/v1",
});
const response = await client.chat.completions.create({
model: "gpt-4o-mini-2024-07-18",
messages: [{ role: "user", content: "Give a concise, practical solution." }],
});
console.log(response.choices[0].message.content);Create an API keyRead the quickstart
Capabilities and endpoints
- Function calling
- Parallel function calling
- PDF input
- Prompt caching
- Structured output
- System messages
- Tool choice
- Vision
Third-party model parameters
Third-party reference data does not change AvalAI endpoint support; use the capabilities and endpoints above.
frequency_penaltylogit_biaslogprobsmax_tokenspredictionpresence_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_logprobstop_pweb_search_options
GPT-4o-mini (2024-07-18) rate limits
| Tier | RPM | TPM |
|---|---|---|
| Basic | 3 | 40,000 |
| Tier 1 | 500 | 500,000 |
| Tier 2 | 5,000 | 3,000,000 |
| Tier 3 | 5,000 | 5,000,000 |
| Tier 4 | 10,000 | 20,000,000 |
| Tier 5 | 30,000 | 150,000,000 |
AvalAI and reference marketplace pricing
| Token role | AvalAI final price0% platform fee | OpenRouter cost breakdown | ||
|---|---|---|---|---|
| Model rate | Credit purchase fee | Effective OpenRouter cost after fee | ||
| Input price | $0.15 / 1M tokens | $0.15 / 1M tokens | 5.5% | $0.16 / 1M tokens |
| Output price | $0.60 / 1M tokens | $0.60 / 1M tokens | 5.5% | $0.63 / 1M tokens |
AvalAI pricing is final with a 0% platform fee. The effective OpenRouter prices add its 5.5% credit purchase fee to the highest reported model rate, including long-context tiers. OpenRouter charges a 5.5% fee when credits are purchased ($0.80 minimum), so small credit purchases can have a higher effective cost. OpenRouter FAQ
Third-party model data
Source: OpenRouterIndependent sources supplement this context; AvalAI remains authoritative for pricing, availability, and endpoint support.
Reviewed snapshot 2026-07-31 · This snapshot was checked more than 14 days ago.
- Provider count
- —
- Context window range
- 128,000
- Maximum output range
- 16,384
- OpenRouter model input rate range
- $0.15 / 1M
- OpenRouter model output rate range
- $0.6 / 1M
- OpenRouter fee
- 5.5% credit purchase fee · OpenRouter fee details
- Effective input price range after fee
- $0.1582 / 1M
- Effective output price range after fee
- $0.633 / 1M
Related models
- gpt-4o-mini-tts
- gpt-4o-mini-transcribe
- gpt-4o-mini-search-preview-2025-03-11
- gpt-4o-mini-search-preview
Frequently asked questions
What is the GPT-4o-mini (2024-07-18) API on AvalAI?
Create an AvalAI API key and send the model ID gpt-4o-mini-2024-07-18 to one of the supported endpoints shown on this page.
How much does the GPT-4o-mini (2024-07-18) API cost on AvalAI?
The GPT-4o-mini (2024-07-18) API on AvalAI costs $0.15 per 1M input tokens and $0.60 per 1M output tokens.
What is the token cost for GPT-4o-mini (2024-07-18)?
Current AvalAI pricing for GPT-4o-mini (2024-07-18) is $0.15 / 1M tokens for input and $0.60 / 1M tokens for output.
What is the GPT-4o-mini (2024-07-18) context window?
The recorded maximum input context for GPT-4o-mini (2024-07-18) is 128,000 tokens.
What are the rate limits for GPT-4o-mini (2024-07-18) on AvalAI?
On tier 5, GPT-4o-mini (2024-07-18) on AvalAI supports up to 30,000 requests per minute and 150,000,000 tokens per minute.
What features does GPT-4o-mini (2024-07-18) support on AvalAI?
GPT-4o-mini (2024-07-18) on AvalAI supports Function calling, Parallel function calling, PDF input, Prompt caching, Structured output, System messages, Tool choice, Vision.
What account tier is required to use GPT-4o-mini (2024-07-18) on AvalAI?
Calling GPT-4o-mini (2024-07-18) on AvalAI requires the Basic tier (tier 0) or higher.
Third-party data and freshness
- Pricing, availability, endpoints, and rate limits: AvalAI structured data (authoritative)
- OpenRouter marketplace · Last checked: 2026-07-31