Gemini 3.7 Flash API on AvalAI
gemini-3.7-flashGemini 3.7 Flash is an AI model from Google that is available through AvalAI's OpenAI-compatible API.
Use Gemini 3.7 Flash with the AvalAI API
Keep the API key in an environment variable and send requests from server-side code.
curl https://api.avalai.ir/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $AVALAI_API_KEY" \
-d '{"model": "gemini-3.7-flash", "messages": [{"role": "user", "content": "Give a concise, practical solution."}]}'import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["AVALAI_API_KEY"],
base_url="https://api.avalai.ir/v1",
)
response = client.chat.completions.create(
model="gemini-3.7-flash",
messages=[{"role": "user", "content": "Give a concise, practical solution."}],
)
print(response.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
apiKey: process.env.AVALAI_API_KEY,
baseURL: "https://api.avalai.ir/v1",
});
const response = await client.chat.completions.create({
model: "gemini-3.7-flash",
messages: [{ role: "user", content: "Give a concise, practical solution." }],
});
console.log(response.choices[0].message.content);Create an API keyRead the quickstart
Capabilities and endpoints
audio_inputfunction_callingnative_streamingparallel_function_callingpdf_inputprompt_cachingreasoningresponse_schemasystem_messagestool_choiceurl_contextvideo_inputvisionweb_search
Endpoints
/v1/chat/completions/v1/completions/v1/batch
Third-party model parameters
Third-party reference data does not change AvalAI endpoint support; use the capabilities and endpoints above.
include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_p
Supplemental reasoning effort values
high, low, medium
Gemini 3.7 Flash rate limits
| Tier | RPM | TPM |
|---|---|---|
| 0 | 1 | 40,000 |
| 1 | 50 | 500,000 |
| 2 | 250 | 1,000,000 |
| 3 | 1,000 | 2,000,000 |
| 4 | 3,500 | 5,000,000 |
| 5 | 25,000 | 30,000,000 |
AvalAI and reference marketplace pricing
| Token role | AvalAI final price0% platform fee | OpenRouter cost breakdown | ||
|---|---|---|---|---|
| Model rate | Credit purchase fee | Effective OpenRouter cost after fee | ||
| Input price | $0.75 / 1M tokens | $0.38 / 1M tokens | 5.5% | $0.40 / 1M tokens |
| Output price | $3.75 / 1M tokens | $1.88 / 1M tokens | 5.5% | $1.98 / 1M tokens |
AvalAI pricing is final with a 0% platform fee. The effective OpenRouter prices add its 5.5% credit purchase fee to the highest reported model rate, including long-context tiers. OpenRouter charges a 5.5% fee when credits are purchased ($0.80 minimum), so small credit purchases can have a higher effective cost. OpenRouter FAQ
Peer benchmark performance
Independent, sourced scores with no composite rating.
Artificial Analysis Agentic Index
| Score | Peer rank | Direction | Source and date |
|---|---|---|---|
| 45.1 index | No comparable data | Higher is better | OpenRouter · 2026-08-14 |
Artificial Analysis Coding Index
| Score | Peer rank | Direction | Source and date |
|---|---|---|---|
| 76.1 index | No comparable data | Higher is better | OpenRouter · 2026-08-14 |
Artificial Analysis Intelligence Index
| Score | Peer rank | Direction | Source and date |
|---|---|---|---|
| 56 index | No comparable data | Higher is better | OpenRouter · 2026-08-14 |
Third-party model data
Source: OpenRouterIndependent sources supplement this context; AvalAI remains authoritative for pricing, availability, and endpoint support.
Reviewed snapshot 2026-08-14
- Provider count
- —
- Context window range
- 1,048,576
- Maximum output range
- 65,536
- OpenRouter model input rate range
- $0.375 / 1M
- OpenRouter model output rate range
- $1.875 / 1M
- OpenRouter fee
- 5.5% credit purchase fee · OpenRouter fee details
- Effective input price range after fee
- $0.3956 / 1M
- Effective output price range after fee
- $1.9781 / 1M
Related models
Frequently asked questions
What is the Gemini 3.7 Flash API on AvalAI?
Create an AvalAI API key and send the model ID gemini-3.7-flash to one of the supported endpoints shown on this page.
How much does the Gemini 3.7 Flash API cost on AvalAI?
The Gemini 3.7 Flash API on AvalAI costs $0.75 per 1M input tokens and $3.75 per 1M output tokens.
What is the token cost for Gemini 3.7 Flash?
Current AvalAI pricing for Gemini 3.7 Flash is $0.75 / 1M tokens for input and $3.75 / 1M tokens for output.
What is the Gemini 3.7 Flash context window?
The recorded maximum input context for Gemini 3.7 Flash is 1,048,576 tokens.
Third-party data and freshness
- Pricing, availability, endpoints, and rate limits: AvalAI structured data (authoritative)
- OpenRouter marketplace · Last checked: 2026-08-14