Developer Dashboard

Back to Model Explorer

GoogleAvailable on AvalAI

Gemini 3.1 Flash Lite Preview API on AvalAI

gemini-3.1-flash-lite-preview

Gemini 3.1 Flash Lite Preview is an AI model from Google that is available through AvalAI's OpenAI-compatible API.

Price display scale
Input price$0.25 / 1M tokens
Output price$1.50 / 1M tokens
Context window1,048,576 tokens
Maximum output65,536 tokens
Release date2026-03-03

Use Gemini 3.1 Flash Lite Preview with the AvalAI API

Keep the API key in an environment variable and send requests from server-side code.

bash
curl https://api.avalai.ir/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $AVALAI_API_KEY" \
  -d '{"model": "gemini-3.1-flash-lite-preview", "messages": [{"role": "user", "content": "Give a concise, practical solution."}]}'
python
import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["AVALAI_API_KEY"],
    base_url="https://api.avalai.ir/v1",
)

response = client.chat.completions.create(
    model="gemini-3.1-flash-lite-preview",
    messages=[{"role": "user", "content": "Give a concise, practical solution."}],
)

print(response.choices[0].message.content)
javascript
import OpenAI from "openai";

const client = new OpenAI({
  apiKey: process.env.AVALAI_API_KEY,
  baseURL: "https://api.avalai.ir/v1",
});

const response = await client.chat.completions.create({
  model: "gemini-3.1-flash-lite-preview",
  messages: [{ role: "user", content: "Give a concise, practical solution." }],
});

console.log(response.choices[0].message.content);

Create an API keyRead the quickstart

Capabilities and endpoints

  • audio_input
  • function_calling
  • native_streaming
  • parallel_function_calling
  • pdf_input
  • prompt_caching
  • reasoning
  • response_schema
  • system_messages
  • tool_choice
  • url_context
  • video_input
  • vision
  • web_search

Endpoints

  • /v1/chat/completions
  • /v1/completions
  • /v1/batch

Third-party model parameters

Third-party reference data does not change AvalAI endpoint support; use the capabilities and endpoints above.

  • include_reasoning
  • max_tokens
  • reasoning
  • reasoning_effort
  • response_format
  • seed
  • structured_outputs
  • temperature
  • tool_choice
  • tools
  • top_p

Supplemental reasoning effort values

high, low, medium, minimal

Gemini 3.1 Flash Lite Preview rate limits

TierRPMTPM
0340,000
150500,000
25001,000,000
31,0003,000,000
43,5005,000,000
525,00020,000,000

AvalAI and reference marketplace pricing

Token roleAvalAI final price0% platform feeOpenRouter cost breakdown
Model rateCredit purchase feeEffective OpenRouter cost after fee
Input price$0.25 / 1M tokens$0.25 / 1M tokens5.5%$0.26 / 1M tokens
Output price$1.50 / 1M tokens$1.50 / 1M tokens5.5%$1.58 / 1M tokens

AvalAI pricing is final with a 0% platform fee. The effective OpenRouter prices add its 5.5% credit purchase fee to the highest reported model rate, including long-context tiers. OpenRouter charges a 5.5% fee when credits are purchased ($0.80 minimum), so small credit purchases can have a higher effective cost. OpenRouter FAQ

Peer benchmark performance

Independent, sourced scores with no composite rating.

Artificial Analysis Agentic Index

6.5index
ScorePeer rankDirectionSource and date
6.5 indexNo comparable dataHigher is betterOpenRouter · 2026-08-15

Artificial Analysis Coding Index

34.7index
ScorePeer rankDirectionSource and date
34.7 indexNo comparable dataHigher is betterOpenRouter · 2026-08-15

Artificial Analysis Intelligence Index

25.6index
ScorePeer rankDirectionSource and date
25.6 indexNo comparable dataHigher is betterOpenRouter · 2026-08-15

Third-party model data

Source: OpenRouter

Independent sources supplement this context; AvalAI remains authoritative for pricing, availability, and endpoint support.

Reviewed snapshot 2026-08-15

Provider count
Context window range
1,048,576
Maximum output range
65,536
OpenRouter model input rate range
$0.25 / 1M
OpenRouter model output rate range
$1.5 / 1M
OpenRouter fee
5.5% credit purchase fee · OpenRouter fee details
Effective input price range after fee
$0.2638 / 1M
Effective output price range after fee
$1.5825 / 1M

Frequently asked questions

What is the Gemini 3.1 Flash Lite Preview API on AvalAI?

Create an AvalAI API key and send the model ID gemini-3.1-flash-lite-preview to one of the supported endpoints shown on this page.

How much does the Gemini 3.1 Flash Lite Preview API cost on AvalAI?

The Gemini 3.1 Flash Lite Preview API on AvalAI costs $0.25 per 1M input tokens and $1.50 per 1M output tokens.

What is the token cost for Gemini 3.1 Flash Lite Preview?

Current AvalAI pricing for Gemini 3.1 Flash Lite Preview is $0.25 / 1M tokens for input and $1.50 / 1M tokens for output.

What is the Gemini 3.1 Flash Lite Preview context window?

The recorded maximum input context for Gemini 3.1 Flash Lite Preview is 1,048,576 tokens.

Third-party data and freshness

  • Pricing, availability, endpoints, and rate limits: AvalAI structured data (authoritative)
  • OpenRouter marketplace · Last checked: 2026-08-15