Developer Dashboard

Back to Model Explorer

GoogleAvailable on AvalAI

gemini-flash-lite-latest API on AvalAI

gemini-flash-lite-latest

gemini-flash-lite-latest is an AI model from Google that is available through AvalAI's OpenAI-compatible API.

Price display scale
Input price$0.10 / 1M tokens
Output price$0.40 / 1M tokens
Context window1,048,576 tokens
Maximum output65,535 tokens

Use gemini-flash-lite-latest with the AvalAI API

Keep the API key in an environment variable and send requests from server-side code.

bash
curl https://api.avalai.ir/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $AVALAI_API_KEY" \
  -d '{"model": "gemini-flash-lite-latest", "messages": [{"role": "user", "content": "Give a concise, practical solution."}]}'
python
import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["AVALAI_API_KEY"],
    base_url="https://api.avalai.ir/v1",
)

response = client.chat.completions.create(
    model="gemini-flash-lite-latest",
    messages=[{"role": "user", "content": "Give a concise, practical solution."}],
)

print(response.choices[0].message.content)
javascript
import OpenAI from "openai";

const client = new OpenAI({
  apiKey: process.env.AVALAI_API_KEY,
  baseURL: "https://api.avalai.ir/v1",
});

const response = await client.chat.completions.create({
  model: "gemini-flash-lite-latest",
  messages: [{ role: "user", content: "Give a concise, practical solution." }],
});

console.log(response.choices[0].message.content);

Create an API keyRead the quickstart

Capabilities and endpoints

  • Function calling
  • Parallel function calling
  • PDF input
  • Prompt caching
  • Reasoning
  • Structured output
  • System messages
  • Tool choice
  • URL context
  • Vision
  • Web search

Endpoints

  • /v1/chat/completions
  • /v1/completions
  • /v1/batch

gemini-flash-lite-latest rate limits

TierRPMTPM
Basic340,000
Tier 150800,000
Tier 21002,000,000
Tier 35005,000,000
Tier 47508,000,000
Tier 510,00015,000,000

Frequently asked questions

What is the gemini-flash-lite-latest API on AvalAI?

Create an AvalAI API key and send the model ID gemini-flash-lite-latest to one of the supported endpoints shown on this page.

How much does the gemini-flash-lite-latest API cost on AvalAI?

The gemini-flash-lite-latest API on AvalAI costs $0.10 per 1M input tokens and $0.40 per 1M output tokens.

What is the token cost for gemini-flash-lite-latest?

Current AvalAI pricing for gemini-flash-lite-latest is $0.10 / 1M tokens for input and $0.40 / 1M tokens for output.

What is the gemini-flash-lite-latest context window?

The recorded maximum input context for gemini-flash-lite-latest is 1,048,576 tokens.

What are the rate limits for gemini-flash-lite-latest on AvalAI?

On tier 5, gemini-flash-lite-latest on AvalAI supports up to 10,000 requests per minute and 15,000,000 tokens per minute.

What features does gemini-flash-lite-latest support on AvalAI?

gemini-flash-lite-latest on AvalAI supports Function calling, Parallel function calling, PDF input, Prompt caching, Reasoning, Structured output, System messages, Tool choice, URL context, Vision, Web search.

What account tier is required to use gemini-flash-lite-latest on AvalAI?

Calling gemini-flash-lite-latest on AvalAI requires the Basic tier (tier 0) or higher.

Third-party data and freshness

  • Pricing, availability, endpoints, and rate limits: AvalAI structured data (authoritative)