Developer Dashboard

Back to Model Explorer

GoogleAvailable on AvalAI

Gemma 4 31B API on AvalAI

gemma-4-31b-it

Gemma 4 31B is an AI model from Google that is available through AvalAI's OpenAI-compatible API.

Price display scale
Input price$0.14 / 1M tokens
Output price$0.40 / 1M tokens
Context window262,144 tokens
Maximum output131,072 tokens
Release date2026-04-02

Use Gemma 4 31B with the AvalAI API

Keep the API key in an environment variable and send requests from server-side code.

bash
curl https://api.avalai.ir/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $AVALAI_API_KEY" \
  -d '{"model": "gemma-4-31b-it", "messages": [{"role": "user", "content": "Give a concise, practical solution."}]}'
python
import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["AVALAI_API_KEY"],
    base_url="https://api.avalai.ir/v1",
)

response = client.chat.completions.create(
    model="gemma-4-31b-it",
    messages=[{"role": "user", "content": "Give a concise, practical solution."}],
)

print(response.choices[0].message.content)
javascript
import OpenAI from "openai";

const client = new OpenAI({
  apiKey: process.env.AVALAI_API_KEY,
  baseURL: "https://api.avalai.ir/v1",
});

const response = await client.chat.completions.create({
  model: "gemma-4-31b-it",
  messages: [{ role: "user", content: "Give a concise, practical solution." }],
});

console.log(response.choices[0].message.content);

Create an API keyRead the quickstart

Capabilities and endpoints

  • Function calling
  • Tool choice
  • Vision

Third-party model parameters

Third-party reference data does not change AvalAI endpoint support; use the capabilities and endpoints above.

  • frequency_penalty
  • include_reasoning
  • logit_bias
  • logprobs
  • max_tokens
  • min_p
  • presence_penalty
  • reasoning
  • repetition_penalty
  • response_format
  • seed
  • stop
  • structured_outputs
  • temperature
  • tool_choice
  • tools
  • top_k
  • top_logprobs
  • top_p

Gemma 4 31B rate limits

TierRPMTPM
Basic15,000
Tier 1510,000
Tier 21015,000
Tier 31520,000
Tier 42025,000
Tier 560010,000,000

AvalAI and reference marketplace pricing

Token roleAvalAI final price0% platform feeOpenRouter cost breakdown
Model rateCredit purchase feeEffective OpenRouter cost after fee
Input price$0.14 / 1M tokens$0.09 / 1M tokens5.5%$0.09 / 1M tokens
Output price$0.40 / 1M tokens$0.34 / 1M tokens5.5%$0.36 / 1M tokens

AvalAI pricing is final with a 0% platform fee. The effective OpenRouter prices add its 5.5% credit purchase fee to the highest reported model rate, including long-context tiers. OpenRouter charges a 5.5% fee when credits are purchased ($0.80 minimum), so small credit purchases can have a higher effective cost. OpenRouter FAQ

Peer benchmark performance

Independent, sourced scores with no composite rating.

Artificial Analysis Agentic Index

6.8index
ScorePeer rankDirectionSource and date
6.8 indexNo comparable dataHigher is betterOpenRouter · 2026-09-05

Artificial Analysis Coding Index

43.4index
ScorePeer rankDirectionSource and date
43.4 indexNo comparable dataHigher is betterOpenRouter · 2026-09-05

Third-party model data

Source: OpenRouter

Independent sources supplement this context; AvalAI remains authoritative for pricing, availability, and endpoint support.

Reviewed snapshot 2026-09-05

Provider count
Context window range
262,144
Maximum output range
16,384
OpenRouter model input rate range
$0.09 / 1M
OpenRouter model output rate range
$0.34 / 1M
OpenRouter fee
5.5% credit purchase fee · OpenRouter fee details
Effective input price range after fee
$0.0949 / 1M
Effective output price range after fee
$0.3587 / 1M

Frequently asked questions

What is the Gemma 4 31B API on AvalAI?

Create an AvalAI API key and send the model ID gemma-4-31b-it to one of the supported endpoints shown on this page.

How much does the Gemma 4 31B API cost on AvalAI?

The Gemma 4 31B API on AvalAI costs $0.14 per 1M input tokens and $0.40 per 1M output tokens.

What is the token cost for Gemma 4 31B?

Current AvalAI pricing for Gemma 4 31B is $0.14 / 1M tokens for input and $0.40 / 1M tokens for output.

What is the Gemma 4 31B context window?

The recorded maximum input context for Gemma 4 31B is 262,144 tokens.

What are the rate limits for Gemma 4 31B on AvalAI?

On tier 5, Gemma 4 31B on AvalAI supports up to 600 requests per minute and 10,000,000 tokens per minute.

What features does Gemma 4 31B support on AvalAI?

Gemma 4 31B on AvalAI supports Function calling, Tool choice, Vision.

What account tier is required to use Gemma 4 31B on AvalAI?

Calling Gemma 4 31B on AvalAI requires the Basic tier (tier 0) or higher.

Third-party data and freshness

  • Pricing, availability, endpoints, and rate limits: AvalAI structured data (authoritative)
  • OpenRouter marketplace · Last checked: 2026-09-05