Developer Dashboard

Back to Model Explorer

Together AiAvailable on AvalAI

llama-4-scout-17b-16e-instruct API on AvalAI

llama-4-scout-17b-16e-instruct

llama-4-scout-17b-16e-instruct is an AI model from Meta that is available through AvalAI's OpenAI-compatible API.

Price display scale
Input price$0.20 / 1M tokens
Output price$0.78 / 1M tokens
Context window10,000,000 tokens
Maximum output16,384 tokens

Use llama-4-scout-17b-16e-instruct with the AvalAI API

Keep the API key in an environment variable and send requests from server-side code.

bash
curl https://api.avalai.ir/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $AVALAI_API_KEY" \
  -d '{"model": "llama-4-scout-17b-16e-instruct", "messages": [{"role": "user", "content": "Give a concise, practical solution."}]}'
python
import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["AVALAI_API_KEY"],
    base_url="https://api.avalai.ir/v1",
)

response = client.chat.completions.create(
    model="llama-4-scout-17b-16e-instruct",
    messages=[{"role": "user", "content": "Give a concise, practical solution."}],
)

print(response.choices[0].message.content)
javascript
import OpenAI from "openai";

const client = new OpenAI({
  apiKey: process.env.AVALAI_API_KEY,
  baseURL: "https://api.avalai.ir/v1",
});

const response = await client.chat.completions.create({
  model: "llama-4-scout-17b-16e-instruct",
  messages: [{ role: "user", content: "Give a concise, practical solution." }],
});

console.log(response.choices[0].message.content);

Create an API keyRead the quickstart

Capabilities and endpoints

  • Function calling
  • Structured output
  • Tool choice
  • Vision

llama-4-scout-17b-16e-instruct rate limits

TierRPMTPM
Basic340,000
Tier 150400,000
Tier 21501,000,000
Tier 32502,000,000
Tier 45004,000,000
Tier 51,50010,000,000

Frequently asked questions

What is the llama-4-scout-17b-16e-instruct API on AvalAI?

Create an AvalAI API key and send the model ID llama-4-scout-17b-16e-instruct to one of the supported endpoints shown on this page.

How much does the llama-4-scout-17b-16e-instruct API cost on AvalAI?

The llama-4-scout-17b-16e-instruct API on AvalAI costs $0.20 per 1M input tokens and $0.78 per 1M output tokens.

What is the token cost for llama-4-scout-17b-16e-instruct?

Current AvalAI pricing for llama-4-scout-17b-16e-instruct is $0.20 / 1M tokens for input and $0.78 / 1M tokens for output.

What is the llama-4-scout-17b-16e-instruct context window?

The recorded maximum input context for llama-4-scout-17b-16e-instruct is 10,000,000 tokens.

What are the rate limits for llama-4-scout-17b-16e-instruct on AvalAI?

On tier 5, llama-4-scout-17b-16e-instruct on AvalAI supports up to 1,500 requests per minute and 10,000,000 tokens per minute.

What features does llama-4-scout-17b-16e-instruct support on AvalAI?

llama-4-scout-17b-16e-instruct on AvalAI supports Function calling, Structured output, Tool choice, Vision.

What account tier is required to use llama-4-scout-17b-16e-instruct on AvalAI?

Calling llama-4-scout-17b-16e-instruct on AvalAI requires the Basic tier (tier 0) or higher.

Third-party data and freshness

  • Pricing, availability, endpoints, and rate limits: AvalAI structured data (authoritative)