← Model catalog

Gemini 2.5 Flash

everyais/gemini-2-5-flash

Gemini 2.5 Flash (everyais/gemini-2-5-flash) availability, capabilities, context limits, and public reference pricing on everyais.

Model family geminiInput Text · ImageOutput Text
Available now

Method

The page starts with current public catalog metadata and reference costs; operational metrics are added only when measured samples are available.

Source and update

GET /models/catalog

Updated

Model type
Chat
Released
2025-06-17
Context window
1.0M
Price unit
Per 1M tokens
Maximum output
65,536
Model reference price (USD)

Standard (≤200K)

Input / 1M tokens
$0.15
Output / 1M tokens
$0.6
Cache read / 1M tokens
$0.037

Long context (>200K, full request)

Input / 1M tokens
$0.3
Output / 1M tokens
$2.5
Cache read / 1M tokens
$0.03

Capabilities

  • Streaming
  • Tool use
  • Vision
  • JSON
  • Reasoning
  • Sampling controls
  • Structured outputs

Supported endpoints

  • /v1/chat/completions

Benchmarks

Evaluation results published by external sources. Only mappings approved by a human administrator are shown.

Model benchmark scores
BenchmarkScoreSource
frontiermath4.8%Epoch AI Benchmarking HubCC-BYSource model label: gemini-2.5-flash
terminal-bench17.1%Epoch AI Benchmarking HubCC-BYSource model label: gemini-2.5-flash

Scores are reproduced as published; everyais does not re-measure them. Scores using different units cannot be compared. See benchmark sources and licenses.

Browse benchmark rankings

Usage and availability trend

Daily tokensInputOutput
2K1K02026-08-272026-09-032026-08-27 · Requests 1 · Input 49 · Output 494 · Availability Insufficient sample2026-08-28 · Requests 0 · Input 0 · Output 0 · Availability Insufficient sample2026-08-29 · Requests 0 · Input 0 · Output 0 · Availability Insufficient sample2026-08-30 · Requests 0 · Input 0 · Output 0 · Availability Insufficient sample2026-08-31 · Requests 1 · Input 197 · Output 60 · Availability Insufficient sample2026-09-01 · Requests 3 · Input 59 · Output 30 · Availability Insufficient sample2026-09-02 · Requests 1 · Input 7 · Output 22 · Availability Insufficient sample2026-09-03 · Requests 1 · Input 1,071 · Output 140 · Availability Insufficient sample

Last 8 days · 2026-09-03 Input 1,071 · Output 140 tokens.

Daily availability (success rate %)

Insufficient sample — not enough requests to publish a success rate.

Code example

Use the OpenAI SDK by changing only base_url.

from openai import OpenAI

client = OpenAI(
    api_key="everyais_...",
    base_url="https://api.everyais.com/v1",
)

response = client.chat.completions.create(
    model="everyais/gemini-2-5-flash",
    messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)
Full API docs

Other models in this family

Compare models side by side

Compare capabilities, context windows, and reference prices to find a model for your use case. Your actual billed rate can differ by account.