Gemini 3.8 Flash
everyais/gemini-3-8-flash
Gemini 3.8 Flash (everyais/gemini-3-8-flash) availability, capabilities, context limits, and public reference pricing on everyais.
Model family geminiInput Text · ImageOutput Text
Available now
Method
The page starts with current public catalog metadata and reference costs; operational metrics are added only when measured samples are available.
Source and update
GET /models/catalogUpdated
- Model type
- Chat
- Released
- Not published
- Context window
- 1.0M
- Price unit
- Per 1M tokens
- Maximum output
- 65,536
- Model reference price (USD)
- Input / 1M tokens
- $0.75
- Output / 1M tokens
- $3.75
- Cache read / 1M tokens
- $0.075
- Web search reference price (USD)
- $0.014 /queryAdded to token pricing
Capabilities
- ✓ Streaming
- ✓ Tool use
- ✓ Vision
- ✓ JSON with tools
- ✓ Web search with tools
- ✓ JSON
- ✓ Reasoning
- ✓ Structured outputs
- ✓ Web search
Supported endpoints
- /v1/chat/completions
Usage and availability trend
Collecting data. Trends appear after more than one day of usage is recorded.
Code example
Use the OpenAI SDK by changing only base_url.
from openai import OpenAI
client = OpenAI(
api_key="everyais_...",
base_url="https://api.everyais.com/v1",
)
response = client.chat.completions.create(
model="everyais/gemini-3-8-flash",
messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)Full API docsOther models in this family
Compare models side by side
Compare capabilities, context windows, and reference prices to find a model for your use case. Your actual billed rate can differ by account.