Skip to content

Gemini 3.7 Flash

ReasoningGoogle

Fast multimodal model for agentic work and coding.

Input
$0.225
per 1M tokens
Output
$1.125
per 1M tokens
Context window
1M
tokens
Max output
64K
tokens

Pricing

Prices are in USD per 1M tokens. Cached input is billed at the cache read price.

Input
$0.225/ 1M
Output
$1.125/ 1M
Cache read
$0.0225/ 1M
Audio input
$0.225/ 1M

Use it from code

The samples read your key from the ORYNX_API_KEY environment variable.

Gemini
curl "https://api.orynx.dev/v1beta/models/gemini-3.7-flash:generateContent" \
  -H "x-goog-api-key: $ORYNX_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "contents": [{"parts": [{"text": "Explain what an API gateway does in one sentence."}]}]
  }'
gemini-3.8-flash

Google's most intelligent Flash model, with an automatic thinking budget.

Input
$0.225/ 1M
Output
$1.125/ 1M
Gemini1M ctx
gemini-3.6-flash

High-efficiency Gemini for coding and agents.

Input
$0.225/ 1M
Output
$1.125/ 1M
Gemini1M ctx
gemini-3.6-flash-tiered

Gemini 3.6 Flash priced per token instead of per context tier.

Input
$0.45/ 1M
Output
$2.25/ 1M
Gemini1M ctx