ModelsOk

gemini-3.5-flash API pricing & access

gemini-3.5-flash API pricing: $1.5 per 1M tokens. Compare supported endpoints, capabilities and access options on ModelsOk.

gemini-3.5-flash API access

gemini-3.5-flash。支持文本、图像、视频、音频、文档输入,输出文本。所列规格依据厂商公开资料,实际可用功能与上限以所选服务渠道为准。

Compare supported API endpoints and capabilities, then access gemini-3.5-flash through the unified API gateway.

Category
text
Context length
1048576 tokens
Maximum output
65536 tokens
Supported APIs
Gemini API, OpenAI Chat Completions API, Anthropic Messages API
API endpoints
POST /v1beta/models/{model}:generateContent, POST /v1/chat/completions, POST /v1/messages
Input modalities
text, image, video, audio, file
Output modalities
text
Capabilities
function calling, structured output, reasoning, caching, vision

gemini-3.5-flash API pricing

Base input price
$1.5 per 1M tokens
Base output price
$9 per 1M tokens

Base prices are shown in USD before group-specific adjustments. Open the live pricing page for current access-group prices.

How to call gemini-3.5-flash

gemini-3.5-flash is served through ModelsOk's unified gateway. Existing OpenAI or Anthropic SDK code keeps working: point base_url at https://modelsok.com and use your ModelsOk API key.

curl (OpenAI-compatible)

curl https://modelsok.com/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model": "gemini-3.5-flash", "messages": [{"role": "user", "content": "Hello"}]}'

Python (openai SDK)

from openai import OpenAI

client = OpenAI(base_url="https://modelsok.com/v1", api_key="YOUR_API_KEY")
response = client.chat.completions.create(
    model="gemini-3.5-flash",
    messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)

Python (anthropic SDK)

from anthropic import Anthropic

client = Anthropic(base_url="https://modelsok.com", api_key="YOUR_API_KEY")
message = client.messages.create(
    model="gemini-3.5-flash",
    max_tokens=1024,
    messages=[{"role": "user", "content": "Hello"}],
)
print(message.content[0].text)

Frequently asked questions

Do I need to change my code to use gemini-3.5-flash here?

No. The gateway speaks the OpenAI and Anthropic wire formats. Change base_url to https://modelsok.com and the API key; requests, streaming and tool calls stay the same.

How is gemini-3.5-flash billed?

Usage is billed per token at the base price shown above ($1.5 per 1M tokens), deducted from a prepaid balance. Group-specific rates and cached-token discounts are applied on the live pricing page.

Is gemini-3.5-flash available right now?

The status page publishes real-traffic availability, first-token latency and throughput for every model over the last 90 days, refreshed every minute.

Back to AI model API pricing