Gemini 3.5 Flash Lite

Google 最具成本效益的正式版模型,针对高调用量智能体任务、翻译与简单数据处理优化

Model ID: gemini-3.5-flash-lite · Type: chat · Provider: Google

Endpoints: /v1/chat/completions · /v1/messages

Pricing

Input (per 1M tokens)$0.225 USD
Output (per 1M tokens)$1.875 USD
Cache read (per 1M tokens)$0.0225 USD
from openai import OpenAI

client = OpenAI(api_key="sk-...", base_url="https://api.router.ai/v1")
resp = client.chat.completions.create(
    model="gemini-3.5-flash-lite",
    messages=[{"role": "user", "content": "Hello"}],
)
print(resp.choices[0].message.content)

FAQ

How much does Gemini 3.5 Flash Lite cost on OpenCrow?

Gemini 3.5 Flash Lite (`gemini-3.5-flash-lite`) is billed per usage at $0.225/1M in · $1.875/1M out, in USD. Current pricing is always listed at https://opencrow.ai/models/gemini-3.5-flash-lite.

How do I call Gemini 3.5 Flash Lite through OpenCrow?

Send a request to https://api.router.ai/v1/v1/chat/completions with the header `Authorization: Bearer <your API key>` and `"model": "gemini-3.5-flash-lite"`. The API is OpenAI-compatible, so any OpenAI SDK works by changing base_url to https://api.router.ai/v1 — no other code change.

Which endpoints does Gemini 3.5 Flash Lite support?

Gemini 3.5 Flash Lite can be called on: /v1/chat/completions; /v1/messages.

What is Gemini 3.5 Flash Lite's context window?

Gemini 3.5 Flash Lite accepts up to 1,048,576 input tokens and can return up to 65,536 output tokens. Requests exceeding the input limit are rejected before reaching the model.

What can Gemini 3.5 Flash Lite do?

Gemini 3.5 Flash Lite supports: vision, function_calling, prompt_caching.

Who makes Gemini 3.5 Flash Lite?

Gemini 3.5 Flash Lite is a chat model from Google, available through the OpenCrow gateway with the same API key as every other model.

Call it through the OpenCrow OpenAI-compatible endpoint (API base: https://api.router.ai/v1). AI agents can discover and call every model on this gateway through MCP (https://mcp.router.ai/mcp) with no manual integration.

API reference · All models