Gemini

gemini-3-flash

Google按量收費
Alias:gemini-3-flash-preview

Google 的 gemini-3-flash 模型,支持文本、图像、音频、视频、文件输入和文本输出。支持推理、工具调用、结构化输出。

textimageaudiovideofilecachingcode_interpreterfunction_callingtoolsweb_searchstructured_outputjson_modereasoningvisioncontext:1048576
起始價格
輸入 / 輸出 · 1M
上下文
1M
最大輸入窗口
最大輸出
65.5K
單次回應最大 token 數
模態
→
發佈於
Dec 2025

Pricing by Supplier

Google AI Studio
省60%
谷歌官方接口
輸入$0.5$0.2/ 1M
輸出$3$1.2/ 1M
緩存讀取$0.05$0.02/ 1M
Google Vertex
省55%
谷歌官方接口
輸入$0.5$0.225/ 1M
輸出$3$1.35/ 1M
緩存讀取$0.05$0.0225/ 1M
Gemini Cli
省80%
Gemini cli 号池,适合 vibe coding学习等场景
輸入$0.5$0.1/ 1M
輸出$3$0.6/ 1M
緩存讀取$0.05$0.01/ 1M

能力 / 支援的模態

提示詞緩存代碼解釋器函數呼叫工具網絡搜尋結構化輸出JSON 模式推理視覺
輸入
輸出

廠商與數據私隱

供應商
Google文件
分詞器
SentencePiece (Gemini)
許可證
Proprietary (commercial)商業閉源
數據保留73 日預設不會用於上游訓練

效能

About gemini-3-flash

Google 的 gemini-3-flash 模型,支持文本、图像、音频、视频、文件输入和文本输出。支持推理、工具调用、结构化输出。

Use cases and prompting

Starting points for evaluation; supported inputs and options are listed in API access.

Use cases to explore

  • Draft and revise text with explicit audience, tone and format requirements.
  • Summarize supplied documents and compare answers against the original sources.

Practical tips

Separate instructions from source material, describe the desired output and provide an example. Test tool use or structured output only when listed as supported.

Example prompt

Summarize the following document in five bullet points. Separate confirmed facts from open questions, cite the relevant passages and do not invent missing information. Document: [paste your text]

API access

呼叫示例

請求POST/v1beta/models/gemini-3-flash:generateContent
請求範例
參數
參數類型預設值 / 範圍說明資訊
temperature
number
=10 ~ 2
採樣溫度;越低越穩定
top_p
number
=10 ~ 1
核採樣累積概率
max_tokens
integer>= 1回應中最大 token 數
frequency_penalty
number
=0-2 ~ 2
懲罰高頻 token 的重複出現
presence_penalty
number
=0-2 ~ 2
鼓勵引入新話題
stop
array—最多 4 個停止生成的字串
seed
integer—盡量保證可復現的採樣種子
n
integer
=1>= 1
生成的候選條數
stream
boolean
=false
透過 SSE 串流返回 token
response_format
object—強制輸出 JSON 物件或符合 Schema 的結果
tools
array—模型可呼叫的工具 / 函數聲明
tool_choice
string
autononerequired
工具選擇策略或具體工具名
logprobs
boolean
=false
返回每個 token 的對數概率
top_logprobs
integer0 ~ 20每個 token 返回的 top 概率數量
logit_bias
object—按 token 的 logit 偏置映射
user
string—用於風險審計的終端用戶標識

替換 <YOUR_API_KEY> 替換為令牌設定中的 API Key。

身份驗證

所有請求必須攜帶 Authorization: Bearer <TOKEN> 請求頭。Anthropic 格式的端點也接受 x-api-key 請求頭。

在「令牌」頁面生成 API Key,可以按模型、分組、IP、速率等維度精細化授權。

支援的參數

Generation parameters
參數類型預設值 / 範圍說明資訊
temperature
number
=10 ~ 2
採樣溫度;越低越穩定
top_p
number
=10 ~ 1
核採樣累積概率
max_tokens
integer>= 1回應中最大 token 數
frequency_penalty
number
=0-2 ~ 2
懲罰高頻 token 的重複出現
presence_penalty
number
=0-2 ~ 2
鼓勵引入新話題
stop
array—最多 4 個停止生成的字串
seed
integer—盡量保證可復現的採樣種子
n
integer
=1>= 1
生成的候選條數
stream
boolean
=false
透過 SSE 串流返回 token
response_format
object—強制輸出 JSON 物件或符合 Schema 的結果
tools
array—模型可呼叫的工具 / 函數聲明
tool_choice
string
autononerequired
工具選擇策略或具體工具名
logprobs
boolean
=false
返回每個 token 的對數概率
top_logprobs
integer0 ~ 20每個 token 返回的 top 概率數量
logit_bias
object—按 token 的 logit 偏置映射
user
string—用於風險審計的終端用戶標識

速率限制

供應商RPMTPMRPD
Gemini Cli920370K19K
Google AI Studio830334K17K
Google Vertex650260K13K

RPM = 每分鐘請求數,TPM = 每分鐘 token 數,RPD = 每日請求數。限制按令牌分組生效。

Frequently asked questions about gemini-3-flash

What is gemini-3-flash?

Google 的 gemini-3-flash 模型,支持文本、图像、音频、视频、文件输入和文本输出。支持推理、工具调用、结构化输出。

How do I call gemini-3-flash?

Create an API key with access to gemini-3-flash, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is gemini-3-flash priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

What is the context window of gemini-3-flash?

The model catalog lists a context window of 1048576 tokens. Check the selected endpoint for request limits.

What is the maximum output of gemini-3-flash?

The model catalog lists a maximum output of 65536 tokens. Your request settings may set a lower limit.

How should I evaluate gemini-3-flash for my project?

Separate instructions from source material, describe the desired output and provide an example. Test tool use or structured output only when listed as supported.