MODELS
模型目录
搜索、按类型筛选,查看已发布模型和价格。
- text
Gemini Flash
AVAILABLEgoogle/gemini-flashgoogle$1.00/M
出 $2.00/M
上下文 —输入 $1.00/M输出 $2.00/M流式温度caps.messagescaps.modelcaps.system函数视觉 - text
Google: Gemini 3.7 Flash
AVAILABLEgoogle/gemini-3.7-flashgoogle$1.50/M
出 $7.50/M
上下文 1M最大输出 66K输入 $1.50/M输出 $7.50/MGemini 3.7 Flash is Google's multimodal workhorse model for fast agentic workflows, coding, and complex multi-step reasoning, improving on 3.6 Flash in agentic benchmarks (GPQA Diamond ~94%, TAU-Bench ~80%). All input modalities (text/image/video/audio) unified pricing; batch tier at 50% discount. Thinking model with configurable levels. 1M context. Released August 13, 2026.
温度caps.top_pcaps.stop函数JSON推理 - text
Google: Gemini 3.5 Flash Lite
AVAILABLEgoogle/gemini-3.5-flash-litegoogle$0.300/M
出 $2.50/M
上下文 1M最大输出 66K输入 $0.300/M输出 $2.50/MGemini 3.5 Flash-Lite is Google's high-throughput, low-latency multimodal model with upgraded agentic capabilities, suited for subagents executing focused tasks within complex multi-agent workflows, agentic search, and document processing. All input modalities (text/image/video/audio) unified at $0.30/1M; batch tier at 50% discount. Successor to Gemini 3.1 Flash Lite. Released July 21, 2026.
温度caps.top_pcaps.stop函数JSON推理 - text
Google: Gemini 3.6 Flash
AVAILABLEgoogle/gemini-3.6-flashgoogle$1.50/M
出 $7.50/M
上下文 1M最大输出 66K输入 $1.50/M输出 $7.50/MGemini 3.6 Flash is Google's token-efficient workhorse model: 17% fewer output tokens than 3.5 Flash, fewer reasoning steps and tool calls in multi-step agentic workflows, and higher-precision code edits with reduced execution loops. All input modalities (text/image/video/audio) unified at $1.50/1M; batch tier at 50% discount. Thinking model with configurable levels. Released July 21, 2026.
温度caps.top_pcaps.stop函数JSON推理 - image
Google: Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)
AVAILABLEgoogle/gemini-3.1-flash-lite-imagegoogle$30.00/M
图像
上下文 66K最大输出 64KGemini 3.1 Flash Lite Image is Google's most cost-efficient image generation and editing model, supporting text-to-image, image editing, and multi-image composition. Outputs generated at 1K resolution across 14 aspect ratios, at the lowest price point in the Nano Banana family.
温度caps.top_pcaps.stopJSON - image
Google: Nano Banana Pro (Gemini 3 Pro Image)
AVAILABLEgoogle/gemini-3-pro-imagegoogle$120.00/M
图像
上下文 66K最大输出 32KNano Banana Pro is Google's most advanced image-generation and editing model, built on Gemini 3 Pro. GA release of gemini-3-pro-image-preview. It generates context-rich graphics from infographics and diagrams to cinematic composites, with 2K/4K output, multi-image blending, identity preservation, and localized edits.
温度caps.top_pcaps.stopJSON - image
Google: Nano Banana 2 (Gemini 3.1 Flash Image)
AVAILABLEgoogle/gemini-3.1-flash-imagegoogle$60.00/M
图像
上下文 131K最大输出 64KGemini 3.1 Flash Image, a.k.a. "Nano Banana 2," is Google's state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. GA release of gemini-3.1-flash-image-preview, combining advanced contextual understanding with fast, cost-efficient inference.
温度caps.top_pcaps.stopJSON - text
Google: Gemini 3.5 Flash
AVAILABLEgoogle/gemini-3.5-flashgoogle$1.50/M
出 $9.00/M
上下文 1M最大输出 66K输入 $1.50/M输出 $9.00/MGemini 3.5 Flash is Google's efficient multimodal model delivering near-Pro level coding and reasoning at Flash-tier cost and speed. Optimized for coding tasks and parallel agent execution with configurable thinking levels. Released May 20, 2026.
温度caps.top_pcaps.stop函数JSON推理 - text
Google: Gemini 3.1 Flash Lite
AVAILABLEgoogle/gemini-3.1-flash-litegoogle$0.250/M
出 $1.50/M
上下文 1M最大输出 64K输入 $0.250/M输出 $1.50/MGemini 3.1 Flash Lite (GA) is Google's high-efficiency multimodal model optimized for low-latency, high-volume workloads. GA version of the preview model. Supports full thinking levels (minimal, low, medium, high) for cost/performance trade-offs. Priced at half the cost of Gemini 3 Flash. Released May 7, 2026.
温度caps.top_pcaps.stop函数JSON推理 - text
Google: Gemini 3.1 Pro Preview
AVAILABLEgoogle/gemini-3.1-pro-previewgoogle$2.00/M
出 $12.00/M
上下文 1.0M最大输出 66K输入 $2.00/M输出 $12.00/MGemini 3.1 Pro is the next generation in the Gemini series of models, a suite of highly-capable, natively multimodal, reasoning models. Gemini 3 Pro is now Google’s most advanced model for complex tasks, and can comprehend vast datasets, challenging problems from different information sources, including text, audio, images, video, and entire code repositories
温度caps.top_pcaps.stop函数JSON推理 - text
Google: Gemini 3 Flash Preview
AVAILABLEgoogle/gemini-3-flash-previewgoogle$0.500/M
出 $3.00/M
上下文 1.0M最大输出 66K输入 $0.500/M输出 $3.00/MGemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool use performance with substantially lower latency than larger Gemini variants, making it well suited for interactive development, long running agent loops, and collaborative coding tasks. Compared to Gemini 2.5 Flash, it provides broad quality improvements across reasoning, multimodal understanding, and reliability.
温度caps.top_pcaps.stop函数JSON推理 - image
Google: Nano Banana (Gemini 2.5 Flash Image)
AVAILABLEgoogle/gemini-2.5-flash-imagegoogle$30.00/M
图像
上下文 32K最大输出 32KGemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation, edits, and multi-turn conversations.
温度caps.top_pcaps.stopJSON - text
Google: Gemini 2.5 Flash Lite
AVAILABLEgoogle/gemini-2.5-flash-litegoogle$0.100/M
出 $0.400/M
上下文 1.0M最大输出 66K输入 $0.100/M输出 $0.400/MGemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance across common benchmarks compared to earlier Flash models. By default, [thinking] (i.e. multi-pass reasoning) is disabled to prioritize speed, but developers can enable it via the Reasoning API parameter to selectively trade off cost for intelligence.
温度caps.top_pcaps.stop函数JSON - text
Google: Gemini 2.5 Flash
AVAILABLEgoogle/gemini-2.5-flashgoogle$0.300/M
出 $2.50/M
上下文 1.0M最大输出 66K输入 $0.300/M输出 $2.50/MFast and cost-effective Gemini 2.5 with 1M context. Features toggleable reasoning capabilities and full multimodal support at significantly lower cost than Pro.
温度caps.top_pcaps.stop函数JSON推理 - text
Google: Gemini 2.5 Pro
AVAILABLEgoogle/gemini-2.5-progoogle$1.25/M
出 $10.00/M
上下文 1.0M最大输出 66K输入 $1.25/M输出 $10.00/MGoogle's flagship Gemini model with 1M context, thinking/reasoning capabilities, and full multimodal support (text, image, video, audio). Excels at complex analysis and creative tasks.
温度caps.top_pcaps.stop函数JSON推理