google

Gemini 3.6 Flash Lite

Gemini 3.6 Flash Lite 是 Google 推出的高效率 model,具备 1M token context window,且提供每秒 350 个 tokens 的 throughput,专为 agentic 工作流设计。

高效率长上下文AgenticAIGoogleGemini
google logogoogleGemini2026年7月21日
上下文
1.0Mtokens
最大输出
64Ktokens
输入价格
$0.30/ 1M
输出价格
$2.50/ 1M
模态:TextImageAudioVideo
能力:视觉工具流式传输
基准测试
GPQA
39%
GPQA: 研究生级科学问答. 由领域专家创建的448道多选题的严格基准测试,涵盖生物学、物理学和化学。博士专家仅达到65-74%的准确率。 Gemini 3.6 Flash Lite 在此基准测试中得分 39%。
HLE
18%
HLE: 高级专业推理. 测试模型在专业领域展示专家级推理能力的能力。 Gemini 3.6 Flash Lite 在此基准测试中得分 18%。
MMLU
84%
MMLU: 大规模多任务语言理解. 涵盖57个学科的16,000道多选题的综合基准测试。 Gemini 3.6 Flash Lite 在此基准测试中得分 84%。
IFEval
89%
IFEval: 指令遵循评估. 衡量模型遵循特定指令和约束的能力。 Gemini 3.6 Flash Lite 在此基准测试中得分 89%。
MATH
73%
MATH: 数学问题解决. 涵盖代数、几何、微积分等领域的综合数学基准测试。 Gemini 3.6 Flash Lite 在此基准测试中得分 73%。
GSM8k
94%
GSM8k: 小学数学8K. 8,500道需要多步推理的小学水平数学应用题。 Gemini 3.6 Flash Lite 在此基准测试中得分 94%。
SWE-Bench
54%
SWE-Bench: 软件工程基准. AI模型尝试解决开源Python项目中的真实GitHub问题。 Gemini 3.6 Flash Lite 在此基准测试中得分 54%。
HumanEval
85%
HumanEval: Python编程问题. 164道手写编程问题,模型必须生成正确的Python函数实现。 Gemini 3.6 Flash Lite 在此基准测试中得分 85%。
Terminal-Bench
54%
Terminal-Bench: 终端/CLI任务. 测试执行命令行操作和编写shell脚本的能力。 Gemini 3.6 Flash Lite 在此基准测试中得分 54%。

关于 Gemini 3.6 Flash Lite

了解 Gemini 3.6 Flash Lite 的功能、特性以及它如何帮助您获得更好的效果。

高速 Agentic 工作流

Gemini 3.6 Flash Lite 旨在应对吞吐量和成本为主要约束的高频、低 latency 任务。它在官方文档中被称为 Gemini 3.5 Flash-Lite,是性能更强大的 Gemini 3.6 Flash 的优化版搭档。该 model 可实现亚秒级响应,并保持每秒 350 个 tokens 的持续输出。它专为后台任务、实时搜索综合以及大规模数据处理而设计。

海量上下文与工具支持

尽管名为 "lite",但该 model 依然保留了旗舰级的 100 万 token context window。这使得开发者无需复杂的 RAG 流水线即可处理整个代码库或海量文档存档。它原生支持 Computer Use,能够驱动自动化的 UI 交互和 Web 导航工具。其训练数据包含截至 2026 年 3 月的信息,相比之前版本能更好地识别近期的软件版本。

可扩展性与集成

Google 构建该架构是为了在 agentic 循环中实现亚秒级性能。它的成本比标准 Flash model 低约 80%,使得每月需要数百万次交互的项目能够更轻松地落地。它与标准的 Gemini API 集成,并支持包括文本、图像、音频和视频在内的 multimodal 输入。

Gemini 3.6 Flash Lite

Gemini 3.6 Flash Lite 的使用案例

发现使用 Gemini 3.6 Flash Lite 获得出色效果的不同方式。

大规模文档审计

利用 1M context window 同时处理数千页的企业申报文件,提取结构化数据。

Agentic 搜索与检索

凭借高 throughput,驱动实时搜索 agent 在几秒钟内综合数十个 Web 源的信息。

聊天机器人人格持久化

在数百万 token 的长度内保持一致的角色特征和对话历史,实现沉浸式 RPG 体验。

计算机使用自动化

使用原生的 Computer Use 工具自动完成数据录入和应用导航等重复性 UI 任务。

大规模数据标记

以极低的 latency 对数百万条用户生成的内容或支持工单进行分类。

轻量级代码生成

为快速应用开发生成可用于生产环境的 React 组件或 SQL 查询。

优势

局限性

行业领先的 Throughput: 该 model 每秒可输出 350 个 tokens,是 Gemini 3.x 系列中速度最快的 model。
Reasoning 深度: 与更大规模的 Pro 或标准 Flash model 相比,它在处理复杂逻辑和多步 reasoning 时表现较弱。
极致的成本效益: 每 100 万个输入 token 仅需 0.30 美元,比标准的 Gemini 3.6 Flash model 便宜约 80%。
输出缓冲区: 该 model 的输出被限制在 64k tokens 以内,这限制了其生成超长内容的能力。
海量 1M Context: 无需传统 RAG 系统中常见的碎片化处理,即可实现对大型数据集的全上下文处理。
指令遵循能力: 一些用户反馈,与早期的 Pro 版本相比,它在维护复杂格式方面偶尔会出现回归。
较新的知识截止日期: 2026 年 3 月的知识库截止点确保 model 对近期的事件和最新的软件开发趋势有充分了解。
空间视觉: 虽然它是 multimodal 的,但在图像的细粒度目标检测方面,其精确度明显不如 3.6 Flash model。

API快速入门

google/gemini-3.5-flash-lite

查看文档
google SDK
import { GoogleGenAI } from "@google/genai";

const ai = new GoogleGenAI({
  apiKey: process.env.GEMINI_API_KEY
});

async function main() {
  const interaction = await ai.interactions.create({
    model: "gemini-3.5-flash-lite",
    input: "Extract all key dates from this contract transcript.",
    system_instruction: "Output JSON only."
  });
  console.log(interaction.outputText);
}

main();

安装SDK并在几分钟内开始进行API调用。

人们对 Gemini 3.6 Flash Lite 的评价

看看社区对 Gemini 3.6 Flash Lite 的看法

3.6 Flash 在 agent 循环和长篇分析方面显然是最快的,而 Flash-Lite 在文档处理任务上表现出乎意料地接近。
singularity_user
reddit
Gemini 3.5 Flash-Lite 快速、廉价且不知疲倦。非常适合进行那种枯燥乏味的批量处理工作。
WORLD3_AI
twitter
artificialanalysis 上的基准测试已经上线...在 Open Router 上价格为 $0.09/$0.18,而对比之下其他产品为 $1/$5
hn_reader_99
hackernews
Lite 版本速度更快、成本更低,但在需要深度思考的任务上表现糟糕。
AI Coding Daily
youtube
它的后台任务处理速度快得惊人,对于大多数 UI 交互来说,几乎是亚秒级的 latency。
DevFlowX
twitter
它处理 100 万个 tokens 的能力与更昂贵的同类产品一样,但成本只是其中的一小部分。
United Top Tech
youtube

关于 Gemini 3.6 Flash Lite 的视频

观看关于 Gemini 3.6 Flash Lite 的教程、评测和讨论

Google 目前一口气发布了三个 model... 3.5 Flash lite... 定价非常便宜。

它专为实际任务优化,速度更快,成本更低。

在处理基础 prompt 时表现非常迅捷。

对于预算有限的开发者来说,Flash-Lite 似乎是一个绝佳的选择。

它能像更强大的同系列产品一样处理 100 万个 tokens。

Lite 版本速度更快、成本更低...与 Grok 4.5 价格相同。

在 React 项目上获得了 5 分(满分),组件生成现在已是标配。

当我给它非常复杂的逻辑任务时,我注意到了一些幻觉。

对于简单的 UI 组件,它与 Pro model 表现一样好。

对于实时编码助手来说,latency 是它最大的卖点。

这个 model 基本就是速度更快...相比之前版本,它的 token 利用效率更高。

它比以前减少了约 17% 的输出 tokens,却能提供更高质量的工作成果。

Flash-Lite 是处理后台任务的小型主力。

视觉能力略有下降,但对于 OCR 来说是可以接受的。

这是大规模自动化进程中的一大进步。

不仅仅是提示词

用以下方式提升您的工作流程 AI自动化

Automatio结合AI代理、网页自动化和智能集成的力量,帮助您在更短的时间内完成更多工作。

AI代理
网页自动化
智能工作流

Gemini 3.6 Flash Lite专业提示

专家提示助您充分利用Gemini 3.6 Flash Lite。

设置思考级别

对于简单的分类任务,将 thinking_level 设置为 minimal 以最大限度提高速度;对于涉及 tool-calling 的任务,则提高该设置。

移除采样参数

3.x API 会忽略 temperature 和 top_p 参数,因此请使用明确的系统指令来控制输出的确定性。

利用 1M Context

直接将完整的相关知识库输入到 prompt 中,这比碎片化的 RAG 系统效果更好。

使用系统指令

使用 system_instruction 而不是预填充 model 的对话轮次,以防止 JSON 输出中出现前言。

用户评价

用户怎么说

加入数千名已改变工作流程的满意用户

Jonathan Kogan

Jonathan Kogan

Co-Founder/CEO, rpatools.io

Automatio is one of the most used for RPA Tools both internally and externally. It saves us countless hours of work and we realized this could do the same for other startups and so we choose Automatio for most of our automation needs.

Mohammed Ibrahim

Mohammed Ibrahim

CEO, qannas.pro

I have used many tools over the past 5 years, Automatio is the Jack of All trades.. !! it could be your scraping bot in the morning and then it becomes your VA by the noon and in the evening it does your automations.. its amazing!

Ben Bressington

Ben Bressington

CTO, AiChatSolutions

Automatio is fantastic and simple to use to extract data from any website. This allowed me to replace a developer and do tasks myself as they only take a few minutes to setup and forget about it. Automatio is a game changer!

Sarah Chen

Sarah Chen

Head of Growth, ScaleUp Labs

We've tried dozens of automation tools, but Automatio stands out for its flexibility and ease of use. Our team productivity increased by 40% within the first month of adoption.

David Park

David Park

Founder, DataDriven.io

The AI-powered features in Automatio are incredible. It understands context and adapts to changes in websites automatically. No more broken scrapers!

Emily Rodriguez

Emily Rodriguez

Marketing Director, GrowthMetrics

Automatio transformed our lead generation process. What used to take our team days now happens automatically in minutes. The ROI is incredible.

Jonathan Kogan

Jonathan Kogan

Co-Founder/CEO, rpatools.io

Automatio is one of the most used for RPA Tools both internally and externally. It saves us countless hours of work and we realized this could do the same for other startups and so we choose Automatio for most of our automation needs.

Mohammed Ibrahim

Mohammed Ibrahim

CEO, qannas.pro

I have used many tools over the past 5 years, Automatio is the Jack of All trades.. !! it could be your scraping bot in the morning and then it becomes your VA by the noon and in the evening it does your automations.. its amazing!

Ben Bressington

Ben Bressington

CTO, AiChatSolutions

Automatio is fantastic and simple to use to extract data from any website. This allowed me to replace a developer and do tasks myself as they only take a few minutes to setup and forget about it. Automatio is a game changer!

Sarah Chen

Sarah Chen

Head of Growth, ScaleUp Labs

We've tried dozens of automation tools, but Automatio stands out for its flexibility and ease of use. Our team productivity increased by 40% within the first month of adoption.

David Park

David Park

Founder, DataDriven.io

The AI-powered features in Automatio are incredible. It understands context and adapts to changes in websites automatically. No more broken scrapers!

Emily Rodriguez

Emily Rodriguez

Marketing Director, GrowthMetrics

Automatio transformed our lead generation process. What used to take our team days now happens automatically in minutes. The ROI is incredible.

相关 AI Models

openai

GPT-4o mini

OpenAI

OpenAI's most cost-efficient small model, GPT-4o mini offers multimodal intelligence and high-speed performance at a significantly lower price point.

128K context
$0.15/$0.60/1M
alibaba

Qwen3-Coder-Next

alibaba

Qwen3-Coder-Next is Alibaba Cloud's elite Apache 2.0 coding model, featuring an 80B MoE architecture and 256k context window for advanced local development.

262K context
$0.12/$0.75/1M
zhipu

GLM-4.7

Zhipu (GLM)

GLM-4.7 by Zhipu AI is a flagship 358B MoE model featuring a 200K context window, elite 73.8% SWE-bench performance, and native Deep Thinking for agentic...

200K context
$0.60/$2.20/1M
google

Gemini 3.6 Flash

Google

Gemini 3.6 Flash is Google's high-speed model featuring a 17% reduction in token consumption, $1.50/M input pricing, and advanced 3D visualization.

1M context
$1.50/$7.50/1M
minimax

MiniMax M2.5

minimax

MiniMax M2.5 is a SOTA MoE model featuring a 1M context window and elite agentic coding capabilities at disruptive pricing for autonomous agents.

1M context
$0.15/$1.20/1M
other

MiMo V2.5 Pro

Other

MiMo V2.5 Pro is Xiaomi's open-source 1.02T parameter MoE model featuring a 1M context window, native multimodality, and elite agentic coding performance.

1M context
$1.00/$3.00/1M
moonshot

Kimi K3

Moonshot

Kimi K3 is Moonshot AI's 2.8T MoE model with a 1M token context window, native multimodal vision, and frontier-tier coding performance for complex agents.

1M context
$3.00/$15.00/1M
zhipu

GLM-5.2

Zhipu (GLM)

GLM-5.2 is Zhipu AI's flagship open-weight model featuring a 1M context window and specialized agentic coding capabilities under an MIT license.

1M context
$1.40/$4.40/1M

关于Gemini 3.6 Flash Lite的常见问题

查找关于Gemini 3.6 Flash Lite的常见问题答案