google

Gemini 3.6 Flash Lite

Gemini 3.6 Flash Lite는 100만 토큰의 context window와 초당 350토큰의 처리량을 갖춘 구글의 고효율 모델로, agentic 워크플로우에 최적화되어 있습니다.

고효율긴ContextAgenticAIGoogleGemini
google logogoogleGemini2026년 7월 21일
컨텍스트
1.0M토큰
최대 출력
64K토큰
입력 가격
$0.30/ 1M
출력 가격
$2.50/ 1M
모달리티:TextImageAudioVideo
기능:비전도구스트리밍
벤치마크
GPQA
39%
GPQA: 대학원 수준 과학 Q&A. 생물학, 물리학, 화학 분야의 448개 객관식 문제로 구성된 엄격한 벤치마크. 박사 전문가도 65-74%의 정확도만 달성합니다. Gemini 3.6 Flash Lite이 이 벤치마크에서 39%점을 기록했습니다.
HLE
18%
HLE: 고급 전문 추론. 전문 분야에서 전문가 수준의 추론을 보여주는 모델의 능력을 테스트합니다. Gemini 3.6 Flash Lite이 이 벤치마크에서 18%점을 기록했습니다.
MMLU
84%
MMLU: 대규모 다중 작업 언어 이해. 57개 학술 과목에 걸쳐 16,000개의 객관식 문제로 구성된 종합 벤치마크. Gemini 3.6 Flash Lite이 이 벤치마크에서 84%점을 기록했습니다.
IFEval
89%
IFEval: 지시 따르기 평가. 모델이 특정 지시와 제약 조건을 얼마나 잘 따르는지 측정합니다. Gemini 3.6 Flash Lite이 이 벤치마크에서 89%점을 기록했습니다.
MATH
73%
MATH: 수학 문제 해결. 대수, 기하, 미적분 등의 분야를 테스트하는 종합 수학 벤치마크. Gemini 3.6 Flash Lite이 이 벤치마크에서 73%점을 기록했습니다.
GSM8k
94%
GSM8k: 초등학교 수학 8K. 다단계 추론이 필요한 8,500개의 초등학교 수준 수학 문장제. Gemini 3.6 Flash Lite이 이 벤치마크에서 94%점을 기록했습니다.
SWE-Bench
54%
SWE-Bench: 소프트웨어 엔지니어링 벤치마크. AI 모델이 오픈소스 Python 프로젝트의 실제 GitHub 이슈를 해결하려고 시도합니다. Gemini 3.6 Flash Lite이 이 벤치마크에서 54%점을 기록했습니다.
HumanEval
85%
HumanEval: Python 프로그래밍 문제. 모델이 올바른 Python 함수 구현을 생성해야 하는 164개의 수작업 프로그래밍 문제. Gemini 3.6 Flash Lite이 이 벤치마크에서 85%점을 기록했습니다.
Terminal-Bench
54%
Terminal-Bench: 터미널/CLI 작업. 명령줄 작업을 수행하고 셸 스크립트를 작성하는 능력을 테스트합니다. Gemini 3.6 Flash Lite이 이 벤치마크에서 54%점을 기록했습니다.

Gemini 3.6 Flash Lite 소개

Gemini 3.6 Flash Lite의 기능, 특징 및 더 나은 결과를 얻는 방법에 대해 알아보세요.

고속 Agentic 워크플로우

Gemini 3.6 Flash Lite는 처리량과 비용이 주요 제약 사항인 대규모 저지연 작업에 초점을 맞추고 있습니다. 공식적으로 Gemini 3.5 Flash-Lite로 명명된 이 모델은 더 강력한 Gemini 3.6 Flash의 최적화된 파트너 역할을 합니다. 이 모델은 1초 미만의 응답 시간을 달성하고 초당 350개의 출력 토큰을 유지합니다. 특히 백그라운드 작업, 실시간 검색 합성, 대규모 데이터 처리를 위해 설계되었습니다.

방대한 Context와 도구 지원

Lite라는 명칭에도 불구하고 이 모델은 플래그십급인 100만 토큰의 context window를 유지합니다. 이를 통해 개발자는 복잡한 RAG 파이프라인 없이도 전체 코드 베이스나 방대한 문서 아카이브를 처리할 수 있습니다. 또한 Computer Use를 기본적으로 지원하여 자동화된 UI 상호 작용 및 웹 탐색 도구를 사용할 수 있습니다. 학습 데이터에는 2026년 3월까지의 정보가 포함되어 있어 초기 버전보다 최신 소프트웨어 버전에 대한 이해도가 높습니다.

확장성 및 통합

구글은 agentic 루프에서 1초 미만의 성능을 내기 위해 이 아키텍처를 구축했습니다. 표준 Flash 모델보다 약 80% 저렴하여 매월 수백만 건의 상호 작용이 필요한 프로젝트에 적합합니다. 표준 Gemini API와 통합되어 텍스트, 이미지, 오디오, 비디오를 포함한 다양한 입력을 지원합니다.

Gemini 3.6 Flash Lite

Gemini 3.6 Flash Lite 사용 사례

Gemini 3.6 Flash Lite을 사용하여 훌륭한 결과를 얻는 다양한 방법을 발견하세요.

대용량 문서 감사

100만 context window를 사용하여 수천 페이지의 기업 보고서를 동시에 처리하고 구조화된 데이터를 추출합니다.

Agentic 검색 및 검색 서비스

높은 처리량 덕분에 수십 개의 웹 소스에서 정보를 몇 초 만에 종합하는 실시간 검색 에이전트를 구동합니다.

챗봇 페르소나 유지

몰입형 RPG를 위해 수백만 토큰에 걸쳐 일관된 캐릭터 특성과 대화 기록을 유지합니다.

Computer Use 자동화

기본 제공되는 Computer Use 도구를 사용하여 데이터 입력 및 앱 탐색과 같은 반복적인 UI 작업을 자동화합니다.

대규모 데이터 태깅

매우 낮은 latency로 수백만 개의 사용자 생성 콘텐츠나 고객 지원 티켓을 분류합니다.

경량 코드 생성

신속한 애플리케이션 개발을 위해 운영 수준의 React 컴포넌트나 SQL 쿼리를 생성합니다.

강점

제한

업계 최고 수준의 처리량: 초당 350개의 출력 토큰을 제공하여 Gemini 3.x 클래스 중 가장 빠른 모델입니다.
Reasoning 깊이: 더 큰 Pro 모델이나 표준 Flash 모델에 비해 복잡한 논리 및 다단계 reasoning에서 어려움을 겪습니다.
극강의 비용 효율성: 입력 토큰 100만 개당 0.30달러로, 표준 Gemini 3.6 Flash 모델보다 약 80% 저렴합니다.
출력 버퍼: 출력 토큰이 64k로 제한되어 있어 매우 긴 콘텐츠 생성에는 사용이 제한됩니다.
대규모 1M context: 기존 RAG 시스템의 파편화 문제 없이 대규모 데이터 세트의 전체 context 처리가 가능합니다.
지시사항 준수: 일부 사용자들은 이전 Pro 버전에 비해 복잡한 포맷팅을 유지하는 능력이 다소 떨어졌다고 보고합니다.
최신 knowledge cutoff: 2026년 3월까지의 cutoff를 통해 최근 사건과 최신 소프트웨어 개발 트렌드를 파악하고 있습니다.
공간 인식 비전: multimodal 모델이긴 하지만, 이미지 내 세밀한 객체 감지 정밀도는 3.6 Flash 모델보다 다소 낮습니다.

API 빠른 시작

google/gemini-3.5-flash-lite

문서 보기
google SDK
import { GoogleGenAI } from "@google/genai";

const ai = new GoogleGenAI({
  apiKey: process.env.GEMINI_API_KEY
});

async function main() {
  const interaction = await ai.interactions.create({
    model: "gemini-3.5-flash-lite",
    input: "이 계약 녹취록에서 모든 주요 날짜를 추출해줘.",
    system_instruction: "JSON 형식으로만 출력하세요."
  });
  console.log(interaction.outputText);
}

main();

SDK를 설치하고 몇 분 안에 API 호출을 시작하세요.

Gemini 3.6 Flash Lite에 대한 사람들의 의견

커뮤니티가 Gemini 3.6 Flash Lite에 대해 어떻게 생각하는지 확인하세요

3.6 Flash는 에이전트 루프와 긴 형식의 분석에 가장 빠르고, Flash-Lite는 문서 작업에서 놀라울 정도로 잘 따라옵니다.
singularity_user
reddit
Gemini 3.5 Flash-Lite는 빠르고 저렴하며 지치지 않습니다. 배치 처리를 위한 반복적인 작업에 완벽합니다.
WORLD3_AI
twitter
artificialanalysis에 벤치마크가 올라왔는데... Open Router 기준 1달러/5달러 대비 0.09달러/0.18달러입니다.
hn_reader_99
hackernews
Lite 모델이 더 빠르고 훨씬 저렴하지만, 깊은 사고가 필요한 작업에서는 형편없었습니다.
AI Coding Daily
youtube
백그라운드 작업 속도가 엄청납니다. 대부분의 UI 상호 작용에서 거의 1초 미만의 latency를 보여줍니다.
DevFlowX
twitter
상위 모델과 똑같이 100만 토큰을 처리하면서도 비용은 훨씬 저렴합니다.
United Top Tech
youtube

Gemini 3.6 Flash Lite에 대한 동영상

Gemini 3.6 Flash Lite에 대한 튜토리얼, 리뷰 및 토론 시청

구글이 지금 모델 3개를 출시했습니다... 3.5 Flash lite... 가격이 정말 저렴합니다.

더 빠른 속도와 낮은 비용으로 실무 작업에 최적화되었습니다.

기본적인 prompt에서의 성능이 매우 즉각적입니다.

Flash-Lite는 예산이 제한적인 개발자들에게 최적의 선택지인 것 같습니다.

더 큰 상위 모델들과 똑같이 100만 토큰을 처리합니다.

Lite 모델이 더 빠르고 훨씬 저렴합니다... Grok 4.5와 같은 가격입니다.

React 프로젝트에서 5점 만점에 5점을 받았고, 컴포넌트 생성은 이제 기본입니다.

매우 복잡한 논리를 주었을 때 약간의 환각 현상이 보였습니다.

단순 UI 컴포넌트의 경우 Pro 모델만큼 좋습니다.

latency는 실시간 코딩 보조 도구에서 가장 큰 강점입니다.

이 모델은 기본적으로 더 빠를 뿐입니다... 이전 버전에 비해 토큰 효율성이 더 높은 모델입니다.

더 높은 품질의 결과물을 내기 위해 출력 토큰을 약 17% 적게 사용합니다.

Flash-Lite는 백그라운드 작업을 처리하는 든든한 일꾼입니다.

비전 성능은 약간 떨어지지만 OCR 용도로는 충분합니다.

대규모 자동화를 향한 큰 도약입니다.

단순한 프롬프트 이상

워크플로를 강화하세요 AI 자동화

Automatio는 AI 에이전트, 웹 자동화 및 스마트 통합의 힘을 결합하여 더 짧은 시간에 더 많은 것을 달성할 수 있도록 도와줍니다.

AI 에이전트
웹 자동화
스마트 워크플로

Gemini 3.6 Flash Lite 프로 팁

Gemini 3.6 Flash Lite을 최대한 활용하기 위한 전문가 팁.

Thinking Level 설정

단순 분류 작업 시 속도를 극대화하려면 thinking_level을 최소로 설정하고, tool-calling이 필요할 때는 높이십시오.

샘플링 파라미터 제외

3.x API는 temperature와 top_p를 무시하므로, 출력의 결정론적 결과를 제어하려면 명시적인 system instruction을 사용하십시오.

1M context 활용

파편화된 RAG 시스템보다 나은 결과를 얻으려면 관련 지식 베이스 전체를 prompt에 직접 입력하십시오.

System Instructions 활용

JSON 출력에서 서론이 포함되는 것을 방지하려면 모델의 턴을 미리 채우는 대신 system_instruction을 사용하십시오.

후기

사용자 후기

워크플로를 혁신한 수천 명의 만족한 사용자와 함께하세요

Jonathan Kogan

Jonathan Kogan

Co-Founder/CEO, rpatools.io

Automatio is one of the most used for RPA Tools both internally and externally. It saves us countless hours of work and we realized this could do the same for other startups and so we choose Automatio for most of our automation needs.

Mohammed Ibrahim

Mohammed Ibrahim

CEO, qannas.pro

I have used many tools over the past 5 years, Automatio is the Jack of All trades.. !! it could be your scraping bot in the morning and then it becomes your VA by the noon and in the evening it does your automations.. its amazing!

Ben Bressington

Ben Bressington

CTO, AiChatSolutions

Automatio is fantastic and simple to use to extract data from any website. This allowed me to replace a developer and do tasks myself as they only take a few minutes to setup and forget about it. Automatio is a game changer!

Sarah Chen

Sarah Chen

Head of Growth, ScaleUp Labs

We've tried dozens of automation tools, but Automatio stands out for its flexibility and ease of use. Our team productivity increased by 40% within the first month of adoption.

David Park

David Park

Founder, DataDriven.io

The AI-powered features in Automatio are incredible. It understands context and adapts to changes in websites automatically. No more broken scrapers!

Emily Rodriguez

Emily Rodriguez

Marketing Director, GrowthMetrics

Automatio transformed our lead generation process. What used to take our team days now happens automatically in minutes. The ROI is incredible.

Jonathan Kogan

Jonathan Kogan

Co-Founder/CEO, rpatools.io

Automatio is one of the most used for RPA Tools both internally and externally. It saves us countless hours of work and we realized this could do the same for other startups and so we choose Automatio for most of our automation needs.

Mohammed Ibrahim

Mohammed Ibrahim

CEO, qannas.pro

I have used many tools over the past 5 years, Automatio is the Jack of All trades.. !! it could be your scraping bot in the morning and then it becomes your VA by the noon and in the evening it does your automations.. its amazing!

Ben Bressington

Ben Bressington

CTO, AiChatSolutions

Automatio is fantastic and simple to use to extract data from any website. This allowed me to replace a developer and do tasks myself as they only take a few minutes to setup and forget about it. Automatio is a game changer!

Sarah Chen

Sarah Chen

Head of Growth, ScaleUp Labs

We've tried dozens of automation tools, but Automatio stands out for its flexibility and ease of use. Our team productivity increased by 40% within the first month of adoption.

David Park

David Park

Founder, DataDriven.io

The AI-powered features in Automatio are incredible. It understands context and adapts to changes in websites automatically. No more broken scrapers!

Emily Rodriguez

Emily Rodriguez

Marketing Director, GrowthMetrics

Automatio transformed our lead generation process. What used to take our team days now happens automatically in minutes. The ROI is incredible.

관련 AI Models

openai

GPT-4o mini

OpenAI

OpenAI's most cost-efficient small model, GPT-4o mini offers multimodal intelligence and high-speed performance at a significantly lower price point.

128K context
$0.15/$0.60/1M
alibaba

Qwen3-Coder-Next

alibaba

Qwen3-Coder-Next is Alibaba Cloud's elite Apache 2.0 coding model, featuring an 80B MoE architecture and 256k context window for advanced local development.

262K context
$0.12/$0.75/1M
zhipu

GLM-4.7

Zhipu (GLM)

GLM-4.7 by Zhipu AI is a flagship 358B MoE model featuring a 200K context window, elite 73.8% SWE-bench performance, and native Deep Thinking for agentic...

200K context
$0.60/$2.20/1M
google

Gemini 3.6 Flash

Google

Gemini 3.6 Flash is Google's high-speed model featuring a 17% reduction in token consumption, $1.50/M input pricing, and advanced 3D visualization.

1M context
$1.50/$7.50/1M
minimax

MiniMax M2.5

minimax

MiniMax M2.5 is a SOTA MoE model featuring a 1M context window and elite agentic coding capabilities at disruptive pricing for autonomous agents.

1M context
$0.15/$1.20/1M
other

MiMo V2.5 Pro

Other

MiMo V2.5 Pro is Xiaomi's open-source 1.02T parameter MoE model featuring a 1M context window, native multimodality, and elite agentic coding performance.

1M context
$1.00/$3.00/1M
moonshot

Kimi K3

Moonshot

Kimi K3 is Moonshot AI's 2.8T MoE model with a 1M token context window, native multimodal vision, and frontier-tier coding performance for complex agents.

1M context
$3.00/$15.00/1M
zhipu

GLM-5.2

Zhipu (GLM)

GLM-5.2 is Zhipu AI's flagship open-weight model featuring a 1M context window and specialized agentic coding capabilities under an MIT license.

1M context
$1.40/$4.40/1M

Gemini 3.6 Flash Lite에 대한 자주 묻는 질문

Gemini 3.6 Flash Lite에 대한 일반적인 질문에 대한 답변 찾기