google

Gemini 3.6 Flash Lite

Gemini 3.6 Flash Lite คือ model ประสิทธิภาพสูงจาก Google ที่มาพร้อม 1M token context window และ throughput 350 tokens/sec สำหรับ agentic workflows

ประสิทธิภาพสูงLongContextAgenticAIGoogleGemini
google logogoogleGemini21 กรกฎาคม 2026
บริบท
1.0Mโทเคน
เอาต์พุตสูงสุด
64Kโทเคน
ราคาอินพุต
$0.30/ 1M
ราคาเอาต์พุต
$2.50/ 1M
โหมด:TextImageAudioVideo
ความสามารถ:การมองเห็นเครื่องมือสตรีมมิ่ง
เกณฑ์มาตรฐาน
GPQA
39%
GPQA: คำถามวิทยาศาสตร์ระดับบัณฑิตศึกษา. เกณฑ์มาตรฐานที่เข้มงวดพร้อม 448 คำถามจากชีววิทยา ฟิสิกส์ และเคมี ผู้เชี่ยวชาญ PhD ทำได้เพียง 65-74% Gemini 3.6 Flash Lite ได้คะแนน 39% ในเกณฑ์มาตรฐานนี้
HLE
18%
HLE: การใช้เหตุผลระดับผู้เชี่ยวชาญ. ทดสอบความสามารถของโมเดลในการแสดงการใช้เหตุผลระดับผู้เชี่ยวชาญในสาขาเฉพาะทาง Gemini 3.6 Flash Lite ได้คะแนน 18% ในเกณฑ์มาตรฐานนี้
MMLU
84%
MMLU: ความเข้าใจภาษาแบบมัลติทาสก์ขนาดใหญ่. เกณฑ์มาตรฐานที่ครอบคลุมพร้อม 16,000 คำถามใน 57 วิชา Gemini 3.6 Flash Lite ได้คะแนน 84% ในเกณฑ์มาตรฐานนี้
IFEval
89%
IFEval: การประเมินการปฏิบัติตามคำสั่ง. วัดว่าโมเดลปฏิบัติตามคำสั่งและข้อจำกัดเฉพาะได้ดีเพียงใด Gemini 3.6 Flash Lite ได้คะแนน 89% ในเกณฑ์มาตรฐานนี้
MATH
73%
MATH: การแก้ปัญหาคณิตศาสตร์. เกณฑ์มาตรฐานคณิตศาสตร์ที่ครอบคลุมทดสอบการแก้ปัญหาในพีชคณิต เรขาคณิต แคลคูลัส Gemini 3.6 Flash Lite ได้คะแนน 73% ในเกณฑ์มาตรฐานนี้
GSM8k
94%
GSM8k: คณิตศาสตร์ประถม 8K. 8,500 โจทย์คณิตศาสตร์ระดับประถมศึกษา Gemini 3.6 Flash Lite ได้คะแนน 94% ในเกณฑ์มาตรฐานนี้
SWE-Bench
54%
SWE-Bench: เกณฑ์มาตรฐานวิศวกรรมซอฟต์แวร์. โมเดล AI พยายามแก้ปัญหา GitHub จริงในโครงการ Python Gemini 3.6 Flash Lite ได้คะแนน 54% ในเกณฑ์มาตรฐานนี้
HumanEval
85%
HumanEval: โจทย์เขียนโปรแกรม Python. 164 โจทย์เขียนโปรแกรมที่โมเดลต้องสร้างการใช้งานฟังก์ชัน Python ที่ถูกต้อง Gemini 3.6 Flash Lite ได้คะแนน 85% ในเกณฑ์มาตรฐานนี้
Terminal-Bench
54%
Terminal-Bench: งาน Terminal/CLI. ทดสอบความสามารถในการดำเนินการ command-line Gemini 3.6 Flash Lite ได้คะแนน 54% ในเกณฑ์มาตรฐานนี้

เกี่ยวกับ Gemini 3.6 Flash Lite

เรียนรู้เกี่ยวกับความสามารถของ Gemini 3.6 Flash Lite คุณสมบัติ และวิธีที่จะช่วยให้คุณได้ผลลัพธ์ที่ดีขึ้น

Agentic Workflows ที่มีความเร็วสูง

Gemini 3.6 Flash Lite มุ่งเน้นไปที่งานที่มีปริมาณข้อมูลมหาศาลและต้องการ latency ต่ำ ซึ่ง throughput และต้นทุนคือข้อจำกัดหลัก โดยมีเอกสารอย่างเป็นทางการระบุว่าเป็น Gemini 3.5 Flash-Lite ทำหน้าที่เป็นพันธมิตรที่ปรับแต่งมาเพื่อเสริมประสิทธิภาพให้กับ Gemini 3.6 Flash ที่ทรงพลังกว่า model นี้ให้เวลาในการตอบสนองต่ำกว่า 1 วินาทีและคงความเร็ว 350 output tokens ต่อวินาทีได้อย่างต่อเนื่อง ออกแบบมาโดยเฉพาะสำหรับงานเบื้องหลัง การสังเคราะห์ผลการค้นหาแบบเรียลไทม์ และการประมวลผลข้อมูลขนาดใหญ่

Context และเครื่องมือที่มหาศาล

แม้จะถูกเรียกว่า Lite แต่ model นี้ยังคงจุดเด่นเรื่อง 1 ล้าน token context window ทำให้เหล่านักพัฒนาสามารถประมวลผลทั้ง codebase หรือคลังเอกสารมหาศาลได้โดยไม่ต้องพึ่งพา pipeline ของ RAG ที่ซับซ้อน นอกจากนี้ยังรองรับ Computer Use โดยกำเนิด ทำให้สามารถโต้ตอบกับ UI และเครื่องมือการนำทางบนเว็บได้โดยอัตโนมัติ ข้อมูลที่ใช้ในการฝึกฝนครอบคลุมถึงเดือนมีนาคม 2026 ทำให้ model มีความเข้าใจเกี่ยวกับซอฟต์แวร์เวอร์ชันล่าสุดได้ดีกว่ารุ่นก่อนหน้า

ความสามารถในการขยายขนาดและการบูรณาการ

Google สร้างสถาปัตยกรรมนี้เพื่อประสิทธิภาพระดับต่ำกว่า 1 วินาทีใน agentic loops มีราคาถูกกว่ารุ่น Flash มาตรฐานประมาณ 80% ทำให้เหมาะสมกับโปรเจกต์ที่ต้องการการโต้ตอบหลายล้านครั้งต่อเดือน สามารถใช้งานร่วมกับ Gemini API มาตรฐานและรองรับ multimodal inputs ได้ทั้งข้อความ รูปภาพ เสียง และวิดีโอ

Gemini 3.6 Flash Lite

กรณีการใช้งานสำหรับ Gemini 3.6 Flash Lite

ค้นพบวิธีต่างๆ ที่คุณสามารถใช้ Gemini 3.6 Flash Lite เพื่อได้ผลลัพธ์ที่ยอดเยี่ยม

การตรวจสอบเอกสารปริมาณมาก

ประมวลผลเอกสารองค์กรหลายพันหน้าพร้อมกันเพื่อสกัดข้อมูลที่มีโครงสร้างโดยใช้ context window ขนาด 1M

Agentic Search & Retrieval

ขับเคลื่อน search agent แบบเรียลไทม์ที่สามารถสังเคราะห์ข้อมูลจากแหล่งที่มาบนเว็บหลายสิบแห่งได้ในไม่กี่วินาทีด้วย throughput ที่สูง

การรักษาบุคลิก Chatbot

คงไว้ซึ่งลักษณะเฉพาะตัวและประวัติการสนทนาที่ยาวนานหลายล้าน tokens สำหรับเกม RPG ที่ต้องการความสมจริง

ระบบอัตโนมัติผ่าน Computer Use

ทำภารกิจบน UI ซ้ำๆ เช่น การป้อนข้อมูลและการนำทางในแอปพลิเคชันโดยใช้เครื่องมือ Computer Use

การจัดหมวดหมู่ข้อมูลขนาดใหญ่

จัดหมวดหมู่เนื้อหาที่ผู้ใช้สร้างขึ้นหรือตั๋วสนับสนุนนับล้านรายการด้วย latency ที่ต่ำมาก

การสร้างโค้ดแบบเบา

สร้าง React components หรือ SQL queries ที่พร้อมใช้งานจริงสำหรับการพัฒนาแอปพลิเคชันที่รวดเร็ว

จุดแข็ง

ข้อจำกัด

Throughput ระดับแนวหน้าของอุตสาหกรรม: model สามารถประมวลผล output ได้ถึง 350 tokens ต่อวินาที ทำให้เป็น model ที่เร็วที่สุดในกลุ่ม Gemini 3.x
ความลึกของการใช้เหตุผล: มีปัญหาด้านตรรกะที่ซับซ้อนและการใช้เหตุผลหลายขั้นตอนเมื่อเทียบกับ model ขนาดใหญ่รุ่น Pro หรือ Flash มาตรฐาน
ความคุ้มค่าสูงสุด: ด้วยราคา $0.30 ต่อล้าน input tokens ทำให้ราคาถูกกว่า Gemini 3.6 Flash รุ่นมาตรฐานประมาณ 80%
Output Buffer: model ถูกจำกัด output ไว้ที่ 64k tokens ซึ่งจำกัดการใช้งานสำหรับการสร้างเนื้อหาที่ยาวมากๆ
Context ขนาดมหาศาล 1M: สามารถประมวลผลชุดข้อมูลขนาดใหญ่ได้เต็มรูปแบบโดยไม่เกิดปัญหาการแยกส่วนข้อมูลเหมือนในระบบ RAG แบบเดิม
การปฏิบัติตามคำสั่ง (Instruction Adherence): ผู้ใช้บางรายรายงานว่ามีการถดถอยเล็กน้อยในการรักษาการจัดรูปแบบที่ซับซ้อนเมื่อเทียบกับรุ่น Pro ก่อนหน้า
Knowledge Cutoff ที่ทันสมัย: ข้อมูลที่ครอบคลุมถึงเดือนมีนาคม 2026 ทำให้ model รับรู้ถึงเหตุการณ์ล่าสุดและแนวโน้มการพัฒนาซอฟต์แวร์ใหม่ๆ
Spatial Vision: แม้จะเป็น multimodal แต่มีความแม่นยำในการตรวจจับวัตถุขนาดเล็กในภาพน้อยกว่าเมื่อเทียบกับ model 3.6 Flash

เริ่มต้นด่วน API

google/gemini-3.5-flash-lite

ดูเอกสาร
google SDK
import { GoogleGenAI } from "@google/genai";

const ai = new GoogleGenAI({
  apiKey: process.env.GEMINI_API_KEY
});

async function main() {
  const interaction = await ai.interactions.create({
    model: "gemini-3.5-flash-lite",
    input: "สกัดวันที่สำคัญทั้งหมดจากเนื้อหาสัญญาฉบับนี้",
    system_instruction: "แสดงผลลัพธ์เป็น JSON เท่านั้น"
  });
  console.log(interaction.outputText);
}

main();

ติดตั้ง SDK และเริ่มเรียก API ภายในไม่กี่นาที

ผู้คนพูดอะไรเกี่ยวกับ Gemini 3.6 Flash Lite

ดูว่าชุมชนคิดอย่างไรเกี่ยวกับ Gemini 3.6 Flash Lite

3.6 Flash เร็วที่สุดสำหรับ agent loops และการวิเคราะห์เนื้อหายาวๆ ในขณะที่ Flash-Lite ก็ทำได้ใกล้เคียงจนน่าประหลาดใจในงานด้านเอกสาร
singularity_user
reddit
Gemini 3.5 Flash-Lite เร็ว ถูก และอึดมาก เหมาะสำหรับงานประมวลผลแบบ batch ที่ไม่ต้องการความหวือหวา
WORLD3_AI
twitter
ผล benchmark บน artificialanalysis ออกมาแล้ว... ราคา $0.09/$0.18 บน Open Router เมื่อเทียบกับ $1/$5
hn_reader_99
hackernews
model Lite เร็วและถูกกว่ามาก แต่มันสอบตกอย่างจังในงานที่ต้องใช้ความคิดเชิงลึก
AI Coding Daily
youtube
ความเร็วในตัวนี้เหลือเชื่อมากสำหรับงานเบื้องหลัง โดยทั่วไปมี latency ต่ำกว่าหนึ่งวินาทีสำหรับการโต้ตอบ UI ส่วนใหญ่
DevFlowX
twitter
มันรองรับ 1 ล้าน tokens ได้เหมือนรุ่นพี่ตัวใหญ่ แต่จ่ายในราคาเพียงเศษเสี้ยว
United Top Tech
youtube

วิดีโอเกี่ยวกับ Gemini 3.6 Flash Lite

ดูบทเรียน รีวิว และการสนทนาเกี่ยวกับ Gemini 3.6 Flash Lite

Google เพิ่งเปิดตัวสาม model พร้อมกัน... 3.5 Flash lite... ราคาถูกมาก

มันถูกปรับแต่งสำหรับงานในโลกจริงด้วยความเร็วที่สูงขึ้นและราคาที่ถูกลง

ประสิทธิภาพในการตอบสนอง prompt พื้นฐานนั้นรวดเร็วมาก

Flash-Lite ดูจะเป็นจุดที่ลงตัวที่สุดสำหรับนักพัฒนาที่มีงบจำกัด

มันรองรับ 1 ล้าน tokens เหมือนรุ่นพี่ตัวใหญ่ๆ

model Lite เร็วและถูกกว่ามาก... ราคาเท่ากับ Grok 4.5

ได้คะแนนเต็ม 5 ในโปรเจกต์ React การสร้าง component ตอนนี้ถือเป็นเรื่องพื้นฐานไปแล้ว

ผมพบอาการหลอนบ้างเล็กน้อยเมื่อให้โจทย์ที่มีตรรกะซับซ้อนมาก

สำหรับ UI components ง่ายๆ มันทำได้ดีพอๆ กับรุ่น Pro เลย

latency คือจุดขายที่สำคัญที่สุดสำหรับผู้ช่วยเขียนโค้ดแบบเรียลไทม์

model นี้โดยพื้นฐานแล้วแค่เร็วขึ้น... มันเป็น model ที่ประหยัด token ได้ดีกว่าเวอร์ชันก่อนๆ

มันใช้ output tokens น้อยลงประมาณ 17% เพื่อให้ได้งานคุณภาพสูงขึ้น

Flash-Lite คือแรงงานตัวเล็กที่คอยขับเคลื่อนงานเบื้องหลัง

ความสามารถด้าน vision ต่ำลงเล็กน้อยแต่ก็เพียงพอสำหรับงาน OCR

มันเป็นก้าวสำคัญสำหรับระบบอัตโนมัติขนาดใหญ่

มากกว่าแค่พรอมต์

เพิ่มพลังให้เวิร์กโฟลว์ของคุณด้วย ระบบอัตโนมัติ AI

Automatio รวมพลังของ AI agents การอัตโนมัติเว็บ และการผสานรวมอัจฉริยะเพื่อช่วยให้คุณทำงานได้มากขึ้นในเวลาน้อยลง

AI Agents
การอัตโนมัติเว็บ
เวิร์กโฟลว์อัจฉริยะ

เคล็ดลับมือโปรสำหรับ Gemini 3.6 Flash Lite

เคล็ดลับจากผู้เชี่ยวชาญเพื่อช่วยให้คุณใช้ประโยชน์สูงสุดจาก Gemini 3.6 Flash Lite และได้ผลลัพธ์ที่ดีขึ้น

การตั้งค่าระดับการใช้ความคิด (Thinking Level)

ควรตั้ง thinking_level ไว้ที่ระดับต่ำสุดสำหรับการจัดหมวดหมู่ข้อมูลทั่วไปเพื่อเพิ่มความเร็วสูงสุด แต่ให้เพิ่มระดับเมื่อต้องการใช้ฟังก์ชันการเรียกใช้เครื่องมือ (tool-calling)

การจัดการ Sampling Params

API เวอร์ชัน 3.x จะไม่สนใจค่า temperature และ top_p ดังนั้นควรใช้ system instructions ที่ชัดเจนเพื่อควบคุมความแน่นอนของผลลัพธ์

ใช้ประโยชน์จาก 1M Context

ส่งฐานความรู้ที่เกี่ยวข้องทั้งหมดเข้าไปใน prompt โดยตรงเพื่อให้ได้ผลลัพธ์ที่ดีกว่าระบบ RAG แบบแยกส่วน

การใช้ System Instructions

ควรใช้ system_instruction แทนการ prefilling ในการโต้ตอบ เพื่อป้องกันไม่ให้มีข้อความเกริ่นนำในผลลัพธ์รูปแบบ JSON

คำรับรอง

ผู้ใช้ของเราพูดอย่างไร

เข้าร่วมกับผู้ใช้ที่พึงพอใจนับพันที่ได้เปลี่ยนแปลงเวิร์กโฟลว์ของพวกเขา

Jonathan Kogan

Jonathan Kogan

Co-Founder/CEO, rpatools.io

Automatio is one of the most used for RPA Tools both internally and externally. It saves us countless hours of work and we realized this could do the same for other startups and so we choose Automatio for most of our automation needs.

Mohammed Ibrahim

Mohammed Ibrahim

CEO, qannas.pro

I have used many tools over the past 5 years, Automatio is the Jack of All trades.. !! it could be your scraping bot in the morning and then it becomes your VA by the noon and in the evening it does your automations.. its amazing!

Ben Bressington

Ben Bressington

CTO, AiChatSolutions

Automatio is fantastic and simple to use to extract data from any website. This allowed me to replace a developer and do tasks myself as they only take a few minutes to setup and forget about it. Automatio is a game changer!

Sarah Chen

Sarah Chen

Head of Growth, ScaleUp Labs

We've tried dozens of automation tools, but Automatio stands out for its flexibility and ease of use. Our team productivity increased by 40% within the first month of adoption.

David Park

David Park

Founder, DataDriven.io

The AI-powered features in Automatio are incredible. It understands context and adapts to changes in websites automatically. No more broken scrapers!

Emily Rodriguez

Emily Rodriguez

Marketing Director, GrowthMetrics

Automatio transformed our lead generation process. What used to take our team days now happens automatically in minutes. The ROI is incredible.

Jonathan Kogan

Jonathan Kogan

Co-Founder/CEO, rpatools.io

Automatio is one of the most used for RPA Tools both internally and externally. It saves us countless hours of work and we realized this could do the same for other startups and so we choose Automatio for most of our automation needs.

Mohammed Ibrahim

Mohammed Ibrahim

CEO, qannas.pro

I have used many tools over the past 5 years, Automatio is the Jack of All trades.. !! it could be your scraping bot in the morning and then it becomes your VA by the noon and in the evening it does your automations.. its amazing!

Ben Bressington

Ben Bressington

CTO, AiChatSolutions

Automatio is fantastic and simple to use to extract data from any website. This allowed me to replace a developer and do tasks myself as they only take a few minutes to setup and forget about it. Automatio is a game changer!

Sarah Chen

Sarah Chen

Head of Growth, ScaleUp Labs

We've tried dozens of automation tools, but Automatio stands out for its flexibility and ease of use. Our team productivity increased by 40% within the first month of adoption.

David Park

David Park

Founder, DataDriven.io

The AI-powered features in Automatio are incredible. It understands context and adapts to changes in websites automatically. No more broken scrapers!

Emily Rodriguez

Emily Rodriguez

Marketing Director, GrowthMetrics

Automatio transformed our lead generation process. What used to take our team days now happens automatically in minutes. The ROI is incredible.

ที่เกี่ยวข้อง AI Models

openai

GPT-4o mini

OpenAI

OpenAI's most cost-efficient small model, GPT-4o mini offers multimodal intelligence and high-speed performance at a significantly lower price point.

128K context
$0.15/$0.60/1M
alibaba

Qwen3-Coder-Next

alibaba

Qwen3-Coder-Next is Alibaba Cloud's elite Apache 2.0 coding model, featuring an 80B MoE architecture and 256k context window for advanced local development.

262K context
$0.12/$0.75/1M
zhipu

GLM-4.7

Zhipu (GLM)

GLM-4.7 by Zhipu AI is a flagship 358B MoE model featuring a 200K context window, elite 73.8% SWE-bench performance, and native Deep Thinking for agentic...

200K context
$0.60/$2.20/1M
google

Gemini 3.6 Flash

Google

Gemini 3.6 Flash is Google's high-speed model featuring a 17% reduction in token consumption, $1.50/M input pricing, and advanced 3D visualization.

1M context
$1.50/$7.50/1M
minimax

MiniMax M2.5

minimax

MiniMax M2.5 is a SOTA MoE model featuring a 1M context window and elite agentic coding capabilities at disruptive pricing for autonomous agents.

1M context
$0.15/$1.20/1M
other

MiMo V2.5 Pro

Other

MiMo V2.5 Pro is Xiaomi's open-source 1.02T parameter MoE model featuring a 1M context window, native multimodality, and elite agentic coding performance.

1M context
$1.00/$3.00/1M
moonshot

Kimi K3

Moonshot

Kimi K3 is Moonshot AI's 2.8T MoE model with a 1M token context window, native multimodal vision, and frontier-tier coding performance for complex agents.

1M context
$3.00/$15.00/1M
zhipu

GLM-5.2

Zhipu (GLM)

GLM-5.2 is Zhipu AI's flagship open-weight model featuring a 1M context window and specialized agentic coding capabilities under an MIT license.

1M context
$1.40/$4.40/1M

คำถามที่พบบ่อยเกี่ยวกับ Gemini 3.6 Flash Lite

ค้นหาคำตอบสำหรับคำถามทั่วไปเกี่ยวกับ Gemini 3.6 Flash Lite