PLAY AI 筆記回 PLAY AI
AI 知識科普發布 · 9 分鐘 · 約 3,869 字

Gemini 3 Pro 深度解析:MoE 架構、Deep Think 與代理能力全面升級

Gemini 3 Pro 是 Google 最新發布的多模態大型語言模型(LLM),其核心定位在於「高階推理」與「自主任務執行」。 Gemini 3 Pro 採用了先進的 MoE(混合專家)架構,在處理複雜邏輯、跨媒體理解及長文本分析上,展現了超越 Gemini 2.5 Pro 的性能。

PLAY
PLAY AI 編輯部把複雜概念拆成一般人也讀得懂的版本,聚焦名詞解釋、比較、來源與實作脈絡。
本篇內容
Gemini 3 Pro 深度解析:MoE 架構、Deep Think 與代理能力全面升級
Gemini 3 Pro 深度解析:MoE 架構、Deep Think 與代理能力全面升級

背景與脈絡

Gemini 3 Pro 是 Google 最新發布的多模態大型語言模型(LLM),其核心定位在於「高階推理」與「自主任務執行」。 Gemini 2.5 has an experimental context window of up to 1 million tokens, surpassing Claude 3 (~200k) and GPT-4 Turbo (~150k). Gemini 3.8 Flash 在 HLE-Verified (跨學科專家推理) 取得 54.9% 的成績,優於 Claude Opus 5 (54.4%) 與 GPT-5.6 Sol (54.5%)。[1][2][3]

Gemini 3 Pro 採用了先進的 MoE(混合專家)架構,在處理複雜邏輯、跨媒體理解及長文本分析上,展現了超越 Gemini 2.5 Pro 的性能。 Gemini 2.5 achieved a score of 18.8% on Humanity's Last Exam (HLE), a benchmark with ~3,000 questions across over 100 fields designed to test expert-level reasoning. Gemini 3.8 Flash 在 BioMysteryBench (生物資訊研究工作流) 的 Human Difficult 難度下取得 56.5% 的成績,優於 Claude Opus 5 (49.4%) 與 GPT-5.6 Terra (49.4%)。[1][2][3]

Gemini 3 Pro 繼承了前代模型處理 100 萬 tokens 以上的能力,可輕鬆閱讀整本書籍或大型程式庫。 Gemini 2.5 demonstrates strong performance in STEM and coding benchmarks such as AIME, often scoring at or near the top compared to leading models. Gemini 3.8 Flash 在 Harvey's Legal Agent Benchmark (複雜法律工作流) 的通過率為 10.0%,在該評測中表現領先於 Gemini 3.7 Flash (8.8%) 與 Claude Opus 5 (6.7%)。[1][2][3]

運作機制與關鍵差異

Gemini 3 Pro 引入了一項突破性功能 Deep Think,允許使用者調整模型的「思考層級(Thinking Level)」,在回應速度與推理深度之間取得最佳平衡,以應對高複雜度的數學或邏輯難題。 Gemini 2.5 shows clear leadership in multimodal tasks, particularly evidenced by its performance on the MMMU benchmark. Gemini 3.8 Flash 的 API 價格包含促銷價與常規價:輸入為 $0.75/1M tokens (常規 $1.50),輸出為 $3.75/1M tokens (常規 $7.50)。[1][2][3]

Gemini 3 Pro Deep Think 運作流程
使用者可調整思考層級以平衡推理深度與回應速度

Gemini 3 Pro 的 Agent 能力使其能自主規劃任務步驟、撰寫並執行程式碼、甚至呼叫外部工具來完成如「Vibe Coding」等複雜工作,不再僅限於文字生成。 Advanced models like GPT-4.0 have achieved very low scores on HLE (around 3%), while a DeepMind research paper reported a higher score of 26% on the same benchmark. Gemini 3.8 Flash 的促銷價格有效期至 2026 年 12 月 31 日,自 2027 年 1 月 1 日起將恢復常規價格。[1][2][3]

Gemini 3 Pro 在多模態理解上新增了「媒體解析度控制(Media Resolution)」功能,讓使用者能根據任務需求,決定模型分析圖片、PDF 或影片時的細緻程度,藉此優化 Token 消耗與精準度。 Gemini 3.8 Flash 專為大規模處理複雜的代理任務 (agentic tasks) 而設計。 Gemini 3.5 Flash-Lite 的執行延遲低於 Gemini 3.5 Flash,適合高容量任務。[1][3]

實際影響與應用

目前 Gemini 3 Pro 預覽版已在 Google AI Studio 開放免費試用,一般大眾可直接體驗其聊天與多模態功能;若需透過 API 串接應用程式,則採用按 Token 計費的模式。 Gemini 3.5 Flash-Lite 適用於需要高效率與智能的高容量任務 (high-volume tasks)。 Gemini 3.1 is the latest version of the Gemini model family as of the Google Cloud Next 2026 event.[1][3][4]

Gemini 3 Pro 預覽版已於 2025 年 11 月中旬陸續在 Google AI Studio 上線,開發者可立即登入試用。 Gemini 3.1 Pro 適用於複雜任務並將創意概念轉化為現實。 Gemini 3.1 Pro is designed for powering agents and coding tasks, particularly in STEM and engineering contexts.[1][3][4]

Gemini 3 Pro API 定價結構
依 token 量級分層計費,輸出成本顯著高於輸入

Gemini 3 Pro API 定價:輸入 < 20萬 tokens 時為 $2.00 USD/1M tokens,輸出 < 20萬 tokens 時為 $12.00 USD/1M tokens;輸入 > 20萬 tokens 時為 $4.00 USD/1M tokens,輸出 > 20萬 tokens 時為 $18.00 USD/1M tokens。 Gemini 3.1 Deep Think 針對科學、研究與工程領域的現代挑戰進行優化。[1][3]

限制與風險

Gemini 3 Pro 在「多模態原生理解」與「長文本處理」上具有傳統優勢,且新增的 Deep Think 模式強化了邏輯推理能力,使其在處理複雜任務時更具競爭力。 Gemini 系列具備代理編碼 (Agentic coding)、進階多模態理解 (Advanced multimodal understanding)、長程任務執行 (Long horizon tasks) 以及多步驟問題解決 (Multi-step problem-solving) 等能力。[1][3]

Gemini 3 Pro 核心能力總覽
多模態、長文本、代理任務與推理為主要優勢

Gemini 2.5 is an experimental model released by Google DeepMind in March 2025, designed with emphasis on deep, structured reasoning rather than solely fluent text generation. Gemini 3.8 Flash 在 Terminal-bench 2.1 (代理終端編碼) 取得 89.4% 的成績,優於 Gemini 3.7 Flash (85.8%)、Claude Opus 5 (89.1%) 與 GPT-5.6 Sol (88.8%)。[2][3]

值得持續觀察的變化

Gemini 2.5 features a multimodal architecture capable of processing text, images, audio, video, and code inputs and outputs. Gemini 3.8 Flash 在 LVBench (長影片理解) 的代理模式 (agentic) 取得 87.8% 的成績,優於 Gemini 3.7 Flash (85.4%) 與 Claude Opus 5 (75.4%)。[2][3]

資料來源與延伸閱讀

以下來源用於核對本文的技術背景與關鍵事實;正文中的編號可直接跳到對應來源。產品規格與時效性資訊仍以原始官方頁面最新版本為準。

查看 10 個來源
  1. Gemini 3 Pro 深度評測:Google 最強 AI Agent 誕生,推理能力與價格完整解析 - YOLO LAB|解構科技邊際與媒體娛樂的數據實驗室
  2. Gemini on Humanity's Last Exam: Benchmark Deep Dive
  3. Gemini — Google DeepMind
  4. What's new with Gemini from Google DeepMind
  5. Google Gemini: Scalable Multimodal Models
  6. Gemini: A Family of Highly Capable Multimodal Models
  7. Gemini 3.1 Pro — Google DeepMind
  8. Introducing Gemini: Google’s most capable AI model yet
  9. Google models  |  Gemini Enterprise Agent Platform  |  Google Cloud Documentation
  10. How Does Google Gemini AI Work? (2026)