gemini-3.5-flash-lite
Vertex AI · Chat model
gemini-3.5-flash-lite は Vertex AI の チャット モデルとして掲載されています。料金、上限、機能、ソースを比較しやすい形で確認できます。
正確性のため、プロバイダー名とモデル ID は原文のまま表示しています。
vertex_ai-language-models-gemini-3-5-flash-lite
カタログ生成日: 2026/08/10
要点
向いている確認
自分のワークロードで試す前に、料金、コンテキスト長、対応機能を短時間で確認したいときに使います。
公式側で確認すること
割引、キャッシュ料金、リージョン条件、公開メタデータに出ていないモデル上限は、必ずプロバイダーの公式ページで確認してください。
料金
| 項目 | 料金 |
|---|---|
| 入力 | $0.3000 / 1M tokens |
| 出力 | $2.5000 / 1M tokens |
| キャッシュ入力 | $0.0300 / 1M tokens |
| reasoning 出力 | $2.5000 / 1M tokens |
| 埋め込み | $0.3000 / 1M tokens |
| バッチ入力 | $0.1500 / 1M tokens |
| バッチ出力 | $1.2500 / 1M tokens |
| priority 入力 | $0.5400 / 1M tokens |
| priority 出力 | $4.5000 / 1M tokens |
| flex 入力 | $0.1500 / 1M tokens |
| flex 出力 | $1.2500 / 1M tokens |
上限
機能
| 機能 | 対応 |
|---|---|
| Vision | 対応 |
| Function calling | 対応 |
| Parallel function calling | 対応 |
| Tool choice | 対応 |
| Prompt caching | 対応 |
| Reasoning | 対応 |
| Response schema | 対応 |
| System messages | 対応 |
| Audio input | 対応 |
| Audio output | - |
| Web search | 対応 |
| PDF input | 対応 |
| Video input | 対応 |
| Native streaming | 対応 |
| Computer use | - |
| Assistant prefill | - |
| Structured output | - |
| Output config | - |
| URL context | 対応 |
近いモデル
下の候補は、確認済みの入出力形式が同じで、トークン料金を比較できるモデルだけです。フィルターで重視する観点を切り替えてから比較できます。
| モデル | 料金 | 入力形状 | 機能 | コンテキスト | 近い理由 |
|---|---|---|---|---|---|
| gemini-3.5-flash-lite Vertex AI | 入力 $0.3000 / 1M tokens 出力 $2.5000 / 1M tokens | pdfurl 出力: text | VisionFunction callingParallel function callingTool choice | 65.5K | 現在のモデル 基準行 |
| モデル | 料金 | 入力形状 | 機能 | コンテキスト | 近い理由 |
|---|---|---|---|---|---|
| gemini-3.5-flash-lite Google | 入力 $0.3000 / 1M tokens 出力 $2.5000 / 1M tokens | pdfurl 出力: text | VisionFunction callingParallel function callingTool choice | 65.5K | 入出力形式が一致 総合 76% |
| gemini-3.5-flash-lite Vertex AI | 入力 $0.3000 / 1M tokens 出力 $2.5000 / 1M tokens | pdfurl 出力: text | VisionFunction callingParallel function callingTool choice | 65.5K | 同じプロバイダー 総合 76% |
| gemini-2.5-flash vertex_ai-language-models | 入力 $0.3000 / 1M tokens 出力 $2.5000 / 1M tokens | pdfurl 出力: text | VisionFunction callingParallel function callingTool choice | 65.5K | 同じプロバイダー 総合 72% |
| gemini-2.5-flash-preview-09-2025 vertex_ai-language-models | 入力 $0.3000 / 1M tokens 出力 $2.5000 / 1M tokens | pdfurl 出力: text | VisionFunction callingParallel function callingTool choice | 65.5K | 同じプロバイダー 総合 72% |
| Gemini 2.5 Flash Google | 入力 $0.3000 / 1M tokens 出力 $2.5000 / 1M tokens | pdfurl 出力: text | VisionFunction callingParallel function callingTool choice | 65.5K | 入出力形式が一致 総合 72% |
| gemini-2.5-flash-preview-09-2025 Google | 入力 $0.3000 / 1M tokens 出力 $2.5000 / 1M tokens | pdfurl 出力: text | VisionFunction callingParallel function callingTool choice | 65.5K | 入出力形式が一致 総合 72% |
| gemini-flash-latest Google | 入力 $0.3000 / 1M tokens 出力 $2.5000 / 1M tokens | pdfurl 出力: text | VisionFunction callingParallel function callingTool choice | 65.5K | 入出力形式が一致 総合 72% |
| gemini-3.1-flash-lite-preview vertex_ai-language-models | 入力 $0.2500 / 1M tokens 出力 $1.5000 / 1M tokens | pdfurl 出力: text | VisionFunction callingParallel function callingTool choice | 65.5K | 同じプロバイダー 総合 70% |
| gemini-3.1-flash-lite vertex_ai-language-models | 入力 $0.2500 / 1M tokens 出力 $1.5000 / 1M tokens | pdfurl 出力: text | VisionFunction callingParallel function callingTool choice | 65.5K | 同じプロバイダー 総合 70% |
| gemini-3.1-flash-lite-preview Google | 入力 $0.2500 / 1M tokens 出力 $1.5000 / 1M tokens | pdfurl 出力: text | VisionFunction callingParallel function callingTool choice | 65.5K | 入出力形式が一致 総合 70% |
このフィルターに合うモデルはありません。
関連記事
このモデルのプロバイダー、機能、ユースケースに関連する記事です。
Reasoning LLM API Pricing Guide
Which reasoning models are cheapest, which are most expensive, and when to use reasoning vs non-reasoning models.
Model Routing Cascade
A four-tier Budget, Mid, Premium, and Reasoning model selection framework for routing queries to the lowest-cost suitable tier.
Hidden Costs of LLM APIs
Fine-tuning, rate limits, latency tradeoffs, evaluation, integration, and vendor risk — the LLM API costs that don't show up in per-token comparisons.
Output vs Input Pricing Multiplier
Output tokens cost 3.6× more than input on average across 2,076 chat models. Provider breakdowns, reasoning vs non-reasoning, and real workload estimates.
ソース
料金とメタデータの出典は英語版と同じデータベースを使っています。公式の提供元の事実情報は翻訳や推測で補完していません。
カタログ生成日: 2026-08-10T21:27:29.020Z