← ホーム

要約

長い入力と短い出力のコストを比較します。レポート、記事、メモ、会議録など、テキスト量の多い処理に向いたモデルを探すためのページです。

見るべき点

要約は入力が重くなりやすい用途です。要約が短い場合、出力料金より入力料金の差が月額コストに効きます。

初期シナリオ

このページは、10,000 トークンの文書、500 トークンの要約、月 1,000 文書から開始します。

この要約候補を比較する 計算機を開く 要約料金の解説記事 → 要約パイプライン構築ガイドを読む →

入力が重い処理に向いた低コスト候補を比較ツールで並べて確認します。

初期シナリオで低コストな要約向けモデル

この表はサーバー側で生成されるため、スクリプト実行前でも検索エンジンがモデル、入力料金、出力料金、推定月額コストを読めます。

モデル 入力料金 出力料金 月額コスト
Qwen2.5-Coder-3B-Instruct $0.0100 / 1M tokens $0.0300 / 1M tokens $0.12
Qwen2.5-Coder-7B-Instruct $0.0100 / 1M tokens $0.0300 / 1M tokens $0.12
Qwen2.5-Coder-7B $0.0100 / 1M tokens $0.0300 / 1M tokens $0.12
llama3.2-11b-vision-instruct $0.0150 / 1M tokens $0.0250 / 1M tokens $0.16
llama3.2-3b-instruct $0.0150 / 1M tokens $0.0250 / 1M tokens $0.16
gpt-oss-20b $0.0145 / 1M tokens $0.0700 / 1M tokens $0.18
titan-embed-text-v2 $0.0200 / 1M tokens N/A $0.20
Llama-3.2-3B-Instruct $0.0200 / 1M tokens $0.0200 / 1M tokens $0.21

Open calculator to estimate your own workload cost with different token counts and model choices.

Browse providers to compare models from OpenAI, Anthropic, Google, Mistral, and other providers.

Pricing data from catalog last generated 2026/08/10. Verify before production decisions. Data sources.