すべてのモデル

Claude Sonnet 4.6 on VM0. The default agent model

Anthropicのワークホース。VM0 Managedのデフォルトモデル — 単発タスク、高速イテレーション、そして×1クレジット基準。

1M tokens · Text / Vision / Code · Prompt cache

Claude Sonnet 4.6はClaude 4ファミリーのワークホースであり、VM0 Managedのデフォルトモデルです。他のすべてのBuilt-inモデルが評価される×1クレジット基準を定義します。実践的には、大部分のエージェントステップに最適なモデルです。

定価$3/$15/1Mトークン、キャッシュ入力$0.30/1M。VM0では×1リファレンスモデルであるため、モデル間のすべてのコスト比較は最終的にSonnet 4.6との比較になります。

Claude Sonnet 4.6とは?

VM0ローンチ以来利用可能 · Claude 4ファミリーのワークホース。VM0 Managedのデフォルトモデル。

Claude Sonnet 4.6 sits in the middle of Anthropic's Claude 4 family. It is the workhorse model designed to handle the full breadth of typical agent work. Multi-tool routing, code edits, long-running conversations, and structured-output tasks. Without the cost premium of Opus.

Across VM0's Built-in lineup, every other model's credit multiplier is normalised against Sonnet 4.6 (×1). That makes Sonnet the right pick when you want predictable budget conversations: “this agent runs at roughly 2× a Sonnet step” is a more useful sentence than absolute dollar quotes that move every quarter.

Sonnet 4.6 supports Anthropic's prompt caching, which makes a big difference for VM0 agents that ship a stable system prompt and a fixed tool schema. Cached input tokens bill at $0.30 per 1M instead of $3. A 10× saving on the parts of the prompt that don't change between turns.

Claude Sonnet 4.6の注目ポイント

アーキテクチャと機能の主な特徴。

Sonnet 4.6はClaude 4アーキテクチャ(1Mトークンコンテキスト、ネイティブツール使用)を共有します。適応思考努力レベルは使用せず、単発および短いマルチターンタスクでの一貫したスループットと品質に最適化されています。

スペック概要

ファミリーClaude 4世代
モダリティテキスト、画像、コード
言語英語中心、多言語対応
プロンプトキャッシュサポート(Anthropic)
コンテキストウィンドウ1,000Kトークン
最大出力32Kトークン
デフォルトVM0 Managed

Claude Sonnet 4.6のベンチマーク

Sonnet 4.6 sits roughly 3 to 4 percentage points behind Opus 4.6 on Anthropic's headline coding benchmarks while being three to five times cheaper at the vendor level. The typical Opus/Sonnet trade-off.

SWE-bench Verifiedvendor-reported
~77%
Long-context recallinternal observation
Strong across 100K+
Tool routingVM0 internal
Best in class at ×1

Claude Sonnet 4.6の価格

プロバイダー定価、100万トークンあたり。

入力$3.00
出力$15.00
キャッシュ読み取り$0.30
キャッシュ書き込み$3.75

Claude Sonnet 4.6の実践的な挙動

本番エージェント実行で観測された動作。

Tool routing

Best-in-class tool-routing accuracy at this price. On multi-tool flows across Slack, GitHub, Linear, and Notion, Sonnet 4.6 picks the correct tool with the correct arguments more reliably than any model below ×2.

Long-context coherence

Coherent across 100K+ token transcripts. Drops below Opus 4.7 only on the longest, most adversarial runs.

Speed

Opus より速く、Kimi K2.7 Code より遅いモデルです。本番エージェントにとって速度と品質のバランスがよい選択肢です。

Cost predictability

Pricing is the credit baseline; prompt caching makes the on-VM0 cost especially predictable for agents with fixed system prompts.

Claude Sonnet 4.6に最適なエージェントタスク

The Slack agent that knows where things live

Triages incoming questions, follows up on stalled threads, posts status updates, and answers search-style queries ("who's owning the auth refactor?"). Sonnet's tool-routing accuracy means the right tool gets called with the right arguments on the first try, even when the request is ambiguous, so the agent feels reliable instead of flaky.

The PR review agent that doesn't drown in noise

Sonnet handles the bulk of code-aware work — PR review, test scaffolding, refactor suggestions, bug bisection — without leaving stylistic comments that nobody asked for. The 1M-token context window lets it pull in the related files and prior reviews when it matters, and you only escalate to Opus 4.7 for the patches Sonnet visibly struggles with.

安定したシステムプロンプトを持つカスタマーサポートアシスタント

長い会話履歴、CRMへの頻繁なツール呼び出し、毎ターン同じ大きなシステムプロンプトとツールスキーマ。Sonnetのプロンプトキャッシュは、初回呼び出し後にその固定プレフィックスの入力コストを大きく下げ、会話量が増えても1会話あたりのコストを安定させます。

The customer-support assistant with a stable system prompt

Long conversation histories, frequent tool calls into the CRM, the same hefty system prompt and tool schema on every turn. Sonnet's prompt caching turns that fixed prefix into a fraction of the input cost after the first call, which is what keeps per-conversation cost flat as volume grows.

Claude Sonnet 4.6を避けるべきケース

最難関の推論ステップで指示を明らかに落とす場合はOpus 4.7へ上げるべきです。高ボリュームの分類ではGPT-5.4 Miniがより安価な一括処理モデルであり、低レイテンシの短い応答ではKimi K2.7 Codeの方が大幅に高速なため、Sonnet 4.6は避けてください。

Claude Sonnet 4.6 vs 他のモデル

Claude Sonnet 4.6 vs Claude Opus 4.7

Sonnet 4.6は×1、Opus 4.7は×2です。Sonnetはほとんどのエージェントを処理でき、推論の深さがスループットより重要なときにOpusへ上げます。多くのチームはOpusをプランナー、Sonnetをワーカーとして使います。

Claude Sonnet 4.6 vs DeepSeek V4 Pro

DeepSeek V4 Pro(×0.1)はコーディングベンチマークでSonnetに近い品質をかなり低いコストで出します。代償はツールルーティングの信頼性と安全性プロファイルがやや成熟していない点です。

Claude Sonnet 4.6 vs GPT-5.4 Mini

GPT-5.4 MiniはOpenAI側の安価な大量処理オプションです。ツールルーティングの信頼性が重要ならSonnetを使い、高ボリュームの事前フィルタリングやSonnet級のルーティングを必要としない簡単なステップではMiniを使います。

結論: Claude Sonnet 4.6を使うべきか?

Sonnet 4.6はVM0の正しいデフォルトモデルです。どのカテゴリーでも最高ではありませんが、エージェントステップの80%以上で十分な性能です。ここから始め、特定の理由がある場合のみ他のモデルにルーティングしてください。

よくある質問

Sonnet 4.6がVM0 Managedのデフォルトな理由は?

推論品質、速度、コストの最良のバランスを提供するためです。×1クレジット基準として、エージェントワークフローの自然な出発点です。

Sonnet 4.6のコンテキストウィンドウは?

100万トークン(1,000K)で、Opus 4.7およびOpus 4.6と同一です。

Sonnet 4.6は画像入力に対応していますか?

はい。Sonnet 4.6は画像入力を受け付け、ビジュアル推論を実行できます。

Sonnet 4.6から切り替えるべきタイミングは?

高度な推論や長ループにはOpus 4.7に。高ボリューム、レイテンシ重視、コスト重視のタスクにはKimi K2.7 CodeまたはGPT-5.4 Miniに切り替え。

Sonnet 4.6はSonnet 4.5と同じですか?

いいえ。Sonnet 4.6は推論と1Mトークンコンテキストが改善された新世代です。Claude 4ファミリーでSonnet 4.5を置き換えます。

代替モデル

VM0でClaude Sonnet 4.6を使う

VM0でClaude Sonnet 4.6にアクセスする2つの方法

VM0はClaude Sonnet 4.6を、VM0クレジットで課金されるBuilt-inモデル、およびAnthropic API keyを使用したBring-your-ownの2通りでサポートしています。Built-inパスではVM0 Managedルーティングと後述のクレジット倍率が適用され、Bring-your-ownパスでは上流プロバイダーに直接課金され、VM0クレジットへの変換は行われません。

VM0の推奨

VM0はClaude Sonnet 4.6をコアエージェントモデルとして位置付けており、Claude Opus 4.7、Claude Opus 4.6、Claude Sonnet 4.6と並んで、エージェント実行の実際の成果を左右するステップに推奨されます。これらは、オーケストレーター役、コードを扱うエージェント、誤った回答のコストが高いステップに選ぶモデルです。

クレジットと×1倍率

VM0のすべてのBuilt-inモデルは、×1クレジット基準となるClaude Sonnet 4.6の倍数で価格設定されています。Claude Sonnet 4.6は×1クレジットで課金されます。倍率はVM0の請求書に表示されるもので、上記の価格表のベンダー定価はVM0がクレジットに変換する前に上流プロバイダーが請求する金額です。

Claude Sonnet 4.6は×1基準であり、他のすべてのBuilt-inモデルの価格の基準となるため、VM0でモデルを選択する際のコスト比較の単位となります。

VM0でAvailable since launchから利用可能。