一行要約

xhigh から始めるのがこのモデルの出発点(コーディング / agentic)。そして Opus 5 とは逆に、subagent を「少なく」起こす傾向があるので、必要なら明示的に促す必要がある。

要点

Effort is likely to be more important for this model than for any prior Opus, so experiment with it actively when you upgrade. このモデルでは effort が、これまでのどの Opus よりも重要になる可能性が高い。移行の際には積極的に振って試せ。

effort の較正

コーディング / agentic は xhigh から。知能重視の用途は最低 high

レベルOpus 4.8 での指針
max性能向上が出る用途もあるが、トークン増に対して収穫逓減。過剰思考に陥ることもある
xhighほとんどのコーディング / agentic 用途で最良
highトークンと知能のバランス。知能重視の用途の最低線
mediumコスト重視
low短くスコープの狭いタスク、レイテンシ重視で知能を要さないもの

Claude Opus 4.8 respects effort levels strictly, especially at the low end. On moderately complex tasks running at low effort there is some risk of under-thinking. Claude Opus 4.8 は effort レベルを厳密に尊重する。特に低い側でそうである。low の effort で中程度に複雑なタスクを走らせると、思考不足に陥る危険がいくらかある。

推論が浅いときは prompt で回避せず effort を上げる。

thinking は明示的に {type: "adaptive"} を設定するまでオフ。

If you are running at max or xhigh effort, set a large max output token budget. Start at 64k tokens and tune from there. max または xhigh の effort で走らせるなら、出力トークンの上限を大きく取れ64k トークンから始めて、そこから調整せよ。

ツール使用の発火(Opus 5 / Sonnet 5 との違い)

Claude Opus 4.8 has a tendency to favor reasoning over tool calls. This produces better results in most cases. However, increasing the effort setting is a useful lever to increase the level of tool usage, especially in knowledge work. Claude Opus 4.8 にはツール呼び出しより推論を好む傾向がある。多くの場合これは良い結果を生む。ただし、effort 設定を上げることはツール使用量を増やす有効なレバーである。特に知的労働においてそうである。

subagent の起動(Opus 5 と逆)

Claude Opus 4.8 tends to spawn fewer subagents by default. However, this behavior is steerable through prompting; give Claude Opus 4.8 explicit guidance around when subagents are desirable. Claude Opus 4.8 は既定では subagent を立てる数が少ない傾向にある。ただしこの挙動はプロンプトで誘導できる。どういうときに subagent が望ましいかを、Claude Opus 4.8 に明示的に伝えよ。

Opus 5 が「委譲しすぎるので抑える」なのに対し、4.8 は「委譲しないので促す」。移行時に prompt の向きを反転させる必要がある。

より字義通りの指示追従

低い effort で特に顕著。指示を別の項目へ暗黙に一般化しないし、していない要求を推測しない。 広く適用してほしいならスコープを明示する。

トーン

Claude Opus 4.8 tends toward a direct, opinionated style with minimal validation-forward phrasing and sparing emoji use. Claude Opus 4.8 は、直接的で主張のある文体に傾き、相手を持ち上げる言い回しは最小限で、絵文字も控えめである。

デザインの既定(具体的に記述されている)

a consistent default house style: warm cream/off-white backgrounds (~#F4F1EA), serif display type (Georgia, Fraunces, Playfair), italic word-accents, and a terracotta/amber accent. 一貫した既定の「ハウススタイル」がある。温かみのあるクリーム/オフホワイトの背景(およそ #F4F1EA)、セリフ体の見出し(Georgia、Fraunces、Playfair)、イタリックによる語の強調、テラコッタ/琥珀色のアクセントである。

エディトリアル・ホスピタリティ・ポートフォリオには合うが、ダッシュボード・開発ツール・fintech・ヘルスケア・エンタープライズアプリでは外すスライドデッキにも web UI にも現れる。

This default is persistent. Generic instructions (“don’t use cream,” “make it clean and minimal”) tend to shift the model to a different fixed palette rather than producing variety. この既定は根強い。漠然とした指示(「クリーム色は使うな」「清潔でミニマルに」)は、多様性を生むのではなく、モデルを別の固定パレットへ移すだけになりがちである

効くのは2つ — ① 具体的な代替を指定する ② 作る前に選択肢を提案させる

Claude Opus 4.8 requires less frontend design prompting than previous models to avoid “AI slop”. 従来より短いスニペットで足りる。

対話的コーディング製品

it tends to use more tokens in interactive settings, primarily because it reasons more after user turns. 対話的な場面ではより多くのトークンを使う傾向がある。主な理由は、user のターンの後により多く推論するからである。

これは long-horizon の一貫性・指示追従・コーディング能力を改善するが、トークンも増える。xhigh / high を使い、auto mode のような自律機能を足し、必要な人間の介入を減らす。

コードレビュー harness

Opus 4.8 はバグ発見が有意に良く、社内 eval では recall と precision の両方が高い。 ただし旧モデル向けに調整した harness では recall が下がって見えることがある。

This is likely a harness effect, not a capability regression. 「重大な問題だけ」「保守的に」という指示を従来より忠実に守るため、同じ深さで調査した上で報告数が減る。

computer use

最大解像度 2576px / 3.75MP。1080p が性能とコストのバランスが良い。 コスト重視なら 720p / 1366×768。

そのまま使える具体例

冗長さを減らす:

Provide concise, focused responses. Skip non-essential context, and keep examples minimal.

低 effort での推論の浅さへの当て木:

This task involves multistep reasoning. Think carefully through the problem before responding.

thinking の発火を抑える:

Thinking adds latency and should only be used when it will meaningfully improve answer quality — typically for problems that require multistep reasoning. When in doubt, respond directly.

トーンを温かくする:

Use a warm, collaborative tone. Acknowledge the user's framing before answering.

subagent を「促す」プロンプト(Opus 5 の「抑える」プロンプトと対になる):

Do not spawn a subagent for work you can complete directly in a single response (e.g. refactoring a function you can already see).
 
Spawn multiple subagents in the same turn when fanning out across items or reading multiple files.

デザインの既定を破る(選択肢を提案させる):

Before building, propose 4 distinct visual directions tailored to this brief (each as: bg hex / accent hex / typeface — one-line rationale). Ask the user to pick one, then implement only that direction.

“AI slop” 回避(Opus 4.8 ではこれで足りる):

<frontend_aesthetics>
NEVER use generic AI-generated aesthetics like overused font families (Inter, Roboto, Arial, system fonts), cliched color schemes (particularly purple gradients on white or dark backgrounds), predictable layouts and component patterns, and cookie-cutter design that lacks context-specific character. Use unique fonts, cohesive colors and themes, and animations for effects and micro-interactions.
</frontend_aesthetics>

コードレビューの recall を戻す:

Report every issue you find, including ones you are uncertain about or consider low-severity. Do not filter for importance or confidence at this stage - a separate verification step will do that. Your goal here is coverage: it is better to surface a finding that later gets filtered out than to silently drop a real bug. For each finding, include your confidence level and an estimated severity so a downstream filter can rank them.

原典で言及されている関連文書

未取得の派生リンク