Skip to main content
POST
CometAPI は Anthropic メッセージ API をネイティブにサポートしており、Anthropic 固有の機能を備えた Claude モデルへ直接アクセスできます。適応的思考、プロンプトキャッシュ、effort 制御といった Claude の機能には、このエンドポイントを使用してください。
完全なパラメータ一覧、レスポンススキーマ、Claude 固有の動作については、公式の Anthropic Messages API reference を信頼できる情報源として参照してください。この CometAPI ページでは、そのリクエスト形式を CometAPI 経由で送信する方法を説明します。
Claude の機能進化に伴い、Anthropic のリクエストパラメータとレスポンスフィールドは変更される可能性があります。最新の完全なパラメータ一覧とプロバイダー固有の挙動については、Anthropic Messages API documentation を確認してください。
多くの新しい Claude モデルは、Messages API でデフォルト以外の temperaturetop_ptop_k の値を受け付けません。選択したモデルでサポートを確認している場合を除き、これらのサンプリングフィールドは省略してください。モデルが未対応または非推奨パラメータのエラーを返した場合は、そのフィールドをリクエストから削除してください。
認証には x-api-keyAuthorization: Bearer ヘッダーの両方を利用できます。公式の Anthropic SDK はデフォルトで x-api-key を使用します。

クイックスタート

CometAPI で公式の Anthropic SDK を使用するには、ベース URL を設定します。

適応的思考を制御する

output_config.effort を使って適応的思考を利用し、Claude がレスポンスにどれだけの処理をかけるかを制御します。新しい Claude モデルは、従来の手動思考形式 thinking={"type": "enabled", "budget_tokens": ...} を受け付けません。
高い effort レベルを使用する場合、thinking トークン(Token)も max_tokens の上限に含まれます。思考分と最終回答の両方に十分な max_tokens を設定してください。

プロンプトをキャッシュする

後続のリクエストでレイテンシとコストを削減するために、大きな system プロンプトや会話のプレフィックスをキャッシュできます。キャッシュしたい content ブロックに cache_control を追加します。
キャッシュの利用状況は、レスポンスの usage フィールドで報告されます。
  • cache_creation_input_tokens — キャッシュに書き込まれたトークン(Token)(より高い料金で課金)
  • cache_read_input_tokens — キャッシュから読み取られたトークン(Token)(割引料金で課金)
プロンプトキャッシュを利用するには、キャッシュ対象の content ブロックに最低 1,024 tokens が必要です。これより短い content はキャッシュされません。

レスポンスをストリーミングする

Server-Sent Events(SSE)を使ってレスポンスをストリーミングするには、stream: true を設定します。イベントは次の順序で到着します。
  1. message_start — メッセージのメタデータと初期 usage を含みます
  2. content_block_start — 各 content ブロックの開始を示します
  3. content_block_delta — 増分テキストチャンク(text_delta
  4. content_block_stop — 各 content ブロックの終了を示します
  5. message_delta — 最終的な stop_reason と完全な usage
  6. message_stop — ストリームの終了を示します

effort を制御する

Claude がレスポンス生成にどれだけの effort をかけるかを制御するには、output_config.effort を使用します。

サーバーツールを使う

Claude は、Anthropic のインフラ上で実行されるサーバーサイドツールをサポートしています。
URL からコンテンツを取得して分析します。

レスポンス例

CometAPI の Anthropic エンドポイントからの典型的なレスポンス:

OpenAI互換エンドポイントとの比較

承認

x-api-key
string
header
必須

Your CometAPI key passed via the x-api-key header. Authorization: Bearer $COMETAPI_KEY is also supported.

ヘッダー

anthropic-version
string
デフォルト:2023-06-01

The Anthropic API version to use. Defaults to 2023-06-01.

:

"2023-06-01"

anthropic-beta
string

Comma-separated feature identifiers required by a specific beta API feature. Omit this header for the examples on this page.

ボディ

application/json
model
string
デフォルト:claude-opus-5
必須

The Claude model to use. See the Models page for available Claude model IDs.

:

"claude-opus-5"

messages
object[]
必須

Conversation history. Use user and assistant messages with text strings or content-block arrays. Return complete assistant content blocks when continuing a tool call.

max_tokens
integer
必須

The maximum number of tokens to generate. The model may stop before reaching this limit. When using thinking, the thinking tokens count towards this limit.

必須範囲: x >= 1
:

1024

system

System prompt providing context and instructions to Claude. Can be a plain string or an array of content blocks (useful for prompt caching).

temperature
number

Sampling temperature. The examples omit sampling overrides.

必須範囲: 0 <= x <= 1
top_p
number

Nucleus sampling threshold. The examples omit sampling overrides.

必須範囲: 0 <= x <= 1
top_k
integer

Limits sampling to the k most likely tokens. The examples omit sampling overrides.

必須範囲: x >= 0
stream
boolean
デフォルト:false

If true, stream the response incrementally using Server-Sent Events (SSE). Events include message_start, content_block_start, content_block_delta, content_block_stop, message_delta, and message_stop.

stop_sequences
string[]

Custom strings that cause the model to stop generating when encountered. The stop sequence is not included in the response.

thinking
object

Thinking configuration. The adaptive example uses type adaptive and sets output_config.effort separately.

tools
object[]

Client tools define a name and input_schema. Server tools use a versioned type and name, such as the web_fetch and web_search examples on this page.

tool_choice
object

Controls how the model uses tools.

metadata
object

Request metadata for tracking and analytics.

output_config
object

Configuration for reasoning effort and structured output.

service_tier
enum<string>

The service tier to use. auto tries priority capacity first, standard_only uses only standard capacity.

利用可能なオプション:
auto,
standard_only

レスポンス

Successful response. When stream is true, the response is a stream of SSE events.

id
string

Message identifier returned by the API.

type
enum<string>

Always message.

利用可能なオプション:
message
role
enum<string>

Always assistant.

利用可能なオプション:
assistant
content
object[]

The response content blocks. May include text, thinking, tool_use, and other block types.

model
string

Model ID reported by the response.

stop_reason
enum<string>

Why the model stopped generating. refusal can be returned as a successful HTTP response when the model declines a request.

利用可能なオプション:
end_turn,
max_tokens,
stop_sequence,
tool_use,
pause_turn,
refusal
stop_sequence
string | null

The stop sequence that caused the model to stop, if applicable.

usage
object

Token usage statistics.

最終更新日 2026年7月3日