/v1/messages, wenn dein Client bereits das Anthropic-Messages-Format spricht.
Tool-Nutzung
Deklariere Tools mittools (jedes hat name, description, input_schema) und
steuere die Auswahl mit tool_choice. Mehrstufige tool_use- / tool_result-
Content-Blöcke werden unterstützt.
Streaming
Setzestream: true, um Anthropic-Server-Sent-Events zu empfangen:
message_start, content_block_start, content_block_delta,
content_block_stop, message_delta, message_stop.
Parameter
Unterstützt:model, messages, max_tokens (erforderlich), system,
temperature, top_p, top_k, stop_sequences, tools, tool_choice
und stream.
Reasoning-Stärke
Verwende nativesoutput_config.effort für unterstützte Claude-Modelle. Die gültigen Werte hängen von der Version ab; siehe modellabhängige Reasoning-Steuerung. Natives thinking bleibt unabhängig: effort aktiviert Thinking nicht und legt kein Tokenbudget fest. Konfiguriere thinking bei Bedarf separat nach den Modellanforderungen.
max_completion_tokens, max_output_tokens und native Entsprechungen zählen Reasoning- und Antwort-Tokens gemeinsam. Effort ist kein Ausgabetokenlimit. Das Budget kann weiterhin erschöpft werden (finish_reason: "length" in Chat Completions), sodass die endgültige Antwort unvollständig oder leer bleibt.