/v1/messages를 사용하세요.
도구 사용
tools(각각 name, description, input_schema를 가짐)로 도구를 선언하고
tool_choice로 선택을 제어합니다. 멀티턴 tool_use / tool_result
콘텐츠 블록을 지원합니다.
스트리밍
stream: true로 설정하면 Anthropic SSE 이벤트를 받습니다:
message_start, content_block_start, content_block_delta,
content_block_stop, message_delta, message_stop.
파라미터
지원:model, messages, max_tokens(필수), system,
temperature, top_p, top_k, stop_sequences, tools, tool_choice,
stream.
추론 강도
지원되는 Claude 모델에서는 네이티브output_config.effort를 사용합니다. 지원 값은 버전에 따라 다르므로 모델별 추론 제어를 참고하세요. 네이티브 thinking은 독립적이며, effort는 추론을 켜거나 토큰 예산을 설정하지 않습니다. 필요하면 모델 요구 사항에 따라 thinking을 별도로 설정하세요.
max_completion_tokens, max_output_tokens 및 네이티브 대응 필드는 추론과 최종 답변 토큰의 합계를 제한합니다. effort 자체는 출력 토큰 상한이 아닙니다. 예산이 소진되면(Chat Completions의 finish_reason: "length") 최종 답변이 불완전하거나 비어 있을 수 있습니다.