Skip to content

Claude 3.5 Haiku

POST/v1/messages

Claude 3.5 Haiku features offers enhanced capabilities in speed, coding accuracy, and tool use. Engineered to excel in real-time applications, it delivers quick response times that are essential for dynamic tasks such as chat interactions and immediate coding suggestions.

Request body

Parameters supported by this model. Values, defaults, and limits are read from the model registry.

stringmodelrequired

Model identifier. Set to anthropic/claude-3.5-haiku.

Default: anthropic/claude-3.5-haiku

array<object>messagesrequired

Conversation messages in system, user, or assistant order.

Optional<integer>max_tokens

Maximum number of tokens the model may generate in the response.

Range: −∞ to 8192

Optional<number>temperature

Sampling temperature. Lower values are more deterministic; higher values are more creative.

Range: 0 to 2

Default: 1

Optional<number>top_p

Nucleus sampling threshold. Use this or temperature, but usually not both.

Range: 0 to 1

Optional<boolean>stream

When true, returns incremental Server-Sent Events instead of one completed response.

Default: false

Optional<array>tools

Tool definitions that the model may call during the response.

Optional<string>tool_choice

Controls whether the model may call a tool and, when supported, which tool it must call.

Optional<array<string>>stop

Sequences that stop generation when the model produces one of them.

Optional<integer>top_k

Request parameter supported by this model.

Range: 0 to ∞

Response Schema

Fields returned by this model API response.

array<object>contentrequired

Assistant content blocks.

stringidrequired

Unique message identifier.

stringmodelrequired

Model that generated the message.

stringrolerequired

Message role returned by the assistant.

Optional<string>stop_reason

Reason generation stopped.

stringtyperequired

Response object type.

Model capabilities

array<string>capability_tagsrequired

Capabilities declared by the model registry.

Default: chat, vision, function_calling

integercontext_lengthrequired

Maximum context window accepted by this model.

Default: 200000 tokens

stringexecution_moderequired

Execution mode declared by the model registry.

Default: sync