Anthropic: Claude Sonnet 4.6

anthropic/claude-sonnet-4-6

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation,...

visiontools
MODALITIES
INPUT PRICE
$0.15per 1M
OUTPUT PRICE
$0.90per 1M
CONTEXT
1M
RELEASED
Feb 17, 2026
ProviderCacheUptimeChat
ccapi
ccapi
5.00(1)$0.15$0.9
Cache read$0.015
FastAPI
FastAPIClaude Code only
$0.6$3
Cache read$0.06

Capabilities

Input modalities
fileimagetext
Output modalities
text
Features
include_reasoningmax_completion_tokensmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstemperaturetool_choicetoolstop_ktop_pverbosity
1

Get API Key

Create an API key from the Keys page, then set it as an environment variable:

export ONLIST_API_KEY=sk-...
2

Make your first request

Endpoints

POSThttps://onlist.io/v1/chat/completions

OpenAI Chat Completions format

Request Headers
Authorization:Bearer $ONLIST_API_KEY
Content-Type:application/json
Model:anthropic/claude-sonnet-4-6
POSThttps://onlist.io/v1/responses

OpenAI Responses format

Request Headers
Authorization:Bearer $ONLIST_API_KEY
Content-Type:application/json
Model:anthropic/claude-sonnet-4-6
POSThttps://onlist.io/v1/messages

Anthropic Messages format

Request Headers
Authorization:Bearer $ONLIST_API_KEY
Content-Type:application/json
Model:anthropic/claude-sonnet-4-6

Code samples

curl https://onlist.io/v1/chat/completions \
  -H "Authorization: Bearer $ONLIST_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
       "model": "anthropic/claude-sonnet-4-6",
       "messages": [
         {
           "role": "user",
           "content": "Explain quantum entanglement in one paragraph."
         }
       ]
     }'

Replace $ONLIST_API_KEY with the API key from your Keys page.

Authentication

All requests must include an Authorization: Bearer <TOKEN> header. Generate keys from the Keys page; keys can be scoped to specific models, groups, IP ranges, and rate limits.

3

Enable streaming

Add "stream": true to receive partial responses as server-sent events in real time.

Streaming example

curl https://onlist.io/v1/chat/completions \
  -H "Authorization: Bearer $ONLIST_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
       "model": "anthropic/claude-sonnet-4-6",
       "messages": [
         {
           "role": "user",
           "content": "Write a haiku about recursion."
         }
       ],
       "stream": true
     }'

Supported parameters

NameTypeDescription
include_reasoning
max_completion_tokensintegerUpper bound on tokens generated, including visible and reasoning tokens.
max_tokensintegerMaximum number of tokens to generate in the completion.
reasoning
reasoning_effortstringConstrains effort on reasoning. Supported values: "low", "medium", "high".
response_formatobjectSpecifies the output format. Use {"type": "json_object"} for JSON mode.
stopstring | arrayUp to 4 sequences where the API will stop generating tokens.
structured_outputs
temperaturenumberSampling temperature between 0 and 2. Higher values make output more random.
tool_choicestring | objectControls which tool is called. "auto", "none", "required", or a specific function.
toolsarrayA list of tools the model may call. Currently supports functions.
top_kintegerLimits token selection to the k most likely candidates at each step.
top_pnumberNucleus sampling. The model considers tokens with top_p probability mass.
verbosity

These are the request parameters this model accepts. Parameter semantics follow the OpenAI Chat Completions specification.