medium. Leave temperature, top_p and top_k unset; non-default values return a 400 error.
Pricing is tiered by prompt length: when a request’s prompt (input plus cached tokens) is over 100,000 tokens, every token in that request bills at the higher tier. Both tiers are listed on the pricing page.
Try it in the browser: Claude Haiku 5.5 API playground.
Model
Supported Endpoints
All three API formats are supported:- Chat Completions —
POST /v1/chat/completions - Responses —
POST /v1/responses - Messages —
POST /v1/messages
Example
Volume discounts are available for almost all models. Reach out at contact@unifically.com to discuss custom pricing.
