All models

Claude Haiku 4.5 runs at 86% off on Moyi API. .

AnthropicAnthropic: Claude Haiku 4.5

No training

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance across reasoning, coding, and computer-use tasks, Haiku 4.5 brings frontier-level capability to real-time and high-volume applications. It introduces extended thinking to the Haiku line; enabling controllable reasoning depth, summarized or interleaved thought output, and tool-assisted workflows with full support for coding, bash, web search, and computer-use tools. Scoring >73% on SWE-bench Verified, Haiku 4.5 ranks among the world’s best coding models while maintaining exceptional responsiveness for sub-agents, parallelized execution, and scaled deployment.

Modalities

Price (in / out)

86% off

$0.141 / $0.705$1 / $5per 1M

Context

200K

Released

Oct 1, 2025

Providers

Moyi API routes by health and cost within the discount range and provider list you configured. When one returns an error we fall through to the next best — you don't pick a provider in the request. Configure discount range

Need a specific set of providers?

ProviderContextInputOutputCacheLatencyThroughputNo training
AnthropicAnthropic200K$1$0.141$5$0.705$0.1$0.0141No training
BedrockAWS Bedrock200K$1$0.141$5$0.705$0.1$0.0141No training

Live discount

Each provider's average effective discount, refreshed hourly.

No discount data for this model in the last 24 hours.

Performance

Latency and throughput are measured on live gateway traffic.

Latency

Time from sending the request to the first output token, broken down by provider.

    LATENCY · MOYI API GATEWAY

    Throughput

    Tokens per second after the first output token, broken down by provider.

      THROUGHPUT · MOYI API GATEWAY

      Uptime

      Moyi API continuously probes every provider and falls through to the next best one on error.

      Uptime (24h)

      Uptime (3d)

      Uptime, last 3 days

      -72h-48h-24hNow

      Uptime, last 24 hours

      One line per upstream provider, taking the best success rate among that provider's channels. The Moyi API line counts only the final attempt of each request — the end-to-end result after all fallbacks.

        AVAILABILITY · LAST 24H · MOYI API GATEWAY

        Data security

        Among this model's providers, Anthropic, AWS Bedrock do not train on your data.

        For the full data-security policy, learn more.

        Need to pick your own providers? .

        Quickstart

        from anthropic import Anthropic
        
        client = Anthropic(
            base_url="https://api.moyiapi.com",  # the only line that changes
            api_key="<MOYIAPI_API_KEY>",
        )
        message = client.messages.create(
            model="claude-haiku-4-5",
            max_tokens=1024,
            messages=[{"role": "user", "content": "Hello"}],
        )
        print(message.content[0].text)
        Read docs

        Endpoint: https://api.moyiapi.com

        Claude Haiku 4.5 pricing, performance & uptime