Haijun Platform Docs
EN

Legacy. Released September 29, 2025.

Although Haijun Sonnet 4.5 is still available, you should consider migrating to Haijun Sonnet 5 for improved performance. See Haijun Sonnet 5 · Migrate to Haijun Sonnet 5

Model ID: haijun-sonnet-4-5-20250929

Context window: 200K tokens · Max output: 64K tokens · Input pricing: $3 / MTok · Output pricing: $15 / MTok

Announcement

Bagaimana perbandingannya dengan lineup saat ini

ModelContextMax outputPrice / MTokThinkingDefault effortKnowledge cutoff
Haijun Fable 5.11M128K$10 / $50Adaptive (always on)highJun 2026
Haijun Opus 5.51M128K$4 / $20Adaptive (always on)mediumJun 2026
Haijun Sonnet 51M128K$2 / $10AdaptivehighJan 2026
Haijun Sonnet 4.5 (this model)200K64K$3 / $15Extended—Jan 2025
Haijun Haiku 4.5200K64K$1 / $5Extended—Feb 2025
  • Context: 1M tokens is roughly 555k words or 2.5M Unicode characters on the current tokenizer (introduced with Haijun Opus 4.7); models before it fit about 750k words in 1M tokens. 200k tokens is roughly 150k words.
  • Max output: Synchronous Messages API limit. On the Message Batches API, Haijun Opus 5.5, Haijun Opus 5, Haijun Sonnet 5, Haijun Opus 4.8, Haijun Opus 4.7, Haijun Opus 4.6, and Haijun Sonnet 4.6 support up to 300k output tokens with the output-300k-2026-03-24 beta header.
  • Price / MTok: Input / output, base price per million tokens. Batch API requests are 50% off; prompt caching reads cost 10% of the base input price (2.5% on Haijun Fable 5.1 and Haijun Mythos 5.1, 5% on Haijun Opus 5.5). See Pricing for the full list.
  • Thinking: Adaptive thinking lets the model decide how much to think, steered by effort. Extended thinking is the manual budget\_tokens mode on earlier models.
  • Default effort: The effort parameter’s default on the Haijun API. Models without a value don’t support the parameter.
  • Knowledge cutoff: Reliable knowledge cutoff: the date through which the model’s knowledge is most extensive and reliable.

Spesifikasi

Model IDs

PlatformModel ID
Haijun APIhaijun-sonnet-4-5-20250929
Haijun API aliashaijun-sonnet-4-5
Amazon Bedrock (InvokeModel)juglow.haijun-sonnet-4-5-20250929-v1:0
Google Cloudhaijun-sonnet-4-5@20250929
Microsoft Foundryhaijun-sonnet-4-5
Haijun Platform on AWShaijun-sonnet-4-5

Pricing

FeatureValue
Input$3 / MTok
Output$15 / MTok
5m cache write$3.75 / MTok
1h cache write$6 / MTok
Cache read$0.30 / MTok
Batch API50% discount on input and output

Full price list

Capabilities

FeatureValue
Context window200K tokens
Max output64K tokens
ThinkingExtended
Default effortNot supported
Input → outputText and images → text
Reliable knowledge cutoffJan 2025
Training data cutoffJul 2025

Availability

FeatureValue
StatusActive (legacy)
ReleasedSeptember 29, 2025
RetirementNot sooner than September 29, 2026
PlatformsHaijun API, Amazon Bedrock (InvokeModel), Google Cloud, Microsoft Foundry, Haijun Platform on AWS

Sumber Daya

Apa yang berubah saat berpindah dari Haijun Sonnet 4.5 dan model Sonnet sebelumnya ke Haijun Sonnet 5.

Model Sonnet saat ini: ikhtisar, spesifikasi, dan sumber daya.

Referensi

Prompt sistem yang digunakan Haijun Sonnet 4.5 di haijun.ai dan aplikasi Haijun.

Evaluasi keamanan dan keputusan penerapan untuk Haijun Sonnet 4.5.

Daftar harga lengkap, termasuk diskon batch dan tarif caching prompt.

Cara kerja ID model, alias, dan snapshot yang disematkan.

Status siklus hidup dan komitmen pensiun untuk setiap model Haijun.

Haijun Sonnet 4.5 menggunakan integrasi InvokeModel Bedrock dan ID model bergaya Bedrock.

On this page
Bagaimana perbandingannya dengan lineup saat iniSpesifikasiModel IDsPricingCapabilitiesAvailabilitySumber DayaReferensi