Skip to content
Merged
Show file tree
Hide file tree
Changes from all commits
Commits
File filter

Filter by extension

Filter by extension


Conversations
Failed to load comments.
Loading
Jump to
Jump to file
Failed to load files.
Loading
Diff view
Diff view
9 changes: 4 additions & 5 deletions docs/llmservice/models/claude-sonnet-5.md
Original file line number Diff line number Diff line change
Expand Up @@ -40,11 +40,10 @@ Claude Sonnet 5, released by Anthropic on June 30, 2026, is the next generation

## Credits Usage

| Model | Pricing Period | Input (Credits/Token) | 5m Cache Write (Credits/Token) | 1h Cache Write (Credits/Token) | Cache Read (Credits/Token) | Output (Credits/Token) | Web Search (Credits/Use) |
| :--- | :--- | --------------------: | -----------------------------: | -----------------------------: | -------------------------: | ---------------------: | -----------------------: |
| **Claude Sonnet 5** | Through Aug 31, 2026 | `2.00` | `2.50` | `4.00` | `0.20` | `10.00` | `10,000` |
| **Claude Sonnet 5** | From Sep 1, 2026 | `3.00` | `3.75` | `6.00` | `0.30` | `15.00` | `10,000` |
| Model | Input (Credits/Token) | 5m Cache Write (Credits/Token) | 1h Cache Write (Credits/Token) | Cache Read (Credits/Token) | Output (Credits/Token) | Web Search (Credits/Use) |
| :--- | --------------------: | -----------------------------: | -----------------------------: | -------------------------: | ---------------------: | -----------------------: |
| **Claude Sonnet 5** | `2.00` | `2.50` | `4.00` | `0.20` | `10.00` | `10,000` |

:::info Pricing note
The main pricing table shows the currently effective standard reference price. For Claude Sonnet 5, the current standard reference price applies through August 31, 2026. Prices shown in the documentation are B.AI standard reference prices for base billing purposes. B.AI may provide lower actual usage costs through top-up bonuses and account benefits. Specific prices, bonus Credits, and account benefits are subject to the platform display and final billing records.
The main pricing table shows the current standard reference price. Prices shown in the documentation are B.AI standard reference prices for base billing purposes. B.AI may provide lower actual usage costs through top-up bonuses and account benefits. Specific prices, bonus Credits, and account benefits are subject to the platform display and final billing records.
:::
2 changes: 1 addition & 1 deletion docs/llmservice/models/glm-5-1.md
Original file line number Diff line number Diff line change
Expand Up @@ -41,7 +41,7 @@ GLM-5.1 is an open-source flagship AI model developed by Z.ai, formerly Zhipu AI

| Model | Input (Credits/Token) | Cache Write (Credits/Token) | Cache Read (Credits/Token) | Output (Credits/Token) | Web Search (Credits/Use) | Billing Notes |
| :--- | --------------------: | --------------------------: | -------------------------: | ---------------------: | -----------------------: | :--- |
| **GLM-5.1** | `1.40` | `1.40` | `0.26` | `4.40` | `-` | - |
| **GLM-5.1** | `1.40` | `1.40` | `0.28` | `4.40` | `-` | - |

:::info Pricing note
Prices shown in the documentation are B.AI standard reference prices for base billing purposes. B.AI may provide lower actual usage costs through top-up bonuses and account benefits. Specific prices, bonus Credits, and account benefits are subject to the platform display and final billing records.
Expand Down
6 changes: 3 additions & 3 deletions docs/llmservice/models/glm-5-2.md
Original file line number Diff line number Diff line change
Expand Up @@ -4,12 +4,12 @@

GLM-5.2 is a GLM-family text foundation model developed by Z.AI and released on June 16, 2026. It is positioned for long-horizon coding and engineering tasks, with a 1M-token context window, 128K maximum output, and a `reasoning_effort` control for adjusting reasoning depth.

:::tip Limited-time offer: GLM-5.2 API calls at 40% off
:::tip Limited-time offer: GLM-5.2 at 40% off
Offer starts August 12, 2026.

**Eligibility:** This offer applies only to GLM-5.2 requests made through the B.AI API. Non-API usage is not eligible.
**Eligibility:** This offer applies to GLM-5.2 requests made through the B.AI API and B.AI web app.

For a limited time, eligible API requests are billed at 60% of the standard reference price: Input `0.84`, Cache Write `0.84`, Cache Read `0.168`, and Output `2.64` Credits/Token.
For a limited time, eligible requests are billed at 60% of the standard reference price: Input `0.84`, Cache Write `0.84`, Cache Read `0.168`, and Output `2.64` Credits/Token.

The pricing table on this page continues to show standard reference prices. Offer end time, eligibility, actual settlement price, and final billing are subject to the platform display and final billing records.
:::
Expand Down
3 changes: 1 addition & 2 deletions docs/llmservice/models/gpt-5-6-luna.md
Original file line number Diff line number Diff line change
Expand Up @@ -42,8 +42,7 @@ GPT-5.6 Luna is OpenAI's cost-oriented GPT-5.6 tier, made generally available on

| Context | Input (Credits/Token) | Cache Write (Credits/Token) | Cache Read (Credits/Token) | Output (Credits/Token) | Web Search (Credits/Use) | Billing Notes |
| :--- | --------------------: | --------------------------: | -------------------------: | ---------------------: | -----------------------: | :--- |
| Short context | `1.00` | `1.25` | `0.10` | `6.00` | `10,000` | Standard GPT-5.6 Luna pricing |
| Long context (>272K input tokens) | `2.00` | `2.50` | `0.20` | `9.00` | `10,000` | Long-context pricing tier |
| Short context | `0.20` | `0.25` | `0.02` | `1.20` | `10,000` | Standard GPT-5.6 Luna pricing |

:::info Pricing note
Prices shown in the documentation are B.AI standard reference prices for base billing purposes. B.AI may provide lower actual usage costs through top-up bonuses and account benefits. Specific prices, bonus Credits, and account benefits are subject to the platform display and final billing records.
Expand Down
3 changes: 1 addition & 2 deletions docs/llmservice/models/gpt-5-6-terra.md
Original file line number Diff line number Diff line change
Expand Up @@ -42,8 +42,7 @@ GPT-5.6 Terra is OpenAI's balanced GPT-5.6 tier, made generally available on Jul

| Context | Input (Credits/Token) | Cache Write (Credits/Token) | Cache Read (Credits/Token) | Output (Credits/Token) | Web Search (Credits/Use) | Billing Notes |
| :--- | --------------------: | --------------------------: | -------------------------: | ---------------------: | -----------------------: | :--- |
| Short context | `2.50` | `3.125` | `0.25` | `15.00` | `10,000` | Standard GPT-5.6 Terra pricing |
| Long context (>272K input tokens) | `5.00` | `6.25` | `0.50` | `22.50` | `10,000` | Long-context pricing tier |
| Short context | `2.00` | `2.50` | `0.20` | `12.00` | `10,000` | Standard GPT-5.6 Terra pricing |

:::info Pricing note
Prices shown in the documentation are B.AI standard reference prices for base billing purposes. B.AI may provide lower actual usage costs through top-up bonuses and account benefits. Specific prices, bonus Credits, and account benefits are subject to the platform display and final billing records.
Expand Down
2 changes: 1 addition & 1 deletion docs/llmservice/models/kimi-k2.5.md
Original file line number Diff line number Diff line change
Expand Up @@ -41,7 +41,7 @@ With its **256K ultra-long context window**, multimodal understanding, and advan

| Model | Input (Credits/Token) | Cache Write (Credits/Token) | Cache Read (Credits/Token) | Output (Credits/Token) | Web Search (Credits/Use) | Billing Notes |
| :--- | --------------------: | --------------------------: | -------------------------: | ---------------------: | -----------------------: | :--- |
| **Kimi K2.5** | `0.59` | `0.59` | `0.177` | `3.00` | `-` | - |
| **Kimi K2.5** | `0.59` | `0.59` | `0.10` | `3.00` | `-` | - |

:::info Pricing note
Prices shown in the documentation are B.AI standard reference prices for base billing purposes. B.AI may provide lower actual usage costs through top-up bonuses and account benefits. Specific prices, bonus Credits, and account benefits are subject to the platform display and final billing records.
Expand Down
2 changes: 1 addition & 1 deletion docs/llmservice/models/kimi-k2.6.md
Original file line number Diff line number Diff line change
Expand Up @@ -43,7 +43,7 @@ Kimi K2.6 is a Moonshot AI model available on B.AI for multimodal understanding,

| Model | Input (Credits/Token) | Cache Write (Credits/Token) | Cache Read (Credits/Token) | Output (Credits/Token) | Web Search (Credits/Use) | Billing Notes |
| :--- | --------------------: | --------------------------: | -------------------------: | ---------------------: | -----------------------: | :--- |
| **Kimi K2.6** | `0.95` | `0.95` | `0.16` | `4.00` | `-` | - |
| **Kimi K2.6** | `0.95` | `0.95` | `0.1615` | `4.00` | `-` | - |

:::info Pricing note
Prices shown in the documentation are B.AI standard reference prices for base billing purposes. B.AI may provide lower actual usage costs through top-up bonuses and account benefits. Specific prices, bonus Credits, and account benefits are subject to the platform display and final billing records.
Expand Down
24 changes: 15 additions & 9 deletions docs/llmservice/pricing-and-usage.md
Original file line number Diff line number Diff line change
Expand Up @@ -12,12 +12,18 @@ The platform uses a unified Credits system to measure and settle usage across al

**Model pricing:** Different AI models have different pricing based on their capabilities and compute cost. In general, more capable models consume more Credits. Cache-enabled requests may incur separate cache write and cache read usage. Web search incurs an additional per-use charge. Some models do not support web search and are marked with `-`. See the table below for detailed pricing:

:::tip 🎁 Limited-time offer: GLM-5.2 API calls at 40% off
:::caution Planned DeepSeek API Pricing Adjustment
Due to a recent pricing adjustment by DeepSeek, B.AI plans to make a corresponding adjustment to pricing for DeepSeek API services. Please plan your usage accordingly.

The adjustment scope, effective date, and final prices are subject to the formal announcement and platform display.
:::

:::tip 🎁 Limited-time offer: GLM-5.2 at 40% off
Offer starts August 12, 2026.

**Eligibility:** This offer applies only to GLM-5.2 requests made through the B.AI API. Non-API usage is not eligible.
**Eligibility:** This offer applies to GLM-5.2 requests made through the B.AI API and B.AI web app.

For a limited time, eligible API requests are billed at 60% of the standard reference price: Input `0.84`, Cache Write `0.84`, Cache Read `0.168`, and Output `2.64` Credits/Token.
For a limited time, eligible requests are billed at 60% of the standard reference price: Input `0.84`, Cache Write `0.84`, Cache Read `0.168`, and Output `2.64` Credits/Token.

The table below continues to show standard reference prices. Offer end time, eligibility, actual settlement price, and final billing are subject to the platform display and final billing records.
:::
Expand All @@ -27,20 +33,20 @@ The table below continues to show standard reference prices. Offer end time, eli
| MiniMax M3 | 0.30 | 0.30 | 0.06 | 1.20 | - |
| MiniMax M2.7 | 0.30 | 0.375 | 0.06 | 1.20 | - |
| Kimi K3 | 3.00 | 3.00 | 0.30 | 15.00 | - |
| Kimi K2.6 | 0.95 | 0.95 | 0.16 | 4.00 | - |
| Kimi K2.5 | 0.59 | 0.59 | 0.177 | 3.00 | - |
| Kimi K2.6 | 0.95 | 0.95 | 0.1615 | 4.00 | - |
| Kimi K2.5 | 0.59 | 0.59 | 0.10 | 3.00 | - |
| Qwen3.8-Max | 2.00 | 2.00 | 0.25 | 6.00 | - |
| Qwen3.7-Max | 1.65 | 1.65 | 0.33 | 4.951 | - |
| Qwen3.6-27B | 0.19 | 0.19 | 0.019 | 2.99 | - |
| GLM-5.2 | 1.40 | 1.40 | 0.28 | 4.40 | - |
| GLM-5.1 | 1.40 | 1.40 | 0.26 | 4.40 | - |
| GLM-5.1 | 1.40 | 1.40 | 0.28 | 4.40 | - |
| DeepSeek V3.2 | 0.29 | 0.29 | 0.145 | 0.44 | - |
| DeepSeek V4 Flash | 0.28 | 0.28 | 0.0056 | 0.56 | - |
| DeepSeek V4 Pro | 0.87 | 0.87 | 0.0087 | 1.74 | - |
| Grok 4.5 | 2.00 | 2.00 | 0.30 | 6.00 | - |
| GPT-5.6 Sol | 5.00 | 6.25 | 0.50 | 30.00 | 10,000 |
| GPT-5.6 Terra | 2.50 | 3.125 | 0.25 | 15.00 | 10,000 |
| GPT-5.6 Luna | 1.00 | 1.25 | 0.10 | 6.00 | 10,000 |
| GPT-5.6 Terra | 2.00 | 2.50 | 0.20 | 12.00 | 10,000 |
| GPT-5.6 Luna | 0.20 | 0.25 | 0.02 | 1.20 | 10,000 |
| GPT-5.4 | 2.50 | 2.50 | 0.25 | 15.00 | 10,000 |
| GPT-5.5 | 5.00 | 5.00 | 0.50 | 30.00 | 10,000 |
| GPT-5.5 Instant | 5.00 | 5.00 | 0.50 | 30.00 | 10,000 |
Expand Down Expand Up @@ -80,7 +86,7 @@ Auto Mode is not a separately billable model. Each request is billed based on th

### Cache Pricing Notes

- **Cache Write:** The cost when tokens are first written into the prompt cache. Most providers charge no premium and use the same rate as standard input pricing. Claude models and GPT-5.6 models apply a 25% premium on cache writes; GPT-5.6 long-context cache write rates follow the corresponding long-context input tier.
- **Cache Write:** The cost when tokens are first written into the prompt cache. Most providers charge no premium and use the same rate as standard input pricing. Claude models and GPT-5.6 models apply a 25% premium on cache writes.
- **Cache Read:** The discounted cost when cached tokens are reused in subsequent requests.
- **Example with Caching:** If you use Claude Sonnet 4.6 with 1000 tokens cached:
- First request (cache write): `1000 × 3.75 = 3,750 credits`
Expand Down
Original file line number Diff line number Diff line change
Expand Up @@ -40,11 +40,10 @@ Claude Sonnet 5 是 Anthropic 于 2026 年 6 月 30 日发布的新一代 Sonnet

## 积分消耗

| 模型名称 | 价格周期 | 输入 (Credits/Token) | 5m Cache Write (Credits/Token) | 1h Cache Write (Credits/Token) | Cache Read (Credits/Token) | 输出 (Credits/Token) | 网页搜索(Credits/次) |
| :--- | :--- | --------------------: | -----------------------------: | -----------------------------: | -------------------------: | -------------------: | ---------------------: |
| **Claude Sonnet 5** | 截至 2026 年 8 月 31 日 | `2.00` | `2.50` | `4.00` | `0.20` | `10.00` | `10,000` |
| **Claude Sonnet 5** | 2026 年 9 月 1 日起 | `3.00` | `3.75` | `6.00` | `0.30` | `15.00` | `10,000` |
| 模型名称 | 输入 (Credits/Token) | 5m Cache Write (Credits/Token) | 1h Cache Write (Credits/Token) | Cache Read (Credits/Token) | 输出 (Credits/Token) | 网页搜索(Credits/次) |
| :--- | --------------------: | -----------------------------: | -----------------------------: | -------------------------: | -------------------: | ---------------------: |
| **Claude Sonnet 5** | `2.00` | `2.50` | `4.00` | `0.20` | `10.00` | `10,000` |

:::info 价格说明
价格总表展示当前生效的标准参考价。Claude Sonnet 5 当前标准参考价适用至 2026 年 8 月 31 日。文档价格为 B.AI 平台模型标准参考价,仅供基础计费说明使用。B.AI 可能会通过充值赠送及账户权益等方式,为用户提供更低的实际使用成本。具体价格、赠送积分及账户权益请以平台页面展示及最终账单为准。
价格总表展示当前标准参考价。文档价格为 B.AI 平台模型标准参考价,仅供基础计费说明使用。B.AI 可能会通过充值赠送及账户权益等方式,为用户提供更低的实际使用成本。具体价格、赠送积分及账户权益请以平台页面展示及最终账单为准。
:::
Original file line number Diff line number Diff line change
Expand Up @@ -41,7 +41,7 @@ GLM-5.1 是由 Z.ai 开发的开源旗舰 AI 模型。Z.ai 前身为智谱 AI,

| 模型名称 | 输入 (Credits/Token) | Cache Write (Credits/Token) | Cache Read (Credits/Token) | 输出 (Credits/Token) | 网页搜索(Credits/次) | 计费说明 |
| :--- | --------------------: | --------------------------: | -------------------------: | -------------------: | ---------------------: | :--- |
| **GLM-5.1** | `1.40` | `1.40` | `0.26` | `4.40` | `-` | - |
| **GLM-5.1** | `1.40` | `1.40` | `0.28` | `4.40` | `-` | - |

:::info 价格说明
文档价格为 B.AI 平台模型标准参考价,仅供基础计费说明使用。B.AI 可能会通过充值赠送及账户权益等方式,为用户提供更低的实际使用成本。具体价格、赠送积分及账户权益请以平台页面展示及最终账单为准。
Expand Down
Original file line number Diff line number Diff line change
Expand Up @@ -4,12 +4,12 @@

GLM-5.2 是由 Z.AI 开发的 GLM 系列文本基础模型,于 2026 年 6 月 16 日发布。该模型面向长周期代码和工程任务,支持 1M tokens 上下文窗口、128K 最大输出,并提供 `reasoning_effort` 参数用于调整推理深度。

:::tip 限时活动:GLM-5.2 API 调用 6 折
:::tip 限时活动:GLM-5.2 6 折
活动开始时间:2026 年 8 月 12 日。

**适用范围:** 本活动仅适用于通过 B.AI API 发起的 GLM-5.2 调用;非 API 使用不参与本活动
**适用范围:** 本活动适用于通过 B.AI API 和 B.AI 网页端发起的 GLM-5.2 调用。

限时活动期间,符合条件的 API 调用按标准参考价的 60% 结算:输入 `0.84`、缓存写入 `0.84`、缓存读取 `0.168`、输出 `2.64` Credits/Token。
限时活动期间,符合条件的调用按标准参考价的 60% 结算:输入 `0.84`、缓存写入 `0.84`、缓存读取 `0.168`、输出 `2.64` Credits/Token。

本页价格表继续展示标准参考价;活动结束时间、适用规则、实际结算价格及最终账单以平台页面展示为准。
:::
Expand Down
Original file line number Diff line number Diff line change
Expand Up @@ -42,8 +42,7 @@ GPT-5.6 Luna 是 OpenAI GPT-5.6 系列的成本友好层级模型,已于 2026

| 上下文档位 | 输入 (Credits/Token) | Cache Write (Credits/Token) | Cache Read (Credits/Token) | 输出 (Credits/Token) | 网页搜索(Credits/次) | 计费说明 |
| :--- | --------------------: | --------------------------: | -------------------------: | -------------------: | ---------------------: | :--- |
| 短上下文 | `1.00` | `1.25` | `0.10` | `6.00` | `10,000` | GPT-5.6 Luna 标准价格 |
| 长上下文(输入 token >272K) | `2.00` | `2.50` | `0.20` | `9.00` | `10,000` | 长上下文价格档位 |
| 短上下文 | `0.20` | `0.25` | `0.02` | `1.20` | `10,000` | GPT-5.6 Luna 标准价格 |

:::info 价格说明
文档价格为 B.AI 平台模型标准参考价,仅供基础计费说明使用。B.AI 可能会通过充值赠送及账户权益等方式,为用户提供更低的实际使用成本。具体价格、赠送积分及账户权益请以平台页面展示及最终账单为准。
Expand Down
Loading
Loading