> For the complete documentation index, see [llms.txt](https://docs.clore.ai/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://docs.clore.ai/guides/guides_v2-zh/ru-men-zhi-nan/pricing.md).

# GPU 价格与可用性

Clore.ai 市场当前实际列出了什么，以及每个 GPU 小时的价格

{% hint style="info" %}
**快照：2026年8月24日。** 以下数字来自实时市场列表： **2,202 台服务器 / 6,791 块 GPU**，其中 **1,198 台服务器在快照时可供租用** ，价格为 **每 GPU 每小时**，按需，以美元计。市场本身就是权威来源——请先查看 [clore.ai/marketplace](https://clore.ai/marketplace) 再做预算。
{% endhint %}

Clore.ai 按分钟计费。这里显示为 `$0.20/hr` 的价格，意味着你大约支付 `$0.0033` 每分钟运行时间的费用，而订单关闭后则不收取费用。市场界面可以显示为 **每小时** 或 **每天** ——本页以及本系列中的每一篇指南引用的都是 **每 GPU 每小时**.

## 成本是多少

范围为快照中可供出租服务器的第10到第90百分位，因此低端是耐心租户会支付的价格，高端是溢价挂牌价。

| GPU                    | VRAM     | 美元/GPU/小时      | 中位数       | 空闲服务器      |
| ---------------------- | -------- | -------------- | --------- | ---------- |
| RTX 3060               | 12GB     | $0.03–0.07     | $0.04     | 32         |
| RTX 3060 Ti            | 8GB      | $0.03–0.10     | $0.05     | 15         |
| RTX 3070               | 8GB      | $0.03–0.09     | $0.05     | 54         |
| RTX 3080               | 10GB     | $0.05–0.19     | $0.08     | 68         |
| RTX 3080 Ti            | 12GB     | $0.06–0.19     | $0.10     | 48         |
| **RTX 3090**           | **24GB** | **$0.07–0.21** | **$0.09** | **90**     |
| RTX 4060 Ti            | 8GB      | $0.04–0.12     | $0.06     | 16         |
| RTX 4070               | 12GB     | $0.04–0.20     | $0.07     | 24         |
| RTX 4070 Ti            | 12GB     | $0.07–0.25     | $0.12     | 24         |
| RTX 4070 Ti SUPER      | 16GB     | $0.08–0.19     | $0.12     | 28         |
| RTX 4080 SUPER         | 16GB     | $0.09–0.17     | $0.12     | 22         |
| **RTX 4090**           | **24GB** | **$0.14–0.42** | **$0.20** | **98**     |
| RTX 5060 Ti            | 16GB     | $0.04–0.10     | $0.06     | 29         |
| RTX 5070               | 12GB     | $0.07–0.29     | $0.11     | 40         |
| RTX 5070 Ti            | 16GB     | $0.10–0.25     | $0.13     | 103        |
| RTX 5080               | 16GB     | $0.12–0.62     | $0.17     | 56         |
| **RTX 5090**           | **32GB** | **$0.25–0.77** | **$0.29** | **121**    |
| RTX PRO 6000 Blackwell | 96GB     | $0.92–1.38     | $1.21     | 4          |
| Tesla V100 SXM2        | 32GB     | $0.05–0.06     | $0.05     | 稀缺         |
| RTX A6000              | 48GB     | \~$0.36        | $0.36     | 总共 2 块 GPU |
| H100 NVL               | 94GB     | \~$1.04        | $1.04     | 总共 4 块 GPU |
| L40S                   | 48GB     | \~$0.50        | $0.50     | 总共 1 块 GPU |

这支机群几乎清一色是消费级芯片。 **RTX 50 系列已占所有列出 GPU 的 19.5%** （1,321 块显卡），30 系列是主力容量档，40 系列则介于两者之间。

## 市场上没有什么

{% hint style="warning" %}
**A100、H200 和 B200 今天未在市场上列出。** 不是“稀有”——而是零。共有 4 块 H100 NVL GPU，分布在 2 台服务器上，另有 2 块 A6000 和 1 块 L40S。如果某份指南把工作负载设为“8× H100”，那么在这里就无法按需租到这种配置。
{% endhint %}

数据中心容量则单独以 [裸金属](https://clore.ai/bare-metal) ——按需提供专用 H100/H200/B200 节点，以合同形式而非按分钟计费。对其他所有情况来说，实际上的上限就是多 GPU 消费级主机。

**今天列出的最大整机：**

| 整机                             | 总显存   | 整机 $/小时 |
| ------------------------------ | ----- | ------- |
| 4× RTX PRO 6000 Blackwell 96GB | 380GB | \~$5.00 |
| 11× RTX 5090 32GB              | 341GB | \~$5.42 |
| 10× RTX 5090 32GB              | 310GB | \~$5.00 |
| 8× Tesla V100 SXM2 32GB        | 256GB | \~$0.40 |
| 8× RTX 4090 24GB               | 192GB | \~$1.60 |

按规模划分的可用性，来自快照：

| 你需要                       | 已列出服务器 | 快照时空闲 |
| ------------------------- | ------ | ----- |
| ≥ 24GB（1 块 3090/4090 级显卡） | 1,299  | 677   |
| ≥ 48GB                    | 653    | 322   |
| ≥ 80GB                    | 328    | 153   |
| ≥ 96GB                    | 182    | 81    |
| ≥ 192GB                   | 54     | 14    |
| ≥ 242GB                   | 47     | 9     |
| ≥ 352GB                   | 2      | 2     |

## 能装下什么

根据模型权重大小大致对应你应该租什么。量化后的 GGUF 大小在实际中最重要。

| 磁盘上的权重    | 租这个                               | 示例                                                   |
| --------- | --------------------------------- | ---------------------------------------------------- |
| ≤ 16GB    | 1× RTX 3090 / 4090                | Qwen3.8-27B Q4、Mistral Small 3.1、SDXL、大多数 TTS        |
| 16–30GB   | 1× RTX 5090                       | Qwen3.8-27B FP8、32B Q4/Q6、高分辨率 FLUX                  |
| 30–60GB   | 2× RTX 4090 / 5090，或 RTX PRO 6000 | Mistral Small 4 Q3–Q4、70B Q4                         |
| 60–192GB  | 4–8× RTX 5090                     | DeepSeek V4-Flash FP8、MiniMax M3 Q2                  |
| 192–380GB | 8–11× RTX 5090、4× RTX PRO 6000    | Nemotron 3 Ultra Q2、GLM-5.2 Q2                       |
| > 380GB   | 不在市场上                             | 任何量化版本的 Kimi K3 → [裸金属](https://clore.ai/bare-metal) |

## 现货 vs 按需

Clore.ai 提供两种订单类型： **按需** 按主机的固定价格，以及 **现货**，在这里你出价，且可能被更高报价抢走。

{% hint style="info" %}
**现货并不一定更便宜。** 在快照中， **63% 的服务器将现货价格标为与按需价格完全相同**。在那 34% 更便宜的服务器中，中位折扣为 **13.4%**。还有少数价格高于按需价。请查看服务器卡片上的实际两项价格，而不要想当然地认为有折扣。
{% endhint %}

现货仍然适合可中断的工作——批量渲染、数据集预处理、带检查点的长时间微调——因为你也可以出价 *高于* 列出的现货价来锁定别人想要的机器。

## 部分 GPU 租赁

大约 **60% 的列出服务器（2,202 台中的 1,331 台）支持部分租赁**：你可以从一台 8 GPU 整机里取出 1 或 2 块 GPU，只为这些付费，而不是租整台机器。对于单 GPU 工作负载，这通常是获取好显卡的最便宜路径，因为大整机按每 GPU 计算的价格往往低于小整机。参见 [部分 GPU 租赁](/guides/guides_v2-zh/ru-men-zhi-nan/partial-gpu-rental.md).

## 真实成本示例

按上面的中位价格、基于按分钟计费计算。

| 任务                         | GPU                 | 时间      | 成本          |
| -------------------------- | ------------------- | ------- | ----------- |
| 1,000 张 SDXL 图像            | RTX 3090 @ $0.09/hr | \~2 小时  | **\~$0.18** |
| 通过 27B Q4 模型处理 100 万 token | RTX 4090 @ $0.20/hr | \~20 分钟 | **\~$0.07** |
| 10 段 6 秒视频片段               | RTX 4090 @ $0.20/hr | \~20 分钟 | **\~$0.07** |
| LoRA 微调，7B，1 个 epoch       | RTX 4090 @ $0.20/hr | \~3 小时  | **\~$0.60** |
| Whisper 转录，10 小时音频         | RTX 3060 @ $0.04/hr | \~40 分钟 | **\~$0.03** |
| 每月 24/7 运行的 TTS API        | RTX 3060 @ $0.04/hr | 730 小时  | **\~$29**   |
| 持续在线的 27B 聊天端点，一个月         | RTX 5090 @ $0.29/hr | 730 小时  | **\~$212**  |

## 自行查看价格

市场列表是公开的。流中的价格是 **每台服务器每天** 以美元计——按 24 除得小时价，再按 GPU 数量除得每 GPU 价格：

```bash
curl -s https://clore.ai/webapi/marketplace/servers \\
  | python3 -c '
import json,sys,re
for s in json.load(sys.stdin)["all_servers"]:
    gpu = s["specs"]["gpu"]
    n = int(re.match(r"^(\d+)x", gpu).group(1)) if re.match(r"^(\d+)x", gpu) else 1
    usd = (s["price"]["usd"] or {}).get("on_demand_usd") or 0
    if usd and "4090" in gpu and not s["rented"]:
        print(f"{gpu:32} ${usd/n/24:.3f}/GPU/hr")
'
```

## 下一步

* [GPU 对比](/guides/guides_v2-zh/ru-men-zhi-nan/gpu-comparison.md) ——哪张卡适合哪种工作负载
* [部分 GPU 租赁](/guides/guides_v2-zh/ru-men-zhi-nan/partial-gpu-rental.md) ——从整机中租 1 块 GPU
* [成本计算器](/guides/guides_v2-zh/ru-men-zhi-nan/cost-calculator.md) ——为特定任务做预算
* [CUDA 与 PyTorch 兼容性](/guides/guides_v2-zh/ru-men-zhi-nan/cuda-pytorch-compatibility.md) ——选择能在你租用的显卡上运行的镜像

### 链接

* [市场](https://clore.ai/marketplace) · [裸金属](https://clore.ai/bare-metal)
* [租用 RTX 4090](https://clore.ai/rent-4090.html) · [租用 RTX 5090](https://clore.ai/rent-5090.html)
* [GPU 云定价对比——Clore.ai 与主要提供商](https://blog.clore.ai/gpu-cloud-pricing-comparison/)


---

# Agent Instructions
This documentation is published with GitBook. GitBook is the documentation platform designed so that both humans and AI agents can read, navigate, and reason over technical content effectively. Learn more at gitbook.com.

## Querying This Documentation
If you need additional information that is not directly available in this page, you can query the documentation dynamically by asking a question.

Perform an HTTP GET request on the current page URL with the `ask` query parameter, and the optional `goal` query parameter:

```
GET https://docs.clore.ai/guides/guides_v2-zh/ru-men-zhi-nan/pricing.md?ask=<question>&goal=<endgoal>
```

`ask` is the immediate question: it should be specific, self-contained, and written in natural language.
`goal` is optional and describes the broader end goal you are ultimately trying to accomplish on behalf of the user. GitBook uses it to tailor the answer towards what is most useful for that goal.

The response will contain a direct answer to the question and relevant excerpts and sources from the documentation.

Use this mechanism when the answer is not explicitly present in the current page, you need clarification or additional context, or you want to retrieve related documentation sections.
