> For the complete documentation index, see [llms.txt](https://docs.clore.ai/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://docs.clore.ai/guides/guides_v2-de/erste-schritte/pricing.md).

# GPU-Preise & Verfügbarkeit

Was derzeit tatsächlich auf dem Clore.ai-Marktplatz gelistet ist und was es pro GPU-Stunde kostet

{% hint style="info" %}
**Schnappschuss: 24. August 2026.** Die Zahlen unten stammen aus dem Live-Marktplatz-Eintrag: **2.202 Server / 6.791 GPUs**, davon **1.198 Server waren kostenlos zu mieten** zum Zeitpunkt des Schnappschusses. Die Preise sind **pro GPU und Stunde**, auf Abruf, in USD. Der Marktplatz selbst ist die maßgebliche Quelle — prüfe [clore.ai/marketplace](https://clore.ai/marketplace) bevor du dein Budget planst.
{% endhint %}

Clore.ai rechnet minutengenau ab. Ein hier als `$0,20/Stunde` angegebenen Preis bedeutet, dass du ungefähr `$0.0033` pro Minute Betriebszeit zahlst und nichts, wenn die Bestellung geschlossen ist. Die Marktplatz-Oberfläche kann entweder **pro Stunde** oder **pro Tag** — diese Seite und jeder Leitfaden in dieser Sammlung nennt **pro GPU und Stunde**.

## Was es kostet

Die Spannen sind das 10. bis 90. Perzentil der Server, die im Schnappschuss kostenlos zu mieten waren, also ist das untere Ende das, was ein geduldiger Mieter zahlt, und das obere Ende ein Premium-Angebot.

| GPU                    | VRAM      | $/GPU/h        | Median    | Kostenlose Server |
| ---------------------- | --------- | -------------- | --------- | ----------------- |
| RTX 3060               | 12 GB     | $0,03–0,07     | $0.04     | 32                |
| RTX 3060 Ti            | 8 GB      | $0,03–0,10     | $0.05     | 15                |
| RTX 3070               | 8 GB      | $0,03–0,09     | $0.05     | 54                |
| RTX 3080               | 10 GB     | $0,05–0,19     | $0.08     | 68                |
| RTX 3080 Ti            | 12 GB     | $0,06–0,19     | $0.10     | 48                |
| **RTX 3090**           | **24 GB** | **$0,07–0,21** | **$0.09** | **90**            |
| RTX 4060 Ti            | 8 GB      | $0,04–0,12     | $0.06     | 16                |
| RTX 4070               | 12 GB     | $0,04–0,20     | $0.07     | 24                |
| RTX 4070 Ti            | 12 GB     | $0,07–0,25     | $0.12     | 24                |
| RTX 4070 Ti SUPER      | 16 GB     | $0,08–0,19     | $0.12     | 28                |
| RTX 4080 SUPER         | 16 GB     | $0,09–0,17     | $0.12     | 22                |
| **RTX 4090**           | **24 GB** | **$0,14–0,42** | **$0.20** | **98**            |
| RTX 5060 Ti            | 16 GB     | $0,04–0,10     | $0.06     | 29                |
| RTX 5070               | 12 GB     | $0,07–0,29     | $0.11     | 40                |
| RTX 5070 Ti            | 16 GB     | $0,10–0,25     | $0.13     | 103               |
| RTX 5080               | 16 GB     | $0,12–0,62     | $0.17     | 56                |
| **RTX 5090**           | **32 GB** | **$0,25–0,77** | **$0.29** | **121**           |
| RTX PRO 6000 Blackwell | 96GB      | $0,92–1,38     | $1.21     | 4                 |
| Tesla V100 SXM2        | 32 GB     | $0,05–0,06     | $0.05     | selten            |
| RTX A6000              | 48GB      | \~$0.36        | $0.36     | insgesamt 2 GPUs  |
| H100 NVL               | 94GB      | \~$1.04        | $1.04     | insgesamt 4 GPUs  |
| L40S                   | 48GB      | \~$0.50        | $0.50     | insgesamt 1 GPU   |

Die Flotte besteht überwiegend aus Consumer-Hardware. **Die RTX-50er-Serie macht bereits 19,5 % aller gelisteten GPUs aus** (1.321 Karten), die 30er-Serie ist die Volumenklasse, und die 40er-Serie liegt dazwischen.

## Was nicht auf dem Marktplatz ist

{% hint style="warning" %}
**A100, H200 und B200 sind heute nicht auf dem Marktplatz gelistet.** Nicht „selten“ — null. Es gibt 4 H100-NVL-GPUs auf 2 Servern, 2 A6000 und 1 L40S. Wenn ein Leitfaden eine Arbeitslast mit „8× H100“ bemisst, kann diese Konfiguration hier nicht auf Abruf gemietet werden.
{% endhint %}

Rechenzentrumskapazität wird separat verkauft als [Bare Metal](https://clore.ai/bare-metal) — dedizierte H100/H200/B200-Knoten auf Anfrage, vertraglich statt minutengenau. Für alles andere ist die praktische Obergrenze ein Consumer-Rig mit mehreren GPUs.

**Die heute gelisteten größten Rigs:**

| Rig                            | Gesamt-VRAM | Rig $/Stunde |
| ------------------------------ | ----------- | ------------ |
| 4× RTX PRO 6000 Blackwell 96GB | 380 GB      | \~$5.00      |
| 11× RTX 5090 32GB              | 341 GB      | \~$5.42      |
| 10× RTX 5090 32GB              | 310 GB      | \~$5.00      |
| 8× Tesla V100 SXM2 32GB        | 256GB       | \~$0.40      |
| 8× RTX 4090 24GB               | 192 GB      | \~$1.60      |

Verfügbarkeit nach Größe, aus dem Schnappschuss:

| Du brauchst                               | Gelistete Server | Zum Zeitpunkt des Schnappschusses frei |
| ----------------------------------------- | ---------------- | -------------------------------------- |
| ≥ 24 GB (eine Karte der 3090/4090-Klasse) | 1,299            | 677                                    |
| ≥ 48 GB                                   | 653              | 322                                    |
| ≥ 80 GB                                   | 328              | 153                                    |
| ≥ 96 GB                                   | 182              | 81                                     |
| ≥ 192 GB                                  | 54               | 14                                     |
| ≥ 242 GB                                  | 47               | 9                                      |
| ≥ 352 GB                                  | 2                | 2                                      |

## Was passt

Grobe Zuordnung von Modellgewichtsgröße zu dem, was du mieten solltest. Quantisierte GGUF-Größen sind in der Praxis die entscheidenden.

| Gewichte auf der Festplatte | Miete das                            | Beispiel                                                                    |
| --------------------------- | ------------------------------------ | --------------------------------------------------------------------------- |
| ≤ 16 GB                     | 1× RTX 3090 / 4090                   | Qwen3.8-27B Q4, Mistral Small 3.1, SDXL, die meisten TTS                    |
| 16–30 GB                    | 1× RTX 5090                          | Qwen3.8-27B FP8, 32B Q4/Q6, FLUX in hoher Auflösung                         |
| 30–60 GB                    | 2× RTX 4090 / 5090 oder RTX PRO 6000 | Mistral Small 4 Q3–Q4, 70B Q4                                               |
| 60–192 GB                   | 4–8× RTX 5090                        | DeepSeek V4-Flash FP8, MiniMax M3 Q2                                        |
| 192–380 GB                  | 8–11× RTX 5090, 4× RTX PRO 6000      | Nemotron 3 Ultra Q2, GLM-5.2 Q2                                             |
| > 380 GB                    | Nicht auf dem Marktplatz             | Kimi K3 bei jeder Quantisierung → [Bare Metal](https://clore.ai/bare-metal) |

## Spot vs. On-Demand

Clore.ai bietet zwei Bestelltypen an: **On-Demand** zum Festpreis des Hosts, und **Spot**, wo du bietest und von einem höheren Gebot überboten werden kannst.

{% hint style="info" %}
**Spot ist nicht automatisch günstiger.** Im Schnappschuss, **63 % der Server hatten Spot genau zum On-Demand-Preis gelistet**. Von den 34 %, die günstiger waren, lag der Median-Rabatt bei **13.4%**. Einige wenige waren teurer als On-Demand. Prüfe das tatsächliche Paar auf der Serverkarte, statt einen Rabatt anzunehmen.
{% endhint %}

Spot lohnt sich weiterhin für unterbrechbare Arbeit — Batch-Rendering, Datenvorverarbeitung, lange Fine-Tunes mit Checkpointing — weil du auch *über* über den angegebenen Spot-Preis bieten kannst, um eine Maschine zu sichern, die andere Bieter wollen.

## Teilweise GPU-Miete

Etwa **60 % der gelisteten Server (1.331 von 2.202) erlauben Teilmiete**: Du nimmst 1 oder 2 GPUs aus einem 8-GPU-Rig und zahlst nur für diese, statt die ganze Maschine zu mieten. Für Single-GPU-Workloads ist das normalerweise der günstigste Weg zu einer guten Karte, weil große Rigs oft pro GPU günstiger sind als kleine. Siehe [Teilweise GPU-Miete](/guides/guides_v2-de/erste-schritte/partial-gpu-rental.md).

## Echte Kostenbeispiele

Berechnet aus den obigen Medianpreisen, bei minutengenauer Abrechnung.

| Job                                               | GPU                     | Zeit      | Kosten      |
| ------------------------------------------------- | ----------------------- | --------- | ----------- |
| 1.000 SDXL-Bilder                                 | RTX 3090 @ $0,09/Stunde | \~2 Std.  | **\~$0.18** |
| 1 Mio. Tokens durch ein 27B-Q4-Modell             | RTX 4090 @ $0,20/Stunde | \~20 Min. | **\~$0.07** |
| 10 × 6-sekündige Videoclips                       | RTX 4090 @ $0,20/Stunde | \~20 Min. | **\~$0.07** |
| LoRA-Finetuning, 7B, 1 Epoche                     | RTX 4090 @ $0,20/Stunde | \~3 Std.  | **\~$0.60** |
| Whisper-Transkription, 10 Std. Audio              | RTX 3060 @ $0,04/Stunde | \~40 Min. | **\~$0.03** |
| TTS-API läuft einen Monat lang 24/7               | RTX 3060 @ $0,04/Stunde | 730 Std.  | **\~$29**   |
| Immer aktiver 27B-Chat-Endpunkt, einen Monat lang | RTX 5090 @ $0,29/Stunde | 730 Std.  | **\~$212**  |

## Preise selbst prüfen

Der Marktplatz-Eintrag ist öffentlich. Die Preise im Feed sind **pro Server und Tag** in USD — teile durch 24 für den Stundenpreis und durch die GPU-Anzahl für den Preis pro GPU:

```bash
curl -s https://clore.ai/webapi/marketplace/servers \

  | python3 -c '
import json,sys,re
for s in json.load(sys.stdin)["all_servers"]:
    gpu = s["specs"]["gpu"]
    n = int(re.match(r"^(\d+)x", gpu).group(1)) if re.match(r"^(\d+)x", gpu) else 1
    usd = (s["price"]["usd"] or {}).get("on_demand_usd") or 0
    if usd and "4090" in gpu and not s["rented"]:
        print(f"{gpu:32} ${usd/n/24:.3f}/GPU/hr")
'
```

## Nächste Schritte

* [GPU-Vergleich](/guides/guides_v2-de/erste-schritte/gpu-comparison.md) — welche Karte für welche Arbeitslast
* [Teilweise GPU-Miete](/guides/guides_v2-de/erste-schritte/partial-gpu-rental.md) — miete 1 GPU aus einem Rig
* [Kostenrechner](/guides/guides_v2-de/erste-schritte/cost-calculator.md) — budgetiere einen bestimmten Auftrag
* [CUDA- & PyTorch-Kompatibilität](/guides/guides_v2-de/erste-schritte/cuda-pytorch-compatibility.md) — wähle ein Image, das auf der gemieteten Karte läuft

### Links

* [Marktplatz](https://clore.ai/marketplace) · [Bare Metal](https://clore.ai/bare-metal)
* [Miete eine RTX 4090](https://clore.ai/rent-4090.html) · [Miete eine RTX 5090](https://clore.ai/rent-5090.html)
* [GPU-Cloud-Preisvergleich — Clore.ai vs. große Anbieter](https://blog.clore.ai/gpu-cloud-pricing-comparison/)


---

# Agent Instructions
This documentation is published with GitBook. GitBook is the documentation platform designed so that both humans and AI agents can read, navigate, and reason over technical content effectively. Learn more at gitbook.com.

## Querying This Documentation
If you need additional information that is not directly available in this page, you can query the documentation dynamically by asking a question.

Perform an HTTP GET request on the current page URL with the `ask` query parameter, and the optional `goal` query parameter:

```
GET https://docs.clore.ai/guides/guides_v2-de/erste-schritte/pricing.md?ask=<question>&goal=<endgoal>
```

`ask` is the immediate question: it should be specific, self-contained, and written in natural language.
`goal` is optional and describes the broader end goal you are ultimately trying to accomplish on behalf of the user. GitBook uses it to tailor the answer towards what is most useful for that goal.

The response will contain a direct answer to the question and relevant excerpts and sources from the documentation.

Use this mechanism when the answer is not explicitly present in the current page, you need clarification or additional context, or you want to retrieve related documentation sections.
