Deploy GPT-OSS
Text & ChatGPT-OSS is OpenAI's open-weight model family. The 20B model offers native function calling, while the 120B flagship provides visible chain-of-thought reasoning comparable to GPT-4.
Deploy GPT-OSS in minutes
Starting at $0.53/hr on dedicated GPU
Available Variants (2)
| Model | GPU | VRAM | Price | Action |
|---|---|---|---|---|
GPT-OSS 20B Medium (20B) | L4 | 24 GB | $0.53/hr | Deploy |
GPT-OSS 120B Large (120B) | A100 80GB PCIe | 80 GB | $1.85/hr | Deploy |
Prices include the service fee. Charges follow actual running time.
Requirements
ModelPilot assigns 24–80GB cloud GPUs across the listed variants. Actual local VRAM requirements vary with model variant, precision, quantization, resolution, and workflow settings.
On ModelPilot, deploy on a dedicated cloud GPU (up to 80GB VRAM) starting at $0.53/hr with no setup required.
Use Cases
- ✓Function calling and tool use
- ✓Chain-of-thought reasoning
- ✓AI agent development
- ✓Enterprise deployments
Related Models
Frequently Asked Questions
How much GPU memory is allocated for GPT-OSS?
The listed ModelPilot variants use 24–80GB cloud GPUs. Local memory needs vary with the variant, precision, quantization, and workflow settings.
How much does it cost to run GPT-OSS?
Starting at $0.53/hr on a dedicated GPU. Charges are calculated from actual running time, with auto-stop when credits run out.
How long does GPT-OSS take to deploy?
Text models typically deploy in 5–15 minutes including model download.
Can I run GPT-OSS on my local GPU?
It depends on the selected variant, precision, quantization, and workflow settings. Compare the variants below with your available VRAM; the table shows ModelPilot's cloud GPU allocation, not a universal local minimum.
Ready to deploy GPT-OSS?
Pick your GPU and have it running in minutes. No infrastructure setup required.