Skip to main content

Deploy Mistral

Text & Chat

Mistral AI develops language models for chat, document processing and structured text generation. This catalog includes Ministral 8B and Mistral Nemo 12B; compare each variant’s requirements and license before use.

Set up a Mistral workspace

Starting at $0.66/hr on dedicated GPU

Model configurations (3)

ModelGPUVRAMPriceAction
Ministral 8B
8B (Fast)
L424 GB$0.66/hrDeploy
Mistral Nemo 12B
Nemo (12B)
L424 GB$0.66/hrDeploy
Mistral 7B
7B (Legacy)
L424 GB$0.66/hrDeploy

Prices include the service fee. Charges follow actual running time.

Requirements

The listed configuration specifies a 24GB cloud GPU. Actual local VRAM requirements vary with model variant, precision, quantization, resolution, and workflow settings.

On ModelPilot, deploy on a dedicated cloud GPU (up to 80GB VRAM) starting at $0.66/hr. Review the template’s required setup steps and settings before launch.

Includes OpenWebUI chat interface and OpenAI-compatible API endpoint.

Use Cases

  • Fast inference applications
  • Multilingual text processing
  • Document analysis (128K context)
  • Structured output generation

Related Models

Frequently Asked Questions

How much GPU memory is allocated for Mistral?

The listed ModelPilot deployment uses a 24GB cloud GPU. Local memory needs can vary with precision, quantization, and workflow settings.

How much does it cost to run Mistral?

Starting at $0.66/hr on a dedicated GPU. Charges are calculated from actual running time, with auto-stop when credits run out.

How long does Mistral take to deploy?

Startup time depends on model downloads, container setup, cached files and GPU availability. Review the estimate for your selected template; it is not a guaranteed time to a finished result. Dedicated GPU time is billed while allocated, including startup.

Can I run Mistral on my local GPU?

It depends on the selected variant, precision, quantization, and workflow settings. Compare the variants below with your available VRAM; the table shows ModelPilot's cloud GPU allocation, not a universal local minimum.

Ready to deploy Mistral?

Choose a GPU, review the estimated rate, and check the template’s setup requirements before launching.