Skip to main content

Deploy Magistral 24B

Text & Chat

Magistral is a 24B parameter model specialized in legal and financial analysis. It provides transparent reasoning chains suited for compliance review, contract analysis, and regulatory research.

Deploy Magistral 24B in minutes

Starting at $0.66/hr on dedicated GPU

Specifications

ModelGPUVRAMPriceAction
Magistral 24B
24B (Reasoning)
RTX A600048 GB$0.66/hrDeploy

Prices include the service fee. Charges follow actual running time.

Requirements

ModelPilot assigns a 48GB cloud GPU to this deployment. Actual local VRAM requirements vary with model variant, precision, quantization, resolution, and workflow settings.

On ModelPilot, deploy on a dedicated cloud GPU (up to 80GB VRAM) starting at $0.66/hr with no setup required.

Includes OpenWebUI chat interface and OpenAI-compatible API endpoint.

Use Cases

  • Legal document analysis
  • Financial report review
  • Compliance and regulatory research
  • Contract summarization

Related Models

Frequently Asked Questions

How much GPU memory is allocated for Magistral 24B?

The listed ModelPilot deployment uses a 48GB cloud GPU. Local memory needs can vary with precision, quantization, and workflow settings.

How much does it cost to run Magistral 24B?

Starting at $0.66/hr on a dedicated GPU. Charges are calculated from actual running time, with auto-stop when credits run out.

How long does Magistral 24B take to deploy?

Text models typically deploy in 5–15 minutes including model download.

Can I run Magistral 24B on my local GPU?

It depends on the selected variant, precision, quantization, and workflow settings. Compare the variants below with your available VRAM; the table shows ModelPilot's cloud GPU allocation, not a universal local minimum.

Ready to deploy Magistral 24B?

Pick your GPU and have it running in minutes. No infrastructure setup required.