Software Factories at enterprise scale: A federated platform for agentic developmentDownload the free Whitepaper
Article

Running MiniMax M2.5 Locally on NVIDIA DGX Spark

How I ran a 230B parameter open model on desktop hardware with NVIDIA DGX Spark, Unsloth quantization, and llama.cpp at cloud API speed.

Author:Michael Mueller19 Feb 2026
Running MiniMax M2.5 Locally on NVIDIA DGX Spark

Your next read

Subscribe to Our Bi-Weekly AI Native Newsletter

Lessons from our client work, engineering deep-dives, and the AI Native research worth reading.

Bi-weekly. No spam, unsubscribe anytime.

More related Articles