NVIDIA CorporationCrusoe's dedicated Tailored Deployment uses NVIDIA HGX B200 systems and Quantum-2 InfiniBand, a concrete product order driven by the deal.

Crusoe and Thinking Machines Lab have announced a $65 million annual agreement to power inference for the lab's open models on Crusoe Cloud. Under the deal, Thinking Machines Lab will run a range of production workloads, including its Inkling models, GLM 5.2 and 5.3, and its own fine-tuned variants, through Crusoe Managed Inference on a dedicated Tailored Deployment of NVIDIA HGX B200 systems connected with NVIDIA Quantum-2 InfiniBand networking. Crusoe will run and support the cluster directly as a Tailored Deployment, a dedicated, benchmarked, SLA-backed endpoint, giving Thinking Machines Lab the price-performance and headroom to scale without managing its own inference stack. The companies are also looking to expand into batch inference for large-scale synthetic data generation. Thinking Machines Lab joins a roster that has taken Crusoe Managed Inference past $100 million in contracted ARR less than a year from launch.
NVIDIA CorporationCrusoe's dedicated Tailored Deployment uses NVIDIA HGX B200 systems and Quantum-2 InfiniBand, a concrete product order driven by the deal.
Crusoe signs a $65M annual inference deal with Thinking Machines Lab, pushing Managed Inference past $100M contracted ARR.