Crusoe, the industry‘s first vertically integrated AI infrastructure provider, and AI innovator Thinking Machines Lab today announced a $65 million annual agreement to power inference for the lab’s open models on Crusoe Cloud. Through Crusoe Managed Inference, Thinking Machines Lab will serve a range of production workloads – including its Inkling models, GLM 5.2 and 5.3, and its own fine-tuned variants – on a dedicated Tailored Deployment of NVIDIA HGX B200 systems connected with NVIDIA Quantum-2 InfiniBand networking, engineered for high-throughput, cost-efficient serving at scale.
Crusoe will run and support the cluster directly as a Tailored Deployment – a dedicated, benchmarked, SLA-backed endpoint – giving Thinking Machines Lab the price-performance and headroom to scale without managing its own inference stack. The companies are also looking to expand into batch inference for large-scale synthetic data generation, turning inference into a flywheel for research.
Thinking Machines Lab joins a roster that has taken Crusoe Managed Inference past $100 million in contracted ARR less than a year from launch.
Read the full press releasehere.
