Photo by ThisisEngineering on Unsplash. Source: https://unsplash.com/photos/a-woman-sitting-in-front-of-a-laptop-computer-U-Werwf32CM (Unsplash License).

An AI lab with a deep engineering bench rented its serving layer instead of building one. That is the story.

  • Thinking Machines Lab signed a $65 million annual agreement with Crusoe to serve production inference for its open models.
  • The work runs on a dedicated deployment of NVIDIA HGX B200 systems on NVIDIA Quantum-2 InfiniBand networking.
  • Crusoe operates the cluster as a dedicated, benchmarked, SLA-backed endpoint. The lab pays for the throughput.
  • Crusoe Managed Inference passed $100 million in contracted annual recurring revenue less than a year from launch.
  • Training capacity is still bought with capital. Inference capacity is now shopped on price, latency, and reliability.

Read the full analysis here.

By Ivan Tarin

Ivan Tarin is a Principal Product Marketing Manager at SUSE, where he owns go-to-market strategy and positioning for a seven-product cloud-native portfolio spanning Kubernetes, virtualization, storage, security, and observability. A former full-stack developer who shipped production code for enterprise and public-sector clients including U.S. national laboratories, Ivan translates complex infrastructure and AI technology into messaging that lands with developers, platform teams, and enterprise buyers. He has presented at KubeCon, SUSECON, and AWS Developer Week, and is currently pursuing an MS in Artificial Intelligence at the University of Colorado Boulder.

Leave a Reply

Your email address will not be published. Required fields are marked *

Get the next one before it is old news

Independent analysis of cloud-native infrastructure, Kubernetes and data center economics. No vendor spin.