Perplexity, CoreWeave Deal Boosts Inferencing

Perplexity, CoreWeave Deal Boosts Inferencing

By AI Business
Publication Date: 2026-03-05 14:28:00

Neocloud provider CoreWeave and AI search vendor Perplexity have agreed to a multiyear deal to scale Perplexity’s AI search and inference capabilities. 

The agreement, financial terms of which were not disclosed, underscores the broad applicability of inferencing and the ongoing shift from AI training to AI inference.

The vendors revealed on March 4 that Perplexity will migrate its next AI inference workloads to CoreWeave Cloud. The partnership requires Nvidia’s GB200 NVL72 clusters to power Perplexity’s AI model, Sonar, and its Search API ecosystem. Perplexity will also use CKS (CoreWeave Kubernetes Services) and W&B (weights and balances) models for model management and deployment. CKS and W&B are core components of CoreWeave’s AI cloud platform, with CKS a managed service optimized for computationally intensive AI workloads, and W&B Models a specialized “system of record” for managing the lifecycle of machine learning models.

Related:Nvidia Takes on Telco Industry With Open Source Model

Perplexity and CoreWeave’s deal further shows the shift in the AI market toward inference, the process by which an AI uses the knowledge or data it acquired during training. Most recently, there have been deals in which vendors are partnering solely for inference. For instance, OpenAI recently committed to using 2 gigawatts of capacity on AWS’ Trainium3 and Trainium4 chips, following an expansion of its partnership with the cloud provider. Meta also plans to deploy millions of Nvidia…