CoreWeave expands full-stack AI cloud push as inference demand grows
Source: SiliconANGLE (opens in a new tab) · Devony Hof
Intel Summary
SiliconANGLE reports that AI-native cloud provider CoreWeave Inc. is expanding its full-stack cloud offerings to support growing enterprise demand for AI inference workloads. The company recently completed validation and bring-up of Nvidia's Vera Rubin NVL72 architecture on its specialized platform.
Why It Matters
As enterprise priorities transition from model training to large-scale deployment, specialized cloud providers are racing to optimize hardware stacks for inference. Demonstrating early integration with next-generation accelerator architectures positions neoclouds to compete directly against incumbent hyperscalers for operational enterprise workloads.
Part of an ongoing development
SourceCoreWeave expands full-stack AI cloud offerings with Nvidia Vera Rubin NVL72 support
SiliconANGLE reports that AI-native cloud provider CoreWeave Inc. is expanding its full-stack cloud offerings to support growing enterprise demand for AI inference workloads. The company recently completed validation and bring-up of Nvidia's Vera Rubin NVL72 architecture on its specialized platform. Claims are as reported; this summary makes no determination about accuracy or significance.
Organizations & Entities
Topics
Related Intelligence
- ReportAlso involving Nvidia
OpenAI's first custom chip "Jalapeño" reportedly beats Nvidia's Blackwell and Rubin in inference benchmarks
At the Hot Chips conference, OpenAI unveiled 'Jalapeño,' its first custom in-house inference chip. According to benchmark assessments reported by research firm SemiAnalysis, the first-generation silicon reportedly outperforms Nvidia's Blackwell and Rubin architectures in both operational throughput and energy efficiency. While first-generation custom accelerators historically lag incumbent merchant silicon, initial disclosures indicate OpenAI's specialized design targets major bottlenecks in running production large language models at scale.
The Decoder - DevelopmentNewAlso involving Nvidia
Nvidia maps global power and data center capacity
The Information reports that Nvidia is actively mapping global power, land, and data center capacity to address infrastructure bottlenecks. CEO Jensen Huang stated at a Goldman Sachs conference that tracking power availability worldwide helps ensure facilities are prepared to bring AI server chips online immediately upon delivery, amid power constraints affecting developers for Google, Microsoft, and Oracle. Claims are as reported; this summary makes no determination about accuracy or significance.
- ReportAlso involving Nvidia
Nvidia and Cisco push the enterprise AI factory into the rack-scale era
Cisco Systems and Nvidia have expanded their collaborative Cisco Secure AI Factory architecture into rack-scale deployments. The updated infrastructure framework integrates full-stack enterprise artificial intelligence capabilities, incorporating rack-to-fabric liquid cooling designed to support high-density compute systems exceeding 200 kilowatts. The joint reference design targets enterprise data centers deploying generative and agentic AI workloads across their internal networks while standardizing compute, networking fabric, and operational management under unified security controls.
SiliconANGLE - ReportAlso involving Nvidia
Perplexity AI launches Portable Computer on-device AI agent
Perplexity AI has introduced Portable Computer, an artificial intelligence agent designed to run locally on desktop systems equipped with Nvidia hardware. The release transitions Perplexity's capabilities from pure cloud services to on-device agentic execution. The product launch arrives amid reports that Nvidia is contemplating a strategic investment in Perplexity AI that could value the startup at more than $30 billion, alongside potential technology licensing agreements.
SiliconANGLE