CoreWeave Sets New AI Benchmark with NVIDIA GB200 Superchips
CoreWeave has achieved a new record in AI inferencing benchmarks using NVIDIA GB200 Grace Blackwell Superchips, announced in a press release. The company reported delivering 800 tokens per second (TPS) on the Llama 3.1 405B model, one of the largest open-source models, using a CoreWeave instance equipped with two NVIDIA Grace CPUs and four NVIDIA Blackwell GPUs.
Additionally, CoreWeave submitted results for NVIDIA H200 GPU instances, achieving 33,000 TPS on the Llama 2 70B model, marking a 40% improvement over previous NVIDIA H100 instances. These achievements underscore CoreWeave's position as a leading provider of cloud infrastructure services optimized for AI applications.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like Silicon Brief or Daily AI Brief.
Also, consider following us on social media:
More from Data Centers
Sep 19 Virginia Proposes Data Center Rules and Creates AI Task Force Sep 19 House Passes Bill to Shield Ratepayers From Data Center Power Costs Sep 19 Laminar Joins L'Oreal Sustainability Accelerator Sep 19 Dnotitia Begins Testing VDPU ASIC Samples Sep 19 Nscale Files for US Initial Public OfferingSilicon Brief
Weekly coverage of AI hardware developments including chips, GPUs, cloud platforms, and data center technology.
Industry analysis
2025 Global Business Services Agenda: Gen AI Takes Center Stage
This industry analysis by The Hackett Group explores the transformative impact of generative artificial intelligence (Gen AI) on global business services (GBS) in 2025. The study highlights the shift from exploration to acceleration of Gen AI initiatives, with 89% of executives advancing these projects to improve customer satisfaction, innovate products, and reduce costs. The report also discusses the challenges and strategies for successful Gen AI adoption, emphasizing the need for a technology-enabled operating model and the importance of reskilling the workforce.
Read moreYou may also like
Aitech Introduces Rugged AI Boards for Defense Systems
Huawei Introduces OceanStor M900 Storage for AI Inference
d-Matrix Raises $275 Million for AI Inference Chips
Kasm and Intel Expand Private AI Workspaces for Xeon 6
China Merchants Bank Wins CNCF Contest for Kubernetes AI Platform
Daily AI Brief: the AI news that matters, in your inbox.