CoreWeave AI云算力再升级!率先上线多机架英伟达Vera Rubin NVL72集群

智通财经
Yesterday

9月16日,CoreWeave(CRWV.US)宣布,已在CoreWeave Cloud上启用一套多机架的英伟达(NVDA.US)Vera Rubin NVL72集群,将数百颗英伟达Rubin图形处理器(GPU)互联成单一系统,用于智能体AI(agentic AI)工作负载。据CoreWeave称,这一多机架架构可让训练与推理工作负载跨数百颗Rubin GPU调度运行,为更大规模的模型、更高强度的推理与大规模强化学习提供算力支撑。

单个英伟达Vera Rubin NVL72机架整合了72颗Rubin GPU与36颗Vera CPU,并搭配英伟达的网络与数据处理技术。CoreWeave表示,多个机架通过英伟达Spectrum-X以太网互联,作为一套横向扩展集群协同运行。CoreWeave的基础设施可自动化完成机架部署、固件升级、验证以及供能与散热,并在机架投入生产前进行全机架级别的测试。

CoreWeave执行副总裁兼产品与工程负责人Chen Goldberg表示,据其所知,CoreWeave是首家完成Vera Rubin NVL72验证并将其投入运行的AI云服务提供商。他表示:“借助多机架Vera Rubin,我们正在将数百颗Rubin GPU连接成一个单一的横向扩展集群。”他补充称,该系统能够为开发智能体AI的客户提供更大的规模以及更快的迭代速度。

在扩大计算能力的同时,CoreWeave还为其AI对象存储平台推出了跨区域写入加速功能。该功能允许AI工作负载在运行所在的区域写入数据,而CoreWeave则在后台将数据复制到另一个区域。该公司表示,对于在多个区域运行训练工作负载的客户而言,这一功能可以减少暂停时间。

这家AI云计算公司还推出了Archive存储层,为希望长期保留数据的客户提供成本更低的存储选择。该存储层不收取数据取回费用、提前删除费用,也不收取从Archive存储层读取数据的费用。其本地对象传输加速器,即LOTA,还可以将数据缓存到更靠近AI工作负载的位置。CoreWeave表示,与从传统存储集群读取数据相比,该功能可将读取延迟降低8倍。

Disclaimer: Investing carries risk. This is not financial advice. The above content should not be regarded as an offer, recommendation, or solicitation on acquiring or disposing of any financial products, any associated discussions, comments, or posts by author or other users should not be considered as such either. It is solely for general information purpose only, which does not consider your own investment objectives, financial situations or needs. TTM assumes no responsibility or warranty for the accuracy and completeness of the information, investors should do their own research and may seek professional advice before investing.

Most Discussed

  1. 1
     
     
     
     
  2. 2
     
     
     
     
  3. 3
     
     
     
     
  4. 4
     
     
     
     
  5. 5
     
     
     
     
  6. 6
     
     
     
     
  7. 7
     
     
     
     
  8. 8
     
     
     
     
  9. 9
     
     
     
     
  10. 10